Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
59.7
TFLOPS
s
Tom-Neverwinter
44
1
9
Follow
21world's profile picture
beem9988's profile picture
rajafa's profile picture
6 followers
·
22 following
Tom-Neverwinter
AI & ML interests
Making improvements to help the world.
Recent Activity
liked
a model
1 day ago
moonshotai/Kimi-K3
reacted
to
badaoui
's
post
with 🚀
3 days ago
432 GB of ultra-fast HBM4 and up to 23.3 TB/s of memory bandwidth on a single GPU 🤯. Two weeks ago, we got early access to AMD's new Instinct MI455X, and our first goal was simple: make sure 🤗 Transformers works on day one. Over the past few weeks, we worked closely with the AMD team to validate the platform, enable Flash Attention, add torchcodec support for multimodal models, and resolve issues uncovered during testing. The result: ✅ 99.5% success rate across our 24 core Transformers model architectures - already on par with our daily CI on previous AMD and NVIDIA platforms. The hardware is just as exciting. With 432 GB of HBM per GPU, our early capacity experiments showed more than 3× the concurrent long-context requests compared to MI300, thanks to the much larger KV cache capacity. A huge thanks to the AMD team for the early access and the great collaboration! Read the full blog 👇 https://huggingface.co/blog/badaoui/transformers-on-amd-mi455
reacted
to
badaoui
's
post
with 👍
3 days ago
432 GB of ultra-fast HBM4 and up to 23.3 TB/s of memory bandwidth on a single GPU 🤯. Two weeks ago, we got early access to AMD's new Instinct MI455X, and our first goal was simple: make sure 🤗 Transformers works on day one. Over the past few weeks, we worked closely with the AMD team to validate the platform, enable Flash Attention, add torchcodec support for multimodal models, and resolve issues uncovered during testing. The result: ✅ 99.5% success rate across our 24 core Transformers model architectures - already on par with our daily CI on previous AMD and NVIDIA platforms. The hardware is just as exciting. With 432 GB of HBM per GPU, our early capacity experiments showed more than 3× the concurrent long-context requests compared to MI300, thanks to the much larger KV cache capacity. A huge thanks to the AMD team for the early access and the great collaboration! Read the full blog 👇 https://huggingface.co/blog/badaoui/transformers-on-amd-mi455
View all activity
Organizations
None yet
Tom-Neverwinter
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
prism-ml/Ternary-Bonsai-27B-gguf
7 days ago
MOE or changing to A3B
2
#43 opened 7 days ago by
Tom-Neverwinter
New activity in
BeaverAI/Artemis-31B-v1m-GGUF
14 days ago
any smaller models coming any time soon?
3
#1 opened 14 days ago by
Tom-Neverwinter
New activity in
InternScience/Agents-A1
24 days ago
benchmaxxed
6
#11 opened 24 days ago by
Tom-Neverwinter
New activity in
lunahr/Marlin-2B-ungated
2 months ago
ungated?
3
#1 opened 2 months ago by
Tom-Neverwinter
New activity in
IQuestLab/IQuest-Coder-V1-40B-Loop-Instruct
7 months ago
Benchmaxxed
🧠
👍
10
2
#12 opened 7 months ago by
Tom-Neverwinter
New activity in
microsoft/VibeVoice-Realtime-0.5B
7 months ago
Safety or the joke there in
🔥
1
4
#7 opened 8 months ago by
Tom-Neverwinter
New activity in
AaryanK/IQuest-Coder-V1-40B-Instruct-GGUF
7 months ago
Benchmaxxed to the max
1
#1 opened 7 months ago by
Tom-Neverwinter
New activity in
cturan/IQuest-Coder-V1-40B-Instruct-GGUF
7 months ago
Benchmaxxed
#2 opened 7 months ago by
Tom-Neverwinter
New activity in
BeaverAI/Anubis-Mini-8B-v1a-GGUF
10 months ago
recommended prompt
#1 opened 10 months ago by
Tom-Neverwinter
New activity in
Qwen/Qwen3-Omni-30B-A3B-Instruct
10 months ago
Example of how to set up Qwen3-omni for audio-input audio-output
👍
🔥
10
2
#2 opened 10 months ago by
abidlabs
New activity in
nvidia/parakeet-tdt-0.6b-v3
11 months ago
Japanese support plan?
👍
5
6
#5 opened 11 months ago by
sttt
New activity in
ReadyArt/Safeword-Casual-V1-12B-GGUF
11 months ago
Anything special?
😎
1
2
#1 opened 11 months ago by
Tom-Neverwinter
New activity in
nvidia/parakeet-rnnt-1.1b
about 1 year ago
Support translation
2
#3 opened over 2 years ago by
SyedAbdul
New activity in
BeaverAI/Test-12B-v1b-GGUF
about 1 year ago
another new model :)
#1 opened about 1 year ago by
Tom-Neverwinter
New activity in
Qwen/Qwen2.5-Omni-7B
about 1 year ago
GGUF model
3
#51 opened about 1 year ago by
Tom-Neverwinter
New activity in
moonshotai/MoonViT-SO-400M
over 1 year ago
GGUF format
#2 opened over 1 year ago by
Tom-Neverwinter
New activity in
TheDrummer/Gemmasutra-Small-4B-v1-GGUF
over 1 year ago
Review so far
2
#1 opened over 1 year ago by
IggyLux
New activity in
multimodalart/flux-lora-the-explorer
almost 2 years ago
how to make a lora
3
#2 opened almost 2 years ago by
guardiancc
New activity in
meta-llama/Llama-3.1-8B-Instruct
about 2 years ago
Issues loading model with ooabooga textgenwebui
👍
3
5
#20 opened about 2 years ago by
Kenji776
New activity in
lmstudio-community/DeepSeek-Coder-V2-Lite-Instruct-GGUF
about 2 years ago
GGUF for the 236B model
3
#4 opened about 2 years ago by
amarmir
Load more