Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
24.3
TFLOPS
Flavio Catalani
fakezeta
19
7
42
Follow
muriloscezar's profile picture
21world's profile picture
emmacall's profile picture
19 followers
·
36 following
fakezeta
flaviocatalani
AI & ML interests
None yet
Recent Activity
reacted
to
pollix
's
post
with 🚀
2 days ago
stuntd 0.1.2 is out 🎉 stuntd sits in front of your LLM, learns its typed decisions and answers the confident ones locally with a small head on the Laya encoder by @convaiinnovations. About 20ms on GPU and 60ms on CPU, and anything it isn't sure about still goes to the big model. New in 0.1.2: - decisions with several fields, like category + urgency + needs_human in one call, answered locally only when every field is sure - the Anthropic Messages API learns too, not only OpenAI - auto_retrain: the daemon retrains a site in the background once enough new traffic comes in, so collect, train, shadow and live run on their own - serve --lazy loads the checkpoint on the first request Try it in the browser: https://huggingface.co/spaces/pollix/stuntd Code: https://github.com/bladedevoff/stuntd ``` pip install -U stuntd ```
liked
a model
10 days ago
agentionai/Qwen3.8-27B-AP-GGUF
liked
a model
15 days ago
Abiray/MiniMax-H3-Pruned-GGUF
View all activity
Organizations
fakezeta
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF
22 days ago
Task effectiveness maintained – Knowledge loss expected
👍
1
1
#29 opened 22 days ago by
fakezeta
New activity in
bartowski/Qwen_Qwen3.6-35B-A3B-GGUF
5 months ago
MTP GGUF?
4
#4 opened 5 months ago by
fakezeta
New activity in
bartowski/google_gemma-4-26B-A4B-it-GGUF
6 months ago
<unused49> infinite generation after llama.cpp release b8699
5
#2 opened 6 months ago by
fakezeta
New activity in
unsloth/Qwen3-30B-A3B-GGUF
over 1 year ago
Please quantize a Q5_K_S version
👍
1
3
#9 opened over 1 year ago by
fakezeta
New activity in
fakezeta/amoral-Qwen3-4B
over 1 year ago
Adding `safetensors` variant of this model
#1 opened over 1 year ago by
SFconvertbot
New activity in
unsloth/Qwen3-14B-GGUF
over 1 year ago
FIXED: Failed to parse Jinja template
😔
1
8
#2 opened over 1 year ago by
wapxmas
New activity in
mistralai/Mistral-Small-Instruct-2409
about 2 years ago
Please make it CLEAR, this is NOT an OPEN SOURCE MODEL license
4
#15 opened about 2 years ago by
KingBadger
New activity in
NousResearch/Hermes-2-Pro-Llama-3-8B
over 2 years ago
OpenVINO IR model with int8 quantization
👍
1
5
#2 opened over 2 years ago by
fakezeta
New activity in
fakezeta/Yi-1.5-6B-Chat-ov-int8
over 2 years ago
Hi. Thank you very much for this nice model. Does exist any native Windows GUI app that can run it?
1
#1 opened over 2 years ago by
NikolayKozloff
New activity in
HPAI-BSC/Llama3-Aloe-8B-Alpha
over 2 years ago
OpenVINO IR model
2
#1 opened over 2 years ago by
fakezeta
New activity in
fakezeta/Llama3-Aloe-8B-Alpha-ov-int8
over 2 years ago
Add LocalAI configuration
1
#1 opened over 2 years ago by
mudler
New activity in
Nexusflow/Starling-LM-7B-beta
over 2 years ago
Adding proposed model avatar
❤️
1
3
#3 opened over 2 years ago by
Suparious
New activity in
Intel/neural-chat-7b-v3-1
almost 3 years ago
Prompt Template?
👍
4
13
#1 opened almost 3 years ago by
fakezeta
New activity in
fakezeta/neural-chat-7b-v3-1-GGUF
almost 3 years ago
rename the model to neural-chat-7b-v3-1 -GGUF
1
#1 opened almost 3 years ago by
eramax
New activity in
TheBloke/openchat_3.5-GGUF
almost 3 years ago
Prints only reply but when asked something extra it hangs
👍
1
14
#1 opened almost 3 years ago by
Pumba2
New activity in
fakezeta/pdfchat
over 3 years ago
Fork not possible - no licence
1
#1 opened over 3 years ago by
michaelfeil