Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
36.9
TFLOPS
ManniX
PRO
ManniX-ITA
32
6
28
Follow
Fernanduli's profile picture
Omegakenshin's profile picture
Lamsam26's profile picture
83 followers
·
20 following
https://github.com/mann1x
mann1x
AI & ML interests
None yet
Recent Activity
commented
on
a paper
about 1 hour ago
Group Entropy-Controlled Policy Optimization
replied
to
their
post
1 day ago
# opencoti-llamafile 0.10.3-c5 — settled admission for multi-agent serving New cut of the opencoti single-file inference engine (llamafile 0.10.3 / llama.cpp + 87 additive patches). Zero-dependency APE: one executable for Linux, Windows, macOS & BSD. What's new vs c4: **PolyKV fan-out — pool from a live session.** `POST /polykv/pools` gains `from_session`/`from_slot`: the shared prefix is snapshotted server-side from the session's cached KV — no tokens resent, token-exact. Ephemeral pools auto-release when orchestrators die mid-round. **PolyKV P7 — settled admission.** Spawning agents faster than the tps signal settles was oversubscribing pools. Now: a per-pool settle window paces admits just enough for a reliable reading; warming sessions no longer bias the mean; the post-admit forecast uses the measured per-admit drop; idle gaps (agents mid-tool-call) no longer read as free capacity; `guarantee_min_sessions` means a new/nested pool always gets its first agent — capacity checks can never deadlock an orchestrator; the enforced gate applies to new sessions only, with per-request `overcommit`. Benchmark (multi-agent courier, floor 15 tok/s): time-under-floor −52%, deep sub-floor −83%, delivery p50 −43%, 100% task score. **Zero-conf GPU sharing.** Instances on one GPU discover each other over shared memory — no ports, no config — and split compute by `--gpu-share-weight`. Measured (3090): weights 2:1 → 71.7/36.2 tok/s; holds at `--parallel 4` and under MTP. Idle peers cost nothing (solo = full speed), crashes age out in 3 s; `GET /gpu/peers` shows live shares + busy %. From c5 every release ships per-platform side-load DSOs: `dso/<ver>/` with Linux x86_64 + sbsa `.so` and a Windows `.dll`. https://huggingface.co/ManniX-ITA/opencoti-llamafile
replied
to
their
post
1 day ago
# opencoti-llamafile 0.10.3-c5 — settled admission for multi-agent serving New cut of the opencoti single-file inference engine (llamafile 0.10.3 / llama.cpp + 87 additive patches). Zero-dependency APE: one executable for Linux, Windows, macOS & BSD. What's new vs c4: **PolyKV fan-out — pool from a live session.** `POST /polykv/pools` gains `from_session`/`from_slot`: the shared prefix is snapshotted server-side from the session's cached KV — no tokens resent, token-exact. Ephemeral pools auto-release when orchestrators die mid-round. **PolyKV P7 — settled admission.** Spawning agents faster than the tps signal settles was oversubscribing pools. Now: a per-pool settle window paces admits just enough for a reliable reading; warming sessions no longer bias the mean; the post-admit forecast uses the measured per-admit drop; idle gaps (agents mid-tool-call) no longer read as free capacity; `guarantee_min_sessions` means a new/nested pool always gets its first agent — capacity checks can never deadlock an orchestrator; the enforced gate applies to new sessions only, with per-request `overcommit`. Benchmark (multi-agent courier, floor 15 tok/s): time-under-floor −52%, deep sub-floor −83%, delivery p50 −43%, 100% task score. **Zero-conf GPU sharing.** Instances on one GPU discover each other over shared memory — no ports, no config — and split compute by `--gpu-share-weight`. Measured (3090): weights 2:1 → 71.7/36.2 tok/s; holds at `--parallel 4` and under MTP. Idle peers cost nothing (solo = full speed), crashes age out in 3 s; `GET /gpu/peers` shows live shares + busy %. From c5 every release ships per-platform side-load DSOs: `dso/<ver>/` with Linux x86_64 + sbsa `.so` and a Windows `.dll`. https://huggingface.co/ManniX-ITA/opencoti-llamafile
View all activity
Organizations
None yet
ManniX-ITA
's models
41
Sort: Recently updated
ManniX-ITA/opencoti-llamafile
Updated
1 day ago
•
134
•
1
ManniX-ITA/Qwen3.6-27B-A3B-Coder-MTP-GGUF
Text Generation
•
26B
•
Updated
12 days ago
•
38k
•
16
ManniX-ITA/Qwen3.6-27B-A3B-Coder
Text Generation
•
27B
•
Updated
13 days ago
•
463
•
4
ManniX-ITA/gemma-4-A4B-98e-v7-coderx-it-GGUF
20B
•
Updated
13 days ago
•
14.5k
•
1
ManniX-ITA/gemma-4-A4B-98e-v7-coder-it-GGUF
20B
•
Updated
13 days ago
•
16.4k
•
5
ManniX-ITA/Qwen3.5-4B-MicroCoder-GGUF
4B
•
Updated
15 days ago
•
134
•
1
ManniX-ITA/Qwen3.5-4B-M8-GGUF
4B
•
Updated
15 days ago
•
126
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4-MTP-GGUF
27B
•
Updated
15 days ago
•
4.4k
•
10
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4-GGUF
Image-Text-to-Text
•
27B
•
Updated
15 days ago
•
4.06k
•
35
ManniX-ITA/gemma-4-A4B-98e-v6-coder-it-GGUF
20B
•
Updated
15 days ago
•
8.06k
•
4
ManniX-ITA/gemma-4-31B-it-assistant-GGUF
0.5B
•
Updated
28 days ago
•
103
ManniX-ITA/gemma-4-26B-A4B-it-assistant-GGUF
0.4B
•
Updated
28 days ago
•
88
ManniX-ITA/gemma-4-12B-it-assistant-GGUF
0.4B
•
Updated
28 days ago
•
120
ManniX-ITA/gemma-4-E4B-it-assistant-GGUF
78.8M
•
Updated
28 days ago
•
82
ManniX-ITA/gemma-4-E2B-it-assistant-GGUF
78M
•
Updated
28 days ago
•
97
ManniX-ITA/gemma-4-A4B-98e-v7-coder-NVFP4A16
11B
•
Updated
Jun 24
•
29
ManniX-ITA/gemma-4-A4B-98e-v7-coder-it
20B
•
Updated
Jun 24
•
45
ManniX-ITA/gemma-4-A4B-98e-v7-coderx-NVFP4A16
11B
•
Updated
Jun 24
•
23
ManniX-ITA/gemma-4-A4B-98e-v7-coderx-it
20B
•
Updated
Jun 24
•
17
•
2
ManniX-ITA/gemma-4-A4B-98e-v6-coder-it
20B
•
Updated
Jun 10
•
7
•
1
ManniX-ITA/Qwen3.5-4B-MicroCoder
Image-Text-to-Text
•
5B
•
Updated
May 25
•
6
ManniX-ITA/gemma-4-A4B-98e-v5-coder-it
20B
•
Updated
May 24
•
12
•
3
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4
Image-Text-to-Text
•
28B
•
Updated
May 22
•
23
•
14
ManniX-ITA/gemma-4-31b-he1-it
Text Generation
•
31B
•
Updated
May 21
•
5
•
1
ManniX-ITA/gemma-4-A4B-98e-v5-it
20B
•
Updated
May 20
•
10
ManniX-ITA/gemma-4-31b-he1-it-NVFP4A16
Text Generation
•
17B
•
Updated
May 19
•
6
ManniX-ITA/Gemma-4-31B-it-NVFP4A16
Text Generation
•
17B
•
Updated
May 18
•
7
ManniX-ITA/gemma-4-A4B-98e-v5-coder-NVFP4A16
Text Generation
•
11B
•
Updated
May 18
•
8
ManniX-ITA/gemma-4-A4B-98e-v4-it
20B
•
Updated
May 17
•
3
ManniX-ITA/Gemma-4-26B-A4B-it-NVFP4A16
14B
•
Updated
May 14
•
9
Previous
1
2
Next