Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Troy Baker
jtroybaker
8
104
Follow
antgas's profile picture
21world's profile picture
mayafree's profile picture
3 followers
·
20 following
jtroybaker
jtroybaker
AI & ML interests
Predictive Maintenance, Reinforcement Learning, Natural Language Processing
Recent Activity
liked
a model
8 days ago
XiaomiMiMo/MiMo-V2.6-Flash-RL
liked
a model
9 days ago
convaiinnovations/laya
reacted
to
tomaarsen
's
post
with 🔥
about 1 month ago
🚨 I've just published Sentence Transformers v6.0, introducing MultiVectorEncoder: ColBERT-style late interaction models are now a fourth model type, for training, inference, and interpretation, alongside the dense, sparse, and reranker models! Details: Where a regular embedding model compresses a whole text into one vector, a multi-vector model keeps one vector per token and scores query against document with the MaxSim operator. That preserves token-level matching information that a single vector has to average away. It is also the state of the art for visual document retrieval, where a text query is matched against page images directly, charts and tables included, with no OCR step in between. Any PyLate, Stanford ColBERT, or ColPali checkpoint loads straight into the same familiar API: model.encode_query(), model.encode_document(), and model.similarity() just work, whether the documents are texts or page images. Does it help? LightOn trained LateOn (multi-vector) and DenseOn (dense) on the same data with the same 149M ModernBERT backbone, and the multi-vector model wins on 9 of the 13 NanoBEIR datasets: 0.6868 vs 0.6764 mean NDCG@10. The price is a bigger index, and the new HierarchicalTokenPooling module halves it at roughly no retrieval cost. Antoine Chaffin, Raphaël Sourty, and I wrote a blog post walking through multi-vector models in practice: loading the various checkpoint formats, encoding and scoring, plugging them into a search stack, running them on page images, and keeping the index affordable. Check it out if you want to get started, or just point your Agent to the URL: https://huggingface.co/blog/multi-vector-encoder pip install sentence-transformers==6.0.0 Release notes: https://github.com/huggingface/sentence-transformers/releases/tag/v6.0.0
View all activity
Organizations
jtroybaker
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
8 days ago
XiaomiMiMo/MiMo-V2.6-Flash-RL
Text Generation
•
311B
•
Updated
8 days ago
•
39.6k
•
519
liked
a model
9 days ago
convaiinnovations/laya
Text Classification
•
0.4B
•
Updated
6 days ago
•
4.51k
liked
a model
about 1 month ago
lintw/HealthGPT-Pro-27B
Image-Text-to-Image
•
3.05M
•
Updated
Jul 31
•
297
•
5
liked
5 models
about 2 months ago
MiniMaxAI/MiniMax-Music3
Text-to-Audio
•
2B
•
Updated
Aug 14
•
9.1k
•
1.42k
inclusionAI/Ling-3.0-flash
Text Generation
•
127B
•
Updated
6 days ago
•
16.6k
•
•
417
AtomicChat/Ling-3.0-flash-GGUF
Text Generation
•
124B
•
Updated
Aug 9
•
289k
•
65
MiniMaxAI/MiniMax-H3
Image-Text-to-Video
•
33B
•
Updated
Aug 13
•
3.63M
•
•
5.77k
poolside/Laguna-S-2.1
Text Generation
•
118B
•
Updated
Aug 19
•
107k
•
•
1.03k
liked
2 models
2 months ago
BlueBackup/Qwen3.6-35B-A3B-uncensored-heretic-IQ2_M
Text Generation
•
35B
•
Updated
Jul 25
•
839
•
9
unsloth/Laguna-S-2.1-GGUF
Text Generation
•
118B
•
Updated
Jul 27
•
16.1k
•
310
liked
3 models
3 months ago
vmlinux/Qwen3.5-122B-A10B-ROCmFP4-iMatrix-GGUF
Text Generation
•
122B
•
Updated
Jul 26
•
104
•
6
thinkingmachines/Inkling
Image-Text-to-Text
•
952B
•
Updated
Jul 23
•
630k
•
•
1.79k
nvidia/Nemotron-Labs-Audex-30B-A3B
Text Generation
•
Updated
Aug 13
•
418
•
177
liked
6 models
5 months ago
unsloth/Mistral-Medium-3.5-128B-GGUF
125B
•
Updated
May 2
•
13.1k
•
87
BidirLM/BidirLM-Omni-2.5B-Embedding
Sentence Similarity
•
2B
•
Updated
May 12
•
1.9k
•
50
FINAL-Bench/Darwin-36B-Opus
Text Generation
•
35B
•
Updated
Jul 25
•
413
•
97
unsloth/Qwen3.6-27B-GGUF
Image-Text-to-Text
•
27B
•
Updated
Apr 22
•
881k
•
958
nvidia/Lyra-2.0
Image-to-3D
•
Updated
Jul 15
•
500
•
348
moonshotai/Kimi-K2.6
Image-Text-to-Text
•
1T
•
Updated
May 19
•
461k
•
•
1.61k
liked
a model
6 months ago
Qwen/Qwen3.6-35B-A3B
Image-Text-to-Text
•
36B
•
Updated
Apr 24
•
3.23M
•
•
2.89k
Load more