view article Article **Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization** nvidia • 3 days ago • 52
Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable Rewards Paper • 2609.03181 • Published 24 days ago • 11
view article Article Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers tomaarsen • Aug 26 • 144
view article Article Lattice: an 8 MB static retriever that embeds Wikipedia in 7 minutes erikkaum • Aug 7 • 19
ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog Paper • 2607.04438 • Published Jul 5 • 61
F2LLM-v2: Inclusive, Performant, and Efficient Embeddings for a Multilingual World Paper • 2603.19223 • Published Mar 19 • 39
view article Article After the party comes the free lunch: regularizing ColBERT models to enhance pooling capabilities and reduce index footprint lightonai • Jul 6 • 17