DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning Paper • 2501.12948 • Published Jan 22, 2025 • 458
view article Article Mixture of Experts (MoEs) in Transformers +5 ariG23498, pcuenq, merve, IlyasMoutawwakil, ArthurZ, sergiopaniego, Molbap • Feb 26 • 173
view article Article Native-speed vLLM transformers modeling backend hmellor, lysandre • 20 days ago • 60
UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks Paper • 2607.08768 • Published 19 days ago • 34
PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space Paper • 2607.05373 • Published 22 days ago • 66
Image Classification Models Collection LiteRT image-classification models from litert-community. • 89 items • Updated 2 days ago • 7
Google Tensor Collection LiteRT Models that can run on Google Tensor • 5 items • Updated 2 days ago • 9
CohereLabs/cohere-transcribe-arabic-07-2026 Automatic Speech Recognition • 2B • Updated 15 days ago • 39.4k • 137
ABC-Bench Collection Evaluating Agentic Backend Coding Capabilities in Real-World Development Scenarios • 4 items • Updated 17 days ago • 5
MOSS Transcribe Collection A unified multimodal large language model for end-to-end speaker-attributed, time-stamped transcription. • 4 items • Updated 17 days ago • 13
MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization Paper • 2601.01554 • Published Jan 4 • 65
view article Article Hugging Face and Cerebras bring Gemma 4 to real-time voice AI +2 A-Mahla, andito, lvwerra, vyassaurabh • 27 days ago • 87