view article Article Async GRPO with LoRA across HF Jobs: a bucket, a proxy, and no NCCL +2 aminediroHF, qgallouedec, kashif, sergiopaniego • 15 days ago • 52
Hyper3-CLIP: Hierarchy-Conditioned Hyperbolic Vision-Language Training Paper • 2608.29313 • Published 27 days ago • 3
EvoOntology: A Self-Evolving Ontology Layer for Data Agents Paper • 2609.15779 • Published 11 days ago • 150
Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable Rewards Paper • 2609.03181 • Published 23 days ago • 11
The Router Within: Eliciting Native Skill Routing from a Frozen LLM Paper • 2609.15982 • Published 11 days ago • 8
view article Article One sandbox per rollout, or how labs run RL for agents in 2026 sergiopaniego • 13 days ago • 10
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published 15 days ago • 46
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models Paper • 2609.08418 • Published 17 days ago • 135
Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training Paper • 2609.15051 • Published 11 days ago • 14
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published 21 days ago • 116
NeoMME: A Single-Tower Multimodal-Native Multilingual Foundation Encoder for Efficient Fine-Tuning and Inference Paper • 2609.01657 • Published 25 days ago • 34
NeoMME Collection Meet NeoMME: a family of 260M and 800M Multimodal-Native Multilingual Encoders • 12 items • Updated 21 days ago • 33
view article Article NeoMME: an efficient Multimodal-native and Multilingual Encoder Hcompany • 21 days ago • 110