view article Article tokenizers v1: encode, decode and scaling, measured +2 ArthurZ, sbrandeis, mcpotato, lysandre • 10 days ago • 80
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 28 days ago • 141
view article Article Training a coding model to paint watercolours with TRL and OpenEnv sergiopaniego • 28 days ago • 76
view article Article How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code nielsr • Aug 21 • 32
UniSpace: Unified Visual Representation and Scalable Multimodal Modeling Paper • 2608.08676 • Published Aug 9 • 15
RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO Paper • 2605.15190 • Published May 14 • 15
view article Article State of Open Models: Summer 2026 Observations +1 AdinaY, multimodalart, irenesolaiman • Aug 14 • 217
SpotSound: Enhancing Large Audio-Language Models with Fine-Grained Temporal Grounding Paper • 2604.13023 • Published Apr 14 • 2
LTX-2.5 Collection LTX-2.5 base models, quantized models and accompanying LoRAs and IC-LoRAs • 5 items • Updated 21 days ago • 65
Reference-Driven Multi-Speaker Audio Scene Generation from In-the-Wild Priors Paper • 2606.19325 • Published Jun 17 • 2
Scaling Properties of Text Conditioning in Visual Generation Paper • 2607.29679 • Published Jul 31 • 42
view article Article IDEOGRAM-4 for inpainting with Modular Diffusers and Differential Diffusion OzzyGT • Aug 5 • 8
QueenVIS: Rethinking Image-Only Training for Video Instance Segmentation via Query Enrichment Paper • 2607.24598 • Published Jul 27 • 9
StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation Paper • 2607.26754 • Published Jul 29 • 19
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • Jul 27 • 507