UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement Paper • 2609.38721 • Published 5 days ago • 279
view article Article NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction nvidia • 6 days ago • 75
Advancing Open and Reproducible Relational Learning: RelArena-α, TabPFN-Rel and RPI Paper • 2608.16319 • Published Aug 17 • 31
To Think or Not to Think: The Hidden Cost of Meta-Training with Excessive CoT Examples Paper • 2512.05318 • Published Dec 4, 2025 • 3
view article Article 🐯 Liger GRPO meets TRL +4 shisahni, kashif, smohammadi, ShirinYamani, m0m0chen, liberty4321 • May 25, 2025 • 54