Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 6 days ago • 24
Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering Paper • 2607.21848 • Published 5 days ago • 5
DataPrep-Bench: Benchmarking LLMs as Training Data Preparators Paper • 2607.20465 • Published May 19 • 40
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-metrics-bucket 613 kB
Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models Paper • 2607.19604 • Published 7 days ago • 16
SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD Paper • 2607.20145 • Published 6 days ago • 62
Environment-free Synthetic Data Generation for API-Calling Agents Paper • 2607.16900 • Published 10 days ago • 20
SciForma: Structure-Faithful Generation of Scientific Diagrams Paper • 2607.18091 • Published 8 days ago • 22
Running Repro Self Distillation Enables Continual Learning Metrics 🎯 View and manage your tracking data in an interactive dashboard
Running Repro Self Distillation Enables Continual Learning Metrics 🎯 View and manage your tracking data in an interactive dashboard
Running Repro Adversarial Dual On Policy Distillation From Expressive Flow Based Teacher Metrics 🎯 View and manage your tracking data with an interactive dashboard
Running Repro Adversarial Dual On Policy Distillation From Expressive Flow Based Teacher Metrics 🎯 View and manage your tracking data with an interactive dashboard
codemaivanngu/repro-adversarial-dual-on-policy-distillation-from-expressive-flow-based-teacher-metrics-bucket 613 kB