EngiWorld: What Can Frontier Agents Deliver in Professional Engineering Environments? Paper • 2609.37686 • Published 8 days ago • 70
Self-Evolving Coding Agents: From Digital Programs to Physical-World Intelligence Paper • 2609.35432 • Published 9 days ago • 107
RayOrch: Programming and Executing Lineage-Controlled Multi-Grain Dataflows for Foundation-Model Data Preparation Paper • 2609.18703 • Published 21 days ago • 56
OmniEdu: Open Foundation Models for Learning and Teaching Paper • 2609.23088 • Published 18 days ago • 238
DataFlex-RL: An Evaluation Platform for RLVR Data Policies Paper • 2609.06107 • Published Sep 5 • 166
DataPrep-Bench: Benchmarking LLMs as Training Data Preparators Paper • 2607.20465 • Published May 19 • 56
DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines Paper • 2607.16617 • Published Jul 18 • 99
SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks Paper • 2606.09669 • Published Jun 8 • 50
OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Paper • 2604.04707 • Published Apr 6 • 199
DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models Paper • 2603.26164 • Published Mar 27 • 280
One-Eval: An Agentic System for Automated and Traceable LLM Evaluation Paper • 2603.09821 • Published Mar 10 • 11
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization Paper • 2601.05242 • Published Jan 8 • 235
DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI Paper • 2512.16676 • Published Dec 18, 2025 • 225
view changelog Hugging Face Changelog AI-generated Abstract summaries on Hugging Face Papers May 22, 2025 • 78