GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows Paper • 2604.15715 • Published Apr 17 • 4
RTMO: Towards High-Performance One-Stage Real-Time Multi-Person Pose Estimation Paper • 2312.07526 • Published Apr 8, 2024
RMP-SAM: Towards Real-Time Multi-Purpose Segment Anything Paper • 2401.10228 • Published Jan 18, 2024
RouteMoA: Dynamic Routing without Pre-Inference Boosts Efficient Mixture-of-Agents Paper • 2601.18130 • Published Jan 26 • 2
DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning Paper • 2602.11089 • Published Feb 11 • 18
Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization Paper • 2603.28342 • Published Mar 30 • 24
TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration Paper • 2604.14116 • Published Apr 15 • 13
TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration Paper • 2604.14116 • Published Apr 15 • 13
The Curse and Blessing of Mean Bias in FP4-Quantized LLM Training Paper • 2603.10444 • Published Mar 11 • 13
End-to-End Video Character Replacement without Structural Guidance Paper • 2601.08587 • Published Jan 13 • 8
view post Post 363 QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Management See translation 👍 1 1 + Reply
BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping Paper • 2510.18927 • Published Oct 21, 2025 • 86
RoboChallenge: Large-scale Real-robot Evaluation of Embodied Policies Paper • 2510.17950 • Published Oct 20, 2025 • 11
Achieving Sample and Computational Efficient Reinforcement Learning by Action Space Reduction via Grouping Paper • 2306.12981 • Published Jun 22, 2023
Towards Language-Driven Video Inpainting via Multimodal Large Language Models Paper • 2401.10226 • Published Jan 18, 2024 • 2
OMG-Seg: Is One Model Good Enough For All Segmentation? Paper • 2401.10229 • Published Jan 18, 2024 • 1