Collections
Discover the best community collections!
Collections including paper arxiv:2604.06392
-
Qualixar OS: A Universal Operating System for AI Agent Orchestration
Paper • 2604.06392 • Published • 20 -
Personalizing Text-to-Image Generation to Individual Taste
Paper • 2604.07427 • Published • 8 -
Lance: Unified Multimodal Modeling by Multi-Task Synergy
Paper • 2605.18678 • Published • 79 -
Qwen-Image-VAE-2.0 Technical Report
Paper • 2605.13565 • Published • 62
-
Low-probability Tokens Sustain Exploration in Reinforcement Learning with Verifiable Reward
Paper • 2510.03222 • Published • 76 -
In-the-Flow Agentic System Optimization for Effective Planning and Tool Use
Paper • 2510.05592 • Published • 112 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 518 -
Multi-Agent Tool-Integrated Policy Optimization
Paper • 2510.04678 • Published • 31
-
Qualixar OS: A Universal Operating System for AI Agent Orchestration
Paper • 2604.06392 • Published • 20 -
AVO: Agentic Variation Operators for Autonomous Evolutionary Search
Paper • 2603.24517 • Published • 11 -
Seeing the Needle in the Haystack: Towards Weakly-Supervised Log Instance Anomaly Localization via Counterfactual Perturbation
Paper • 2605.10988 • Published • 4 -
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
Paper • 2605.25624 • Published • 35
-
ExGRPO: Learning to Reason from Experience
Paper • 2510.02245 • Published • 83 -
A Practitioner's Guide to Multi-turn Agentic Reinforcement Learning
Paper • 2510.01132 • Published • 6 -
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Paper • 2510.04618 • Published • 134 -
MixReasoning: Switching Modes to Think
Paper • 2510.06052 • Published • 23
-
Qualixar OS: A Universal Operating System for AI Agent Orchestration
Paper • 2604.06392 • Published • 20 -
AVO: Agentic Variation Operators for Autonomous Evolutionary Search
Paper • 2603.24517 • Published • 11 -
Seeing the Needle in the Haystack: Towards Weakly-Supervised Log Instance Anomaly Localization via Counterfactual Perturbation
Paper • 2605.10988 • Published • 4 -
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
Paper • 2605.25624 • Published • 35
-
Qualixar OS: A Universal Operating System for AI Agent Orchestration
Paper • 2604.06392 • Published • 20 -
Personalizing Text-to-Image Generation to Individual Taste
Paper • 2604.07427 • Published • 8 -
Lance: Unified Multimodal Modeling by Multi-Task Synergy
Paper • 2605.18678 • Published • 79 -
Qwen-Image-VAE-2.0 Technical Report
Paper • 2605.13565 • Published • 62
-
Low-probability Tokens Sustain Exploration in Reinforcement Learning with Verifiable Reward
Paper • 2510.03222 • Published • 76 -
In-the-Flow Agentic System Optimization for Effective Planning and Tool Use
Paper • 2510.05592 • Published • 112 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 518 -
Multi-Agent Tool-Integrated Policy Optimization
Paper • 2510.04678 • Published • 31
-
ExGRPO: Learning to Reason from Experience
Paper • 2510.02245 • Published • 83 -
A Practitioner's Guide to Multi-turn Agentic Reinforcement Learning
Paper • 2510.01132 • Published • 6 -
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
Paper • 2510.04618 • Published • 134 -
MixReasoning: Switching Modes to Think
Paper • 2510.06052 • Published • 23