-
R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning
Paper • 2508.21113 • Published • 111 -
Breaking the Exploration Bottleneck: Rubric-Scaffolded Reinforcement Learning for General LLM Reasoning
Paper • 2508.16949 • Published • 25 -
EmbodiedOneVision: Interleaved Vision-Text-Action Pretraining for General Robot Control
Paper • 2508.21112 • Published • 77 -
UItron: Foundational GUI Agent with Advanced Perception and Planning
Paper • 2508.21767 • Published • 12
Jeff Nyzio
TheOneTrueNiz
AI & ML interests
None yet
Recent Activity
liked a model 12 days ago
Agnes-AI/Agnes-3.0-Flash liked a model about 1 month ago
Qwen/Qwen3.8-27B new activity about 1 month ago
Lightricks/LTX-2.3-22b-IC-LoRA-Ingredients:Multiple character references all become invalid.Organizations
None yet