Paint-Anything: Unified Any-Color Control for Image Generation and Editing Paper • 2609.20816 • Published 7 days ago • 53
andersonbcdefg/sharegpt_reward_modeling_pairwise_no_as_an_ai Viewer • Updated Jun 6, 2023 • 11.8k • 102 • 3
Region-Level Policy Optimization for Fine-grained MLLM Perception Paper • 2609.19745 • Published 7 days ago • 41
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 7 days ago • 175
crumb/bloom-560m-RLHF-SD2-prompter-aesthetic Text Generation • 0.6B • Updated Mar 19, 2023 • 247 • 24
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 9 days ago • 44
sabaridsnfuji/repro-off-policy-learning-in-large-action-spaces-optimization-matters-more-than-estimation Viewer • Updated Jul 25 • 1 • 24 • 1
Difficulty-Adaptive Tree-Structured Policy Optimization for Expanding Reasoning Coverage in RLVR Paper • 2609.08650 • Published 16 days ago • 11
AgenticGen: Reward-Guided Agentic Video Generation for Advertising Paper • 2609.09187 • Published 24 days ago • 14