GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay Paper • 2609.25001 • Published 4 days ago • 129
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 7 days ago • 78
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation Paper • 2609.20511 • Published 8 days ago • 108
Heoni/llama-3-KoEn-8b_sft_ep4_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 12 • 1
Heoni/llama-3-KoEn-8b_sft_ep5_merged_red_teaming_20240621_new_template Text Generation • Updated Jun 21, 2024 • 16 • 1
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 10 days ago • 44
Heoni/llama-3-KoEn-8b_sft_ep3_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 18 • 1
HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness Paper • 2609.15195 • Published 11 days ago • 22
Heoni/llama-3-KoEn-8b_sft_ep5_merged_red_teaming_20240623_final_data Text Generation • Updated Jun 23, 2024 • 16 • 1
Another Blueprint In The Wall: How to Ask Frontier AI Like a Kid? Paper • 2609.14803 • Published 12 days ago • 12
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 11 days ago • 214
Grouped Value Attention: Efficient KV Caching via On-Demand Key Reconstruction Paper • 2609.13285 • Published 17 days ago • 81
andersonbcdefg/red_teaming_reward_modeling_pairwise_no_as_an_ai Viewer • Updated Jun 1, 2023 • 35.3k • 135 • 6