li sheng
bambisheng
AI & ML interests
None yet
Recent Activity
upvoted a paper 4 days ago
EasyPPO: Stabilizing the Critic Is Key upvoted a paper 10 days ago
Improving Test-Time Scaling with Adaptive Looped Transformers upvoted a paper 28 days ago
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks