Karthik Iyer
kiyer97
AI & ML interests
Reinforcement learning, reward modeling, RLHF, policy optimization, offline RL
Recent Activity
upvoted a paper about 16 hours ago
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself upvoted a paper about 16 hours ago
Grounded Skill Synthesis from Code at Scale for Agentic Intelligence upvoted a paper about 16 hours ago
StudentSim: Training LLM-based Student SimulatorsOrganizations
None yet