KIMTAEIL
7stars-commander
AI & ML interests
None yet
Recent Activity
upvoted a paper about 2 months ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement upvoted a paper about 2 months ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement LearningOrganizations
None yet