๐ In a Training Loop
Milton Montiel
miltmont
ยท
AI & ML interests
Reinforcement learning
Recent Activity
upvoted a paper about 2 hours ago
The information geometry of large language models is shared, learned, and controllable upvoted a paper about 10 hours ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses upvoted a paper about 10 hours ago
OmniEdu: Open Foundation Models for Learning and Teaching