๐ In a Training Loop
Milton Montiel
miltmont
ยท
AI & ML interests
Reinforcement learning
Recent Activity
upvoted a paper about 23 hours ago
The information geometry of large language models is shared, learned, and controllable upvoted a paper 1 day ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses upvoted a paper 1 day ago
OmniEdu: Open Foundation Models for Learning and Teaching