๐ In a Training Loop
Milton Montiel
miltmont
ยท
AI & ML interests
Reinforcement learning
Recent Activity
upvoted a paper 2 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses upvoted a paper 2 days ago
OmniEdu: Open Foundation Models for Learning and Teaching