๐ In a Training Loop
Milton Montiel
miltmont
ยท
AI & ML interests
Reinforcement learning
Recent Activity
upvoted a paper 1 day ago
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents upvoted a paper 8 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses