๐ In a Training Loop
Milton Montiel
miltmont
ยท
AI & ML interests
Reinforcement learning
Recent Activity
upvoted a paper 6 days ago
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents upvoted a paper 12 days ago
The information geometry of large language models is shared, learned, and controllable upvoted a paper 13 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses