Ben Calvert
bencalvert04
AI & ML interests
Building evals for agents
Recent Activity
upvoted a paper about 6 hours ago
Rethinking Training-Inference Mismatch in LLM Reinforcement Learning: Where It Arises and How to Correct It upvoted a paper 11 days ago
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation new activity 3 months ago
harborframework/parity-experiments:Add parity experiments for programbenchOrganizations
None yet