h Ethan Ma — Writing

Writing

Notes on reinforcement learning, mostly the unglamorous parts.

2026.03 Why offline RL evals lie to youOn the gap between benchmark reward and real performance
2025.09 Notes on reproducible experiment trackingA checklist I wish I'd had two years ago