Projects

Selected work in reinforcement learning and ML infrastructure — mostly tools built to make research reproducible.

2025— Offline RL benchmark suiteReproducible eval harness across 12 continuous-control tasks
2024 GridcraftMulti-agent gridworld environment for emergent-behavior experiments
2023 Policy distillation for edge inferenceCompressing large policies to run under 10ms on-device