Steven Roche
MIT–WHOI Joint Program | Reinforcement Learning · Underwater Robotics · Robotic Perception
rochesh [at] mit [dot] edu
I am a PhD student in the MIT–WHOI Joint Program, based in MIT AeroAstro's Aerospace Controls Laboratory under Prof. Jonathan How and working closely with WARPLab at Woods Hole under Dr. Yogesh Girdhar. I am supported by a DoD SMART Scholarship.
My research aims to make autonomous underwater vehicles capable, robust field instruments for ocean science, with a focus on coral reefs. I work on reinforcement learning for vehicle control, closing the sim-to-real gap with high-fidelity physics such as computational fluid dynamics, and on robotic perception for autonomous sampling, including environmental DNA collection on reefs.
Before MIT, I was a Gabelli Presidential Scholar at Boston College, where I earned two degrees, a B.S. in Computer Science and a B.S. in Mathematics. There I worked on graph algorithms (min-max correlation clustering, my Scholar of the College thesis) and, as a NOAA Hollings Scholar, on machine learning for GPS-denied magnetic navigation. I have also contributed to robot learning for quadruped locomotion and to whole-brain voltage imaging in larval zebrafish.
news
| Sep 2026 | Preprint of Feedback-Modulated Harmonic Policies for Quadruped Locomotion posted to arXiv. Paper |
|---|---|
| Aug 2026 | Voltage imaging of neurons distributed across entire brains of larval zebrafish published in Nature Methods. Paper |
| Jul 2026 | Preprint of CORL-AUV: CFD Oriented Reinforcement Learning for Autonomous Underwater Vehicles posted to arXiv. Paper / Project |
| 2026 | Received a WHOI Ocean Ventures Fund award for Autonomous eDNA Sampling for Coral Reefs. |
| Feb 2026 | New blog post: how to find and fix a leaky end cap on an underwater robot. |
| 2025 | Started graduate school in the MIT–WHOI Joint Program (MIT AeroAstro) as a DoD SMART Scholar. |
selected publications
-
Feedback-Modulated Harmonic Policies for Quadruped LocomotionarXiv preprint arXiv:2609.17946, 2026
-
Voltage imaging of neurons distributed across entire brains of larval zebrafishNature Methods, 2026
-
CORL-AUV: CFD Oriented Reinforcement Learning for Autonomous Underwater VehiclesarXiv preprint arXiv:2607.09557, 2026
-
AI-Driven Magnetic Navigation: Predicting Position from AnomaliesAGU Annual Meeting (AGU24), 2024
outside the lab
When I'm not in the lab, I run trail ultramarathons, including races put on by the Trail Animals Running Club (TARC):
- TARC Fall Classic, 50 miles (September 2026): 15th of 63 men in 10:58:51. Results
- TARC Spring Classic, 50K (April 2025): 16th of 47 men (4th in men 20–29) in 6:31:13. Results