Fundamentals of Arthroscopic Surgery Training and beyond: a reinforcement learning exploration and benchmark
*International Journal of Computer Assisted Radiology and Surgery* · Ivan Ovinnikov, Ami Beuret, Flavia Cavaliere, Joachim M Buhmann
Applied Scientist · RL, Simulation & Physical AI
ANYbotics · Zürich, Switzerland
I connect research in inverse reinforcement learning, imitation learning, and robust objectives with deployed robot locomotion and scalable simulation workflows.
I completed my PhD at ETH Zürich on reinforcement learning from demonstrations in digital-twin simulations and now build RL locomotion systems for industrial quadruped robots at ANYbotics.
Selected work
Deployed locomotion, simulation infrastructure, robust reward learning, and distributional objectives.
Internal deployment work on RL locomotion controllers, sim-to-real training workflows, and safety-critical evaluation for industrial quadrupeds.
ArXiv preprint on reward learning from diverse demonstrations that remain stable under environment shifts.
Peer-reviewed work on reinforcement-learning benchmarks and assistance policies in surgical digital-twin environments.
Turning optimal-transport distances between expert and policy behavior into practical reinforcement-learning rewards.
ArXiv preprint on generative modeling with Wasserstein autoencoders in hyperbolic latent spaces.
Selected publications
Work spanning reinforcement learning, imitation learning, causal invariance, and optimal transport.
*International Journal of Computer Assisted Radiology and Surgery* · Ivan Ovinnikov, Ami Beuret, Flavia Cavaliere, Joachim M Buhmann
TMLR, in review · Ivan Ovinnikov, Alexander Terenin, Joachim M. Buhmann
Preprint / manuscript · Ivan Ovinnikov, Eugene Bykovets, Joachim M. Buhmann
Preprint / manuscript · Ivan Ovinnikov, Joachim M. Buhmann
Preprint / manuscript · Ivan Ovinnikov
Writing
A short technical note on connecting reward learning, simulation-based training, and safety-critical evaluation for physical AI systems.
Successfully passed my PhD examination
Our paper on surgical assistance agents in simulation was published in the