Only 15 observed demonstrations and 40 minutes of training from scratch
If you find this work useful, please cite:
@article{han2026mpail2,
title = {Online World Modeling Enables Real-World Inverse Reinforcement Learning from Observation},
author = {Han, Tyler and Nemekhbold, Bat and Shen, Siyang and Baijal, Rohan and
Ebock, Richard and Ravichandiran, Harine and Jung, Sanghun and
Huang, Kevin and Boots, Byron},
year = {2026},
}