Alexander Pondaven
I’m a PhD student at Oxford’s Torr Vision Group, working on controllable and interactive video generation — teaching world models to be steered, and even played.
Right now I’m a research intern at Odyssey, building multiplayer world models: generative environments that many players can act in at once. My recent paper ActionParty (ECCV 2026) is a step towards this — the first video world model that binds distinct actions to up to seven players in a scene.
- 🎮 Now — Research intern at Odyssey, building multiplayer world models
- 🎓 PhD (2023–) — Controllable video diffusion at Oxford AIMS CDT, funded by a Snap studentship
- 📄 ECCV 2026 — ActionParty: binding actions to multiple players in a world model
- 📄 CVPR 2025 — DiTFlow: motion transfer for diffusion transformers
- 🌍 NeurIPS 2022 — Satellite inpainting with neural processes for climate
Curious about anything below? The publications, projects and experience pages have the details — otherwise the best way to reach me is by email.
news
| Jun 1, 2026 | Started a research internship at Odyssey, building multiplayer world models 🎮 |
|---|---|
| Apr 3, 2026 | ActionParty accepted to ECCV 2026 🎉 — a multi-subject world model that controls up to seven players at once. [arXiv] |
| Feb 27, 2025 | DiTFlow — video motion transfer with diffusion transformers — accepted at CVPR 2025 🎉 |


