MUVO: A Multimodal Generative World Model for Autonomous Driving with Geometric Representations

November 20, 2023 ยท Declared Dead ยท ๐Ÿ› 2025 IEEE Intelligent Vehicles Symposium (IV)

๐Ÿ‘ป CAUSE OF DEATH: Ghosted
No code link whatsoever

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Daniel Bogdoll, Yitian Yang, Tim Joseph, Melih Yazgan, J. Marius Zรถllner arXiv ID 2311.11762 Category cs.LG: Machine Learning Cross-listed cs.RO Citations 34 Venue 2025 IEEE Intelligent Vehicles Symposium (IV) Last Checked 6 months ago
Abstract
World models for autonomous driving have the potential to dramatically improve the reasoning capabilities of today's systems. However, most works focus on camera data, with only a few that leverage lidar data or combine both to better represent autonomous vehicle sensor setups. In addition, raw sensor predictions are less actionable than 3D occupancy predictions, but there are no works examining the effects of combining both multimodal sensor data and 3D occupancy prediction. In this work, we perform a set of experiments with a MUltimodal World Model with Geometric VOxel representations (MUVO) to evaluate different sensor fusion strategies to better understand the effects on sensor data prediction. We also analyze potential weaknesses of current sensor fusion approaches and examine the benefits of additionally predicting 3D occupancy.
Community shame:
Not yet rated
Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

๐Ÿ“œ Similar Papers

In the same crypt โ€” Machine Learning

Died the same way โ€” ๐Ÿ‘ป Ghosted