Computer Science > Computer Vision and Pattern Recognition

arXiv:2407.21126 (cs)

[Submitted on 30 Jul 2024]

Title:Self-supervised Multi-future Occupancy Forecasting for Autonomous Driving

Authors:Bernard Lange, Masha Itkina, Jiachen Li, Mykel J. Kochenderfer

View PDF

Abstract:Environment prediction frameworks are critical for the safe navigation of autonomous vehicles (AVs) in dynamic settings. LiDAR-generated occupancy grid maps (L-OGMs) offer a robust bird's-eye view for the scene representation, enabling self-supervised joint scene predictions while exhibiting resilience to partial observability and perception detection failures. Prior approaches have focused on deterministic L-OGM prediction architectures within the grid cell space. While these methods have seen some success, they frequently produce unrealistic predictions and fail to capture the stochastic nature of the environment. Additionally, they do not effectively integrate additional sensor modalities present in AVs. Our proposed framework performs stochastic L-OGM prediction in the latent space of a generative architecture and allows for conditioning on RGB cameras, maps, and planned trajectories. We decode predictions using either a single-step decoder, which provides high-quality predictions in real-time, or a diffusion-based batch decoder, which can further refine the decoded frames to address temporal consistency issues and reduce compression losses. Our experiments on the nuScenes and Waymo Open datasets show that all variants of our approach qualitatively and quantitatively outperform prior approaches.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
Cite as:	arXiv:2407.21126 [cs.CV]
	(or arXiv:2407.21126v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2407.21126

Submission history

From: Bernard Lange [view email]
[v1] Tue, 30 Jul 2024 18:37:59 UTC (24,089 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Self-supervised Multi-future Occupancy Forecasting for Autonomous Driving

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Self-supervised Multi-future Occupancy Forecasting for Autonomous Driving

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators