Multi-Agent Variational Occlusion Inference Using People as Sensors

Itkina, Masha; Mun, Ye-Ji; Driggs-Campbell, Katherine; Kochenderfer, Mykel J.

Computer Science > Robotics

arXiv:2109.02173 (cs)

[Submitted on 5 Sep 2021 (v1), last revised 2 Mar 2022 (this version, v3)]

Title:Multi-Agent Variational Occlusion Inference Using People as Sensors

Authors:Masha Itkina, Ye-Ji Mun, Katherine Driggs-Campbell, Mykel J. Kochenderfer

View PDF

Abstract:Autonomous vehicles must reason about spatial occlusions in urban environments to ensure safety without being overly cautious. Prior work explored occlusion inference from observed social behaviors of road agents, hence treating people as sensors. Inferring occupancy from agent behaviors is an inherently multimodal problem; a driver may behave similarly for different occupancy patterns ahead of them (e.g., a driver may move at constant speed in traffic or on an open road). Past work, however, does not account for this multimodality, thus neglecting to model this source of aleatoric uncertainty in the relationship between driver behaviors and their environment. We propose an occlusion inference method that characterizes observed behaviors of human agents as sensor measurements, and fuses them with those from a standard sensor suite. To capture the aleatoric uncertainty, we train a conditional variational autoencoder with a discrete latent space to learn a multimodal mapping from observed driver trajectories to an occupancy grid representation of the view ahead of the driver. Our method handles multi-agent scenarios, combining measurements from multiple observed drivers using evidential theory to solve the sensor fusion problem. Our approach is validated on a cluttered, real-world intersection, outperforming baselines and demonstrating real-time capable performance. Our code is available at this https URL .

Comments:	12 pages, 9 figures, International Conference on Robotics and Automation (ICRA) 2022
Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Multiagent Systems (cs.MA)
ACM classes:	I.2.9; I.2.10
Cite as:	arXiv:2109.02173 [cs.RO]
	(or arXiv:2109.02173v3 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2109.02173

Submission history

From: Masha Itkina [view email]
[v1] Sun, 5 Sep 2021 21:56:54 UTC (15,941 KB)
[v2] Wed, 10 Nov 2021 18:31:52 UTC (16,187 KB)
[v3] Wed, 2 Mar 2022 20:37:01 UTC (16,188 KB)

Computer Science > Robotics

Title:Multi-Agent Variational Occlusion Inference Using People as Sensors

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Multi-Agent Variational Occlusion Inference Using People as Sensors

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators