We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.RO

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Robotics

Title: Multi-Agent Variational Occlusion Inference Using People as Sensors

Abstract: Autonomous vehicles must reason about spatial occlusions in urban environments to ensure safety without being overly cautious. Prior work explored occlusion inference from observed social behaviors of road agents, hence treating people as sensors. Inferring occupancy from agent behaviors is an inherently multimodal problem; a driver may behave similarly for different occupancy patterns ahead of them (e.g., a driver may move at constant speed in traffic or on an open road). Past work, however, does not account for this multimodality, thus neglecting to model this source of aleatoric uncertainty in the relationship between driver behaviors and their environment. We propose an occlusion inference method that characterizes observed behaviors of human agents as sensor measurements, and fuses them with those from a standard sensor suite. To capture the aleatoric uncertainty, we train a conditional variational autoencoder with a discrete latent space to learn a multimodal mapping from observed driver trajectories to an occupancy grid representation of the view ahead of the driver. Our method handles multi-agent scenarios, combining measurements from multiple observed drivers using evidential theory to solve the sensor fusion problem. Our approach is validated on a cluttered, real-world intersection, outperforming baselines and demonstrating real-time capable performance. Our code is available at this https URL .
Comments: 12 pages, 9 figures, International Conference on Robotics and Automation (ICRA) 2022
Subjects: Robotics (cs.RO); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Multiagent Systems (cs.MA)
ACM classes: I.2.9; I.2.10
Cite as: arXiv:2109.02173 [cs.RO]
  (or arXiv:2109.02173v3 [cs.RO] for this version)

Submission history

From: Masha Itkina [view email]
[v1] Sun, 5 Sep 2021 21:56:54 GMT (15941kb,D)
[v2] Wed, 10 Nov 2021 18:31:52 GMT (16187kb,D)
[v3] Wed, 2 Mar 2022 20:37:01 GMT (16188kb,D)

Link back to: arXiv, form interface, contact.