PIP-Net: Pedestrian Intention Prediction in the Wild

Mohsen Azarmi; Mahdi Rezaei; He Wang

arXiv:2402.12810·cs.CV·July 8, 2025·1 cites

PIP-Net: Pedestrian Intention Prediction in the Wild

Mohsen Azarmi, Mahdi Rezaei, He Wang

PDF

Open Access

TL;DR

PIP-Net is a new deep learning framework that predicts pedestrian crossing intentions in real-world urban environments, utilizing multi-camera data and novel features to improve accuracy and forecast up to 4 seconds ahead.

Contribution

The paper introduces PIP-Net, a novel recurrent and attention-based model for pedestrian intention prediction, along with the Urban-PIP dataset for real-world multi-camera scenarios.

Findings

01

Outperforms state-of-the-art models in pedestrian intention prediction

02

Predicts crossing intentions up to 4 seconds in advance

03

Enhances scene understanding with multi-camera and depth features

Abstract

Accurate pedestrian intention prediction (PIP) by Autonomous Vehicles (AVs) is one of the current research challenges in this field. In this article, we introduce PIP-Net, a novel framework designed to predict pedestrian crossing intentions by AVs in real-world urban scenarios. We offer two variants of PIP-Net designed for different camera mounts and setups. Leveraging both kinematic data and spatial features from the driving scene, the proposed model employs a recurrent and temporal attention-based solution, outperforming state-of-the-art performance. To enhance the visual representation of road users and their proximity to the ego vehicle, we introduce a categorical depth feature map, combined with a local motion flow feature, providing rich insights into the scene dynamics. Additionally, we explore the impact of expanding the camera's field of view, from one to three cameras…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsVideo Surveillance and Tracking Methods · Autonomous Vehicle Technology and Safety