REVIEW 2 cited by
RMP-YOLO: A Robust Motion Predictor for Partially Observable Scenarios even if You Only Look Once
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We introduce RMP-YOLO, a unified framework designed to provide robust motion predictions even with incomplete input data. Our key insight stems from the observation that complete and reliable historical trajectory data plays a pivotal role in ensuring accurate motion prediction. Therefore, we propose a new paradigm that prioritizes the reconstruction of intact historical trajectories before feeding them into the prediction modules. Our approach introduces a novel scene tokenization module to enhance the extraction and fusion of spatial and temporal features. Following this, our proposed recovery module reconstructs agents' incomplete historical trajectories by leveraging local map topology and interactions with nearby agents. The reconstructed, clean historical data is then integrated into the downstream prediction modules. Our framework is able to effectively handle missing data of varying lengths and remains robust against observation noise, while maintaining high prediction accuracy. Furthermore, our recovery module is compatible with existing prediction models, ensuring seamless integration. Extensive experiments validate the effectiveness of our approach, and deployment in real-world autonomous vehicles confirms its practical utility. In the 2024 Waymo Motion Prediction Competition, our method, RMP-YOLO, achieves state-of-the-art performance, securing third place.
Forward citations
Cited by 2 Pith papers
-
ModeSeq: Taming Sparse Multimodal Motion Prediction with Sequential Mode Modeling
By decoding trajectory modes sequentially with a new Early-Match-Take-All training loss, ModeSeq improves mode diversity and confidence calibration in sparse multimodal motion prediction.
-
AGI-Elo: How Far Are We From Mastering A Task?
AGI-Elo applies Elo/Glicko-style ratings to model-versus-test-case matches, producing joint difficulty and competency scores and competency-gap estimates across six AI benchmarks.
Discussion (0). Continue with ORCID to comment.