Research Papers research paper arxiv machine-learning deep-learning

Out-of-Sight Embodied Agents: Multimodal Tracking, Sensor Fusion, and Trajectory Forecasting

arXivMarch 30, 202610 min read0 views

arXiv:2509.15219v2 Announce Type: replace-cross Abstract: Trajectory prediction is a fundamental problem in computer vision, vision-language-action models, world models, and autonomous systems, with broad impact on autonomous driving, robotics, and surveillance. However, most existing methods assume complete and clean observations, and therefore do not adequately handle out-of-sight agents or noisy sensing signals caused by limited camera coverage, occlusions, and the absence of ground-truth denoised trajectories. These challenges raise safety concerns and reduce robustness in real-world deplo — Haichao Zhang, Yi Xu, Yun Fu

View PDF HTML (experimental)

Abstract:Trajectory prediction is a fundamental problem in computer vision, vision-language-action models, world models, and autonomous systems, with broad impact on autonomous driving, robotics, and surveillance. However, most existing methods assume complete and clean observations, and therefore do not adequately handle out-of-sight agents or noisy sensing signals caused by limited camera coverage, occlusions, and the absence of ground-truth denoised trajectories. These challenges raise safety concerns and reduce robustness in real-world deployment. In this extended study, we introduce major improvements to Out-of-Sight Trajectory (OST), a task for predicting noise-free visual trajectories of out-of-sight objects from noisy sensor observations. Building on our prior work, we expand Out-of-Sight Trajectory Prediction (OOSTraj) from pedestrians to both pedestrians and vehicles, increasing its relevance to autonomous driving, robotics, and surveillance. Our improved Vision-Positioning Denoising Module exploits camera calibration to establish vision-position correspondence, mitigating the lack of direct visual cues and enabling effective unsupervised denoising of noisy sensor signals. Extensive experiments on the Vi-Fi and JRDB datasets show that our method achieves state-of-the-art results for both trajectory denoising and trajectory prediction, with clear gains over prior baselines. We also compare with classical denoising methods, including Kalman filtering, and adapt recent trajectory prediction models to this setting, establishing a stronger benchmark. To the best of our knowledge, this is the first work to use vision-positioning projection to denoise noisy sensor trajectories of out-of-sight agents, opening new directions for future research.

Comments: Published in IEEE Transactions on Pattern Analysis and Machine Intelligence (Early Access), pp. 1-14, March 23, 2026

Subjects:

Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Multiagent Systems (cs.MA); Multimedia (cs.MM); Robotics (cs.RO)

MSC classes: 68T45, 68U10, 68T07, 68T40, 93C85, 93E11, 62M20, 62M10, 68U05, 94A12

ACM classes: F.2.2; I.2.9; I.2.10; I.4.1; I.4.8; I.4.9; I.5.4; I.3.7

Cite as: arXiv:2509.15219 [cs.CV]

(or arXiv:2509.15219v2 [cs.CV] for this version)

https://doi.org/10.48550/arXiv.2509.15219

arXiv-issued DOI via DataCite

Journal reference: IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026

DOI(s) linking to related resources

Submission history

From: Haichao Zhang [view email] [v1] Thu, 18 Sep 2025 17:59:16 UTC (3,235 KB) [v2] Thu, 26 Mar 2026 21:31:12 UTC (3,634 KB)

Original source

arXiv

https://arxiv.org/abs/2509.15219

Was this article helpful?

Ask AI about this article

Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

More about

researchpaperarxiv

AI ToolsLive

Which AI Tool Should You Use for What?

A practical guide for writers, researchers, and creators. Continue reading on Write A Catalyst »

Medium AI

1m44 minutes ago

Frontier ResearchLive

Can we ever trust AI to watch over itself?

Article URL: https://www.transformernews.ai/p/ai-alignment-researchers-want-to-superintelligence Comments URL: https://news.ycombinator.com/item?id=47655420 Points: 1 # Comments: 0

Hacker News AI Top

1mabout 1 hour ago

Research PapersLive

[R] ICML Anonymized git repos for rebuttal

A number of the papers I'm reviewing for have submitted additional figures and code through anonymized git repos (e.g. https://anonymous.4open.science/ ) to help supplement their rebuttal. Is this against any policy? I'm considering submitting additional graphs during the discussion phase for clarity, and would like to make sure that won't cause any issues submitted by /u/drahcirenoob [link] [comments]

Reddit r/MachineLearning

1mabout 1 hour ago

Knowledge Map

TopicsEntitiesSource

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 174 connections

Scroll to zoom · drag to pan · click to open

Discussion

No comments yet — be the first to share your thoughts!

More in Research Papers

Research PapersLive

[R] ICML Anonymized git repos for rebuttal

Reddit r/MachineLearning

1mabout 1 hour ago

Research Papers

Tech Moves: Microsoft execs depart; TerraClear, UserTesting, EchoMark and Read AI add leaders - GeekWire

Tech Moves: Microsoft execs depart; TerraClear, UserTesting, EchoMark and Read AI add leaders GeekWire

GNews AI Microsoft

1m2 days ago

Research PapersFresh

[D] Is research in semantic segmentation saturated?

Nowadays I dont see a lot of papers addressing 2D semantic segmentation problem statements be it supervised, semi-supervised, domain adaptation. Is the problem statement saturated? Are there any promising research directions in segmentation except open-set segmentation? submitted by /u/Hot_Version_6403 [link] [comments]

Reddit r/MachineLearning

1mabout 9 hours ago

Research Papers

AI, quantum computing, fusion energy remain Energy’s top research priorities - Nextgov/FCW

AI, quantum computing, fusion energy remain Energy’s top research priorities Nextgov/FCW

GNews AI quantum

1m4 months ago