Research Papers research paper arxiv computer-vision image-recognition

A.I.R.: Enabling Adaptive, Iterative, and Reasoning-based Frame Selection For Video Question Answering

arXivMarch 30, 202610 min read0 views

arXiv:2510.04428v3 Announce Type: replace Abstract: Effectively applying Vision-Language Models (VLMs) to Video Question Answering (VideoQA) hinges on selecting a concise yet comprehensive set of frames, as processing entire videos is computationally infeasible. However, current frame selection methods face a critical trade-off: approaches relying on lightweight similarity models, such as CLIP, often fail to capture the nuances of complex queries, resulting in inaccurate similarity scores that cannot reflect the authentic query-frame relevance, which further undermines frame selection. Meanwhi — Yuanhao Zou, Shengji Jin, Andong Deng, Youpeng Zhao, Jun Wang, Chen Chen

View PDF HTML (experimental)

Abstract:Effectively applying Vision-Language Models (VLMs) to Video Question Answering (VideoQA) hinges on selecting a concise yet comprehensive set of frames, as processing entire videos is computationally infeasible. However, current frame selection methods face a critical trade-off: approaches relying on lightweight similarity models, such as CLIP, often fail to capture the nuances of complex queries, resulting in inaccurate similarity scores that cannot reflect the authentic query-frame relevance, which further undermines frame selection. Meanwhile, methods that leverage a VLM for deeper analysis achieve higher accuracy but incur prohibitive computational costs. To address these limitations, we propose A.I.R., a training-free approach for Adaptive, Iterative, and Reasoning-based frame selection. We leverage a powerful VLM to perform deep, semantic analysis on complex queries, and this analysis is deployed within a cost-effective iterative loop that processes only a small batch of the most high-potential frames at a time. Extensive experiments on various VideoQA benchmarks demonstrate that our approach outperforms existing frame selection methods, significantly boosts the performance of the foundation VLM, and achieves substantial gains in computational efficiency over other VLM-based techniques.

Comments: ICLR 2026 Paper

Subjects:

Computer Vision and Pattern Recognition (cs.CV)

Cite as: arXiv:2510.04428 [cs.CV]

(or arXiv:2510.04428v3 [cs.CV] for this version)

https://doi.org/10.48550/arXiv.2510.04428

arXiv-issued DOI via DataCite

Submission history

From: Yuanhao Zou [view email] [v1] Mon, 6 Oct 2025 01:51:13 UTC (11,986 KB) [v2] Thu, 26 Feb 2026 01:08:09 UTC (11,987 KB) [v3] Fri, 27 Mar 2026 02:48:21 UTC (11,987 KB)

Original source

arXiv

https://arxiv.org/abs/2510.04428

Was this article helpful?

Ask AI about this article

Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

More about

researchpaperarxiv

ModelsLive

Caltech Researchers Claim Compression of High-Fidelity AI Models

Article URL: https://www.wsj.com/cio-journal/caltech-researchers-claim-radical-compression-of-high-fidelity-ai-models-e66f31c9 Comments URL: https://news.ycombinator.com/item?id=47593903 Points: 1 # Comments: 0

Hacker News AI Top

1m6 minutes ago

Research PapersLive

A Retrospective on the ICLR 2026 Review Process

The selection of papers for ICLR 2026 has fully concluded. We extend our congratulations to the authors whose work will appear at the conference. Creating ICLR’s technical program requires immense effort from the authors, reviewers, and area chairs, and we thank you for your contributions and service. For researchers whose work was rejected, we hope […]

blog.iclr.cc

1m25 minutes ago

Products

Bold bet on AI to keep UK at forefront of science and research breakthroughs from healthcare, to better public services - GOV.UK

<a href="https://news.google.com/rss/articles/CBMi6AFBVV95cUxPSU9QQ2Y0NVJHZDNPQ3htWE45R2tfODhYSzJfRm9aRjlzSmV5X1U5cVlKWFVqWmk0ZTZhV0x2VUNmZjg1Z05DVk41MW1hMzZJOE05WEFHNVFBZlg5ZjNHd21sUi1OVEp4SjlnQTV0UVFJLWJDVzdxdFRHelNsNC1yaWRyWGtPRXk1aGU4MHllQzRYNFdHVk1Yc3Fid09uV3VwWFN1Nkc0Yktnam04S0Y4cVJtMlJqY1hYczBpQnlCNUtEejBxaFBHUUN0cXJVcU53VjZoNm05QVlZd2dhek5STEhiQVgyT2t5?oc=5" target="_blank">Bold bet on AI to keep UK at forefront of science and research breakthroughs from healthcare, to better public services</a> GOV.UK

GNews AI UK

1mabout 1 month ago

Knowledge Map

TopicsEntitiesSource

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 89 connections

Scroll to zoom · drag to pan · click to open

Discussion

No comments yet — be the first to share your thoughts!

More in Research Papers

Research PapersLive

A Retrospective on the ICLR 2026 Review Process

blog.iclr.cc

1m25 minutes ago

Research Papers

Vector Researchers present papers at ACL 2024

Vector researchers will be well represented at the 62nd Annual Meeting of the Association for Computational Linguistics in Bangkok, Thailand this year. 14 papers co-authored by Vector-affiliated researchers are being […] The post Vector Researchers present papers at ACL 2024 appeared first on Vector Institute for Artificial Intelligence .

Vector Institute

1mover 1 year ago

Research Papers

Yann LeCun's Team's New Paper: AI Development Mimicking Human Intelligence Hits a Dead End - eu.36kr.com

<a href="https://news.google.com/rss/articles/CBMiU0FVX3lxTFBkbTRhNlhtRnY0cVBERld2OTdWNkRGMXBEaG9Vc21janRUcjJaUlJ4YzZRajVmMGQxNGJYTFB6M3lleUFNakUtWElHdGwzTXBQZjNZ?oc=5" target="_blank">Yann LeCun's Team's New Paper: AI Development Mimicking Human Intelligence Hits a Dead End</a> eu.36kr.com

GNews AI AGI

1m23 days ago

Research Papers

Plans must be made for the welfare of sentient AI, animal consciousness researchers argue - The Hill

<a href="https://news.google.com/rss/articles/CBMiiAFBVV95cUxNNzVaUTkzYkFUaVRsNGtnQVRXS2xsQVZfd1dFQ01RUlNZWUdDbjBNLUNycll2enl2NHp4Z0Ficm9HUnNWUnlvSGFrR3lDVUVxT1QyeE03QWhWcHFDTVJxV3VUQ0FKT3hiTkY3dWZha3JjcjRIM3l3WUtHZVlBUlhxdVBhLW1tdlJ40gGOAUFVX3lxTFBDQnllcVNNa1NRYVMyYlBtVXVxR0VPeHNjTjNMNWNTMFZXRjRkSU1OeXRFNmxvcENqbXkwSERoU1pGdXJYX2g5c214cFJFdEc1WUlkaEE5TlFDTTNoek5yR18tVi1vWUlGUnl4Tk13VWlFMDhzdUUyOUl3RmhNZ0FobTdiVG51N2h1SmJ5Y3c?oc=5" target="_blank">Plans must be made for the welfare of sentient AI, animal consciousness researchers argue</a> The Hill

GNews AI welfare

1mover 1 year ago