Live
Black Hat USAAI BusinessBlack Hat AsiaAI BusinessFrom False Positives to Real Risk: AI‑Driven Compliance in Modern UC - UC TodayGoogle News: Generative AIClaude Code Leak: What Went Wrong at Anthropic? - AI MagazineGoogle News: ClaudeU.S. Reportedly Seeking Access To Three Additional Bases In Greenland, The First Expansion In DecadesInternational Business TimesAnthropic's Claude Code source code got accidentally leaked - qz.comGoogle News: ClaudeAI’s Biggest Opportunity Lies in the 92% of Work It Hasn’t Touched - PYMNTS.comGoogle News: AIWhy is gaming becoming so expensive? The answer is found in AI - The GuardianGoogle News: AIChoosing the Right Model is Hard. Maintaining Accuracy is Harder.AI YouTube Channel 24A YouTuber channeled his distaste for the PS5’s design into slick console coversThe Verge AILess than a month: StrictlyVC San Francisco brings leaders from TDK Ventures, Replit, and more togetherTechCrunch AIThe Strange, Shaky Alliance Taking on Trump and His Big Tech Friends - PoliticoGoogle News: AI SafetyI Asked ChatGPT If It Was A Psychopath—Here’s What It Said - ForbesGoogle News: ChatGPTGoogle’s TurboQuant Marks A Turning Point In AI’s Evolution - ForbesGoogle News: LLMBlack Hat USAAI BusinessBlack Hat AsiaAI BusinessFrom False Positives to Real Risk: AI‑Driven Compliance in Modern UC - UC TodayGoogle News: Generative AIClaude Code Leak: What Went Wrong at Anthropic? - AI MagazineGoogle News: ClaudeU.S. Reportedly Seeking Access To Three Additional Bases In Greenland, The First Expansion In DecadesInternational Business TimesAnthropic's Claude Code source code got accidentally leaked - qz.comGoogle News: ClaudeAI’s Biggest Opportunity Lies in the 92% of Work It Hasn’t Touched - PYMNTS.comGoogle News: AIWhy is gaming becoming so expensive? The answer is found in AI - The GuardianGoogle News: AIChoosing the Right Model is Hard. Maintaining Accuracy is Harder.AI YouTube Channel 24A YouTuber channeled his distaste for the PS5’s design into slick console coversThe Verge AILess than a month: StrictlyVC San Francisco brings leaders from TDK Ventures, Replit, and more togetherTechCrunch AIThe Strange, Shaky Alliance Taking on Trump and His Big Tech Friends - PoliticoGoogle News: AI SafetyI Asked ChatGPT If It Was A Psychopath—Here’s What It Said - ForbesGoogle News: ChatGPTGoogle’s TurboQuant Marks A Turning Point In AI’s Evolution - ForbesGoogle News: LLM

TIR-Agent: Training an Explorative and Efficient Agent for Image Restoration

arXivMarch 31, 20262 min read0 views
Source Quiz

arXiv:2603.27742v1 Announce Type: new Abstract: Vision-language agents that orchestrate specialized tools for image restoration (IR) have emerged as a promising method, yet most existing frameworks operate in a training-free manner. They rely on heuristic task scheduling and exhaustive tool traversal, resulting in sub-optimal restoration paths and prohibitive computational cost. We argue that the core bottleneck lies in the absence of a learned policy to make decision, as a vision-language model cannot efficiently handle degradation-aware task ordering and tool composition. To this end, we pro — Yisheng Zhang, Guoli Jia, Haote Hu, Shanxu Zhao, Kaikai Zhao, Long Sun, Xinwei Long, Kai Tian, Che Jiang, Zhaoxiang Liu, Kai Wang, Shiguo Lian, Kaiyan Zhang, Bowen Zhou

Authors:Yisheng Zhang, Guoli Jia, Haote Hu, Shanxu Zhao, Kaikai Zhao, Long Sun, Xinwei Long, Kai Tian, Che Jiang, Zhaoxiang Liu, Kai Wang, Shiguo Lian, Kaiyan Zhang, Bowen Zhou

View PDF HTML (experimental)

Abstract:Vision-language agents that orchestrate specialized tools for image restoration (IR) have emerged as a promising method, yet most existing frameworks operate in a training-free manner. They rely on heuristic task scheduling and exhaustive tool traversal, resulting in sub-optimal restoration paths and prohibitive computational cost. We argue that the core bottleneck lies in the absence of a learned policy to make decision, as a vision-language model cannot efficiently handle degradation-aware task ordering and tool composition. To this end, we propose TIR-Agent, a trainable image restoration agent that performs a direct tool-calling policy through a two-stage training pipeline of supervised fine-tuning (SFT) followed by reinforcement learning (RL). Two key designs underpin effective RL training: (i) a random perturbation strategy applied to the SFT data, which broadens the policy's exploration over task schedules and tool compositions, and (ii) a multi-dimensional adaptive reward mechanism that dynamically re-weights heterogeneous image quality metrics to mitigate reward hacking. To support high-throughput, asynchronous GPU-based tool invocation during training, we further develop a globally shared model-call pool. Experiments on both in-domain and out-of-domain degradations show that TIR-Agent outperforms 12 baselines, including 6 all-in-one models, 3 training-free agents, and 3 proprietary models, and achieves over 2.5$\times$ inference speedup by eliminating redundant tool executions.

Subjects:

Computer Vision and Pattern Recognition (cs.CV)

Cite as: arXiv:2603.27742 [cs.CV]

(or arXiv:2603.27742v1 [cs.CV] for this version)

https://doi.org/10.48550/arXiv.2603.27742

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Guoli Jia [view email] [v1] Sun, 29 Mar 2026 15:50:36 UTC (8,265 KB)

Was this article helpful?

Sign in to highlight and annotate this article

AI
Ask AI about this article
Powered by AI News Hub · full article context loaded
Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

Knowledge Map

Knowledge Map
TopicsEntitiesSource
TIR-Agent: …researchpaperarxivcomputer-vi…image-recog…arXiv

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Building knowledge graph…

Discussion

Sign in to join the discussion

No comments yet — be the first to share your thoughts!