Research Papers research paper arxiv machine-learning deep-learning

SkillRouter: Skill Routing for LLM Agents at Scale

arXivby [Submitted on 23 Mar 2026 (v1), last revised 31 Mar 2026 (this version, v3)]March 31, 20262 min read1 views

arXiv:2603.22455v2 Announce Type: replace Abstract: Reusable skills let LLM agents package task-specific procedures, tool affordances, and execution guidance into modular building blocks. As skill ecosystems grow to tens of thousands of entries, exposing every skill at inference time becomes infeasible. This creates a skill-routing problem: given a user task, the system must identify relevant skills before downstream planning or execution. Existing agent stacks often rely on progressive disclosure, exposing only skill names and descriptions while hiding the full implementation body. We examine — YanZhao Zheng, ZhenTao Zhang, Chao Ma, YuanQiang Yu, JiHuai Zhu, Yong Wu, Tianze Xu, Baohua Dong, Hangcheng Zhu, Ruohui Huang, Gang Yu

Authors:YanZhao Zheng, ZhenTao Zhang, Chao Ma, YuanQiang Yu, JiHuai Zhu, Yong Wu, Tianze Xu, Baohua Dong, Hangcheng Zhu, Ruohui Huang, Gang Yu

View PDF HTML (experimental)

Abstract:Reusable skills let LLM agents package task-specific procedures, tool affordances, and execution guidance into modular building blocks. As skill ecosystems grow to tens of thousands of entries, exposing every skill at inference time becomes infeasible. This creates a skill-routing problem: given a user task, the system must identify relevant skills before downstream planning or execution. Existing agent stacks often rely on progressive disclosure, exposing only skill names and descriptions while hiding the full implementation body. We examine this design choice on a SkillsBench-derived benchmark with approximately 80K candidate skills, targeting the practically important setting of large skill registries with heavy overlap. Across representative sparse, dense, and reranking baselines on this setting, hiding the skill body causes a 31--44 percentage point drop in routing accuracy, showing that full skill text is a critical routing signal in this setting rather than a minor metadata refinement. Motivated by this finding, we present SkillRouter, a compact 1.2B full-text retrieve-and-rerank pipeline. SkillRouter achieves 74.0% Hit@1 on our benchmark -- the strongest average top-1 routing performance among the baselines we evaluate -- while using 13$\times$ fewer parameters and running 5.8$\times$ faster than the strongest base pipeline. The ranking gains further generalize to a supplementary benchmark independently constructed from three skill sources. In a complementary end-to-end study across four coding agents, routing gains transfer to improved task success, with larger gains for more capable agents.

Subjects:

Machine Learning (cs.LG)

Cite as: arXiv:2603.22455 [cs.LG]

(or arXiv:2603.22455v3 [cs.LG] for this version)

https://doi.org/10.48550/arXiv.2603.22455

arXiv-issued DOI via DataCite

Submission history

From: Yanzhao Zheng [view email] [v1] Mon, 23 Mar 2026 18:23:59 UTC (545 KB) [v2] Mon, 30 Mar 2026 09:19:32 UTC (438 KB) [v3] Tue, 31 Mar 2026 16:28:22 UTC (439 KB)

Original source

arXiv

https://arxiv.org/abs/2603.22455

Was this article helpful?

Ask AI about this article

Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

More about

researchpaperarxiv

Market NewsFresh

[D] The memory chip market lost tens of billions over a paper this community would have understood in 10 minutes

TurboQuant was teased recently and tens of billions gone from memory chip market in 48 hours but anyone in this community who read the paper would have seen the problem with the panic immediately. TurboQuant compresses the KV cache down to 3 bits per value from the standard 16 using polar coordinate quantization. But the KV cache is inference memory. Training memory, activations, gradients, optimizer states, is a completely different thing and completely untouched. And majority of HBM demand comes from training. An inference compression paper doesn't move that number. And the commercial inference baseline already runs at 4 to 8 bit precision. The 6x headline is benchmarked against 16 bit full precision. The real marginal gain over what's actually deployed is considerably smaller than that

Reddit r/MachineLearning

2mabout 4 hours ago

Research PapersFresh

[D] Is research in semantic segmentation saturated?

Nowadays I dont see a lot of papers addressing 2D semantic segmentation problem statements be it supervised, semi-supervised, domain adaptation. Is the problem statement saturated? Are there any promising research directions in segmentation except open-set segmentation? submitted by /u/Hot_Version_6403 [link] [comments]

Reddit r/MachineLearning

1mabout 5 hours ago

Research Papers

AI, quantum computing, fusion energy remain Energy’s top research priorities - Nextgov/FCW

AI, quantum computing, fusion energy remain Energy’s top research priorities Nextgov/FCW

GNews AI quantum

1m4 months ago

Knowledge Map

TopicsEntitiesSource

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 155 connections

Scroll to zoom · drag to pan · click to open

Discussion

No comments yet — be the first to share your thoughts!

SkillRouter: Skill Routing for LLM Agents at Scale

Submission history

Daily AI Digest

More about

[D] The memory chip market lost tens of billions over a paper this community would have understood in 10 minutes

[D] Is research in semantic segmentation saturated?

AI, quantum computing, fusion energy remain Energy’s top research priorities - Nextgov/FCW

Knowledge Map

Connected Articles — Knowledge Graph

Discussion

More in Research Papers

[D] Is research in semantic segmentation saturated?

AI, quantum computing, fusion energy remain Energy’s top research priorities - Nextgov/FCW

This Ancient Roman Game Board Was a Mystery. Researchers Used A.I. to Figure Out How to Play - Smithsonian Magazine

URI Day Highlights Student Research and the Future of AI Education in Rhode Island - uri.edu