cs.IRSep 30, 2026

SkillSeek: Revisiting Agent Skill Retrieval at Marketplace Scale

Authors: Guanqun Yang, Wenlong Zhang, Tian Shi, Ping Wang

Organizations: Stevens Institute of Technology, Hoboken, NJ, USA · Independent Researcher

Abstract

Anthropic's Agent Skills package reusable procedural know-how for an LLM agent into SKILL.md directories, and open-source aggregations have grown past 230,000 skills, making selection rather than authoring the bottleneck. The standing answer in the literature outsources selection to the agent itself: an LLM-mediated retrieval loop that rewrites queries and refines candidates inside the agent's decision loop, paying LLM tokens on every task. We present SkillSeek, an open-source two-stage skill retriever built from the standard IR recipe (a BGE-base bi-encoder feeding a small cross-encoder, exposed over MCP). Across a 4×114 \times 11 grid of pool, backbone, and method on the 89-task SkillsBench benchmark, SkillSeek reaches observed parity with the LLM-mediated loop of Liu et al. at essentially no extra cost: plain bm25 alone records a pass rate at or above their refined loop on three of four settings, and a small cross-encoder covers the remaining difference on the fourth. A first-stage recall ceiling explains the pattern, and total per-trial spend drops from USD 51.30 to USD 27.54 (within fifty cents of the no-skill baseline). Under the SkillsBench tasks and OpenHands harness we tested, this positions the standard IR recipe as a strong default for agent-skill retrieval, with LLM-mediated alternatives a natural fit for cases where deterministic methods fall short.

Figures & tables

Appendix figures & tables3 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. SkillRet: A Large-Scale Benchmark for Skill Retrieval in LLM Agents

    May 7, 2026Ryangkyung Kang, Hongcheol Cho, Youngeun KimSkillsLarge Language Model Agents

  2. SkillSight: Calibrating Generic Content Bias for Skill Retrieval

    Jul 21, 2026Jinying Xiao, Bin Li, Xiaopeng Li +7SkillsRetrievers

  3. SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests

    Aug 9, 2026Donghong Jiang, Endian Lin, Luoping Cui +6Agent Skill RetrievalSkills