2607.18785v1 Jul 21, 2026 cs.AI

SkillSight: Seeing Through Shared Descriptions for Accurate Skill Retrieval

Shasha Li
Shasha Li
Citations: 717
h-index: 11
Bing Ji
Bing Ji
Citations: 121
h-index: 5
Jie Yu
Jie Yu
Citations: 667
h-index: 9
Jinying Xiao
Jinying Xiao
Citations: 8
h-index: 2
Ma Jun
Ma Jun
Citations: 22
h-index: 2
Jiacheng Jie
Jiacheng Jie
Citations: 0
h-index: 0
Chao Wang
Chao Wang
Citations: 17
h-index: 2
N. Tashi
N. Tashi
Citations: 937
h-index: 17
Xiaodong Liu
Xiaodong Liu
Citations: 87
h-index: 5

As large language model agents gain access to increasingly large skill libraries, retrieving the right skill becomes critical to reliable capability selection and execution. Existing retrievers often treat skill descriptions as ordinary documents, overlooking their highly regular structure: shared descriptive patterns recur across many skills while providing little evidence for distinguishing the required capability. We show that this shared descriptive background systematically contributes to dense relevance scores, induces a pronounced energy gap between queries and skill documents, and obscures task-relevant signals. Based on this observation, we propose SkillSight, a training-free retrieval framework that calibrates shared background in both semantic and lexical spaces. Semantic Background Calibration estimates a background subspace from generic tokens identified by IDF, reducing similarity induced by shared descriptive patterns, while Lexical Evidence Calibration downweights shared background tokens to recover discriminative token-level evidence. Experiments on SRA-Bench and SkillBench-Supp demonstrate consistent improvements across retrieval metrics, with SkillSight improving Recall@10 by up to 20.21 percentage points over the original dense retriever. In end-to-end evaluation, SkillSight achieves the best overall performance across three agent models and outperforms LLM Selection by up to 4.97 percentage points. It is also up to 1,248 times faster than the Dense + Reranker baseline. These results identify shared descriptive background as a key source of bias in skill retrieval and demonstrate that explicitly calibrating it enables accurate and efficient skill selection without additional training. Our code is available at https://github.com/xiaojinying/SkillSight.

0 Citations
0 Influential
31.9657359028 Altmetric
159.8 Score
Original PDF
1

No Analysis Report Yet

This paper hasn't been analyzed by Gemini yet.

Log in to request an AI analysis.

댓글

댓글을 작성하려면 로그인하세요.

아직 댓글이 없습니다. 첫 번째 댓글을 남겨보세요!