Mar 17, 2026arXiv:2603.16572

Malicious Or Not: Adding Repository Context to Agent Skill Classification

Florian Holzbauer, David Schmidt, Gabriel Gegenhuber, G. Gegenhuber, S. Schrittwieser, Sebastian Schrittwieser, Johanna Ullrich

AI Summary

This paper presents a large-scale empirical security analysis of the AI agent skill ecosystem, collecting 238,180 skills from distribution platforms and GitHub. By incorporating repository context from GitHub, the study significantly reduces the false positive rate of malicious skill detection compared to solely analyzing skill descriptions. The analysis reveals a more accurate risk surface, identifying only 0.52% of skills in malicious repositories as non-benign and uncovering new attack vectors like hijacking skills on abandoned GitHub repositories.

Key Contribution

Security scanners flag nearly half of AI agent skills as malicious, but adding GitHub repository context reveals that the true number is closer to 0.5%.

Abstract

Agent skills extend local AI agents, such as Claude Code or Open Claw, with additional functionality, and their popularity has led to the emergence of dedicated skill marketplaces, similar to app stores for mobile applications. Simultaneously, automated skill scanners were introduced, analyzing the skill description available in SKILL.md, to verify their benign behavior. The results for individual market places mark up to 46.8% of skills as malicious. In this paper, we present the largest empirical security analysis of the AI agent skill ecosystem, questioning this high classification of malicious skills. Therefore, we collect 238,180 unique skills from three major distribution platforms and GitHub to systematically analyze their type and behavior. This approach substantially reduces the number of skills flagged as non-benign by security scanners to only 0.52% which remain in malicious flagged repositories. Consequently, out methodology substantially reduces false positives and provides a more robust view of the ecosystem's current risk surface. Beyond that, we extend the security analysis from the mere investigation of the skill description to a comparison of its congruence with the GitHub repository the skill is embedded in, providing additional context. Furthermore, our analysis also uncovers several, by now undocumented real-world attack vectors, namely hijacking skills hosted on abandoned GitHub repositories.

Code Generation & Program Synthesis Red-Teaming & Adversarial Robustness Tool Use & Agents

Citation Metrics

Citations0

Influential citations0

References30

Year2026

VenueN/A

Related Papers

Finding related papers...

Search

Malicious Or Not: Adding Repository Context to Agent Skill Classification

Related Papers