Malicious Or Not: Adding Repository Context to Agent Skill Classification

Florian Holzbauer; David Schmidt; Gabriel Gegenhuber; Sebastian Schrittwieser; Johanna Ullrich

arXiv:2603.16572·cs.CR·March 18, 2026

Malicious Or Not: Adding Repository Context to Agent Skill Classification

Florian Holzbauer, David Schmidt, Gabriel Gegenhuber, Sebastian Schrittwieser, Johanna Ullrich

PDF

Open Access

TL;DR

This paper conducts the largest empirical security analysis of AI agent skills, reducing false positives in malicious classification by incorporating repository context and uncovering new attack vectors.

Contribution

It introduces a methodology that combines skill description analysis with repository context to improve malicious skill detection accuracy.

Findings

01

Reduces false positives from 46.8% to 0.52% in security scans.

02

Uncovers undocumented attack vectors like hijacked abandoned repositories.

03

Provides a comprehensive view of the current risk surface in AI agent ecosystems.

Abstract

Agent skills extend local AI agents, such as Claude Code or Open Claw, with additional functionality, and their popularity has led to the emergence of dedicated skill marketplaces, similar to app stores for mobile applications. Simultaneously, automated skill scanners were introduced, analyzing the skill description available in SKILL.md, to verify their benign behavior. The results for individual market places mark up to 46.8% of skills as malicious. In this paper, we present the largest empirical security analysis of the AI agent skill ecosystem, questioning this high classification of malicious skills. Therefore, we collect 238,180 unique skills from three major distribution platforms and GitHub to systematically analyze their type and behavior. This approach substantially reduces the number of skills flagged as non-benign by security scanners to only 0.52% which remain in malicious…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Malware Detection Techniques · Web Application Security Vulnerabilities · Spam and Phishing Detection