Automated Marketplace-Scale Vetting of Agent Skills
Develop automated vetting pipelines for Agent Skills that combine static analysis of bundled scripts, semantic analysis of natural language instructions, dynamic analysis in sandboxed environments, and provenance verification of dependencies, while maintaining low false positive rates suitable for marketplace-scale review.
References
Developing automated vetting pipelines that combine static analysis of bundled scripts, semantic analysis of natural language instructions, dynamic analysis of Skill behavior in sandboxed environments, and provenance verification of declared dependencies---while maintaining low false positive rates to avoid blocking legitimate Skills---is an open engineering and research challenge.
Reliable malicious Agent Skill detection therefore remains open, and progress requires diverse malicious and benign coverage together with source-aware, two-sided evaluation.