arxiv
PublishedMay 26, 2026 at 4:00 AM
—neutral
When Skills Don't Help: A Negative Result on Procedural Knowledge for Tool-Grounded Agents in Offensive Cybersecurity
Publisher summary· verbatim
arXiv:2605.20023v2 Announce Type: replace Abstract: Agent Skills, structured packages of procedural knowledge loaded into an LLM agent at inference time, are widely reported to improve task pass rates by an average of 16.2~percentage points across diverse domains. Yet the same benchmarks show wide v
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivComMem: Complementary Memory Systems for Test-Time Adaptation of Vision-Language Models8harxivCustomized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline8harxivPreventing Error Propagation in Multi-Agent AI through Runtime Monitoring8harxivMemory as an Attack Surface in LLM Agents: A Study on Multiple-Choice Question Answering8hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗