mirror of
https://github.com/mukul975/Anthropic-Cybersecurity-Skills.git
synced 2026-09-03 06:50:51 +03:00
Rewrite 548 skill descriptions to the activation rubric
Each rewritten description now states both what the skill does (concrete capability, named tools/artifacts) and an explicit when-to-use trigger, improving agent discovery/activation. Grounded in each skill's own body; changes confined to the `description` field only (bodies and all other frontmatter untouched). Produced by a gated audit->rewrite->recheck loop (548 -> 0 flagged) with a sampled anti-invention check (0 ungrounded). Schema: 817/817 pass. Framework-ID gate: 0 defects.
This commit is contained in:
@@ -1,6 +1,10 @@
|
||||
---
|
||||
name: testing-for-system-prompt-leakage
|
||||
description: Extract and defend system prompts plus embedded secrets and routing logic.
|
||||
description: Extracts LLM system prompts using direct requests, jailbreak/instruction-override
|
||||
framing, translation/encoding tricks, and few-shot replay, combining manual payloads with
|
||||
automated garak and Promptfoo scanners to surface embedded secrets, routing logic, and
|
||||
policy leakage (OWASP LLM07:2025). Use during LLM application red-team engagements or
|
||||
when validating that no credentials or authorization logic live in the system prompt.
|
||||
domain: cybersecurity
|
||||
subdomain: ai-security
|
||||
tags:
|
||||
|
||||
Reference in New Issue
Block a user