testing-for-system-prompt-leakage
mukul975/Anthropic-Cybersecurity-Skills
Detailed methods for red-teaming and penetration testing of LLM applications. This skill extracts sensitive system prompts, secrets, and embedded logic using advanced techniques like jailbreaking, encoding tricks, and automated scanners (Garak, Promptfoo), ensuring compliance with OWASP LLM07:2025.