defending-llms-with-guardrails
mukul975/Anthropic-Cybersecurity-Skills
Implements comprehensive runtime safety defenses for Large Language Model (LLM), RAG, and agent applications. This skill integrates industry-leading guardrail systems—including Llama Guard 3, NeMo Guardrails, and LLM Guard—to inspect and constrain all input and output data. It is essential for production environments requiring protection against adversarial attacks such as jailbreaks, prompt injection (OWASP LLM01), toxic content, and sensitive data leakage.