技能 人工智能 提示注入安全分析

提示注入安全分析

v20260903
competition-prompt-injection
这是一个专业的下游技能,用于分析复杂智能体(Agent)系统中的安全漏洞。它专注于红队测试和渗透测试,旨在检测多步骤LLM链中的提示注入、数据投毒、内存污染、规划器漂移和工具滥用等边界穿越漏洞。当需要证明不可信输入如何危害系统完整性时使用。
获取技能
370 次下载
概览

Competition Prompt Injection

Use this skill only as a downstream specialization after $ctf-sandbox-orchestrator is already active and has established sandbox assumptions, node ownership, and evidence priorities. If that has not happened yet, return to $ctf-sandbox-orchestrator first.

Use this skill when the challenge is primarily about trust boundaries inside an agentic system.

Reply in Simplified Chinese unless the user explicitly requests English.

Quick Start

  1. Identify the first untrusted content that becomes model-visible.
  2. Map the chain from retrieval, memory, or transcript into planner or executor behavior.
  3. Record the exact point where text becomes a tool argument, file path, network target, or secret request.
  4. Prove one minimal exploit chain before exploring variants.
  5. Keep prompt snippets and tool transitions in compact evidence blocks.

Workflow

1. Map The Control Stack

  • Track system, developer, user, retrieved, memory, planner, and tool-response layers separately.
  • Distinguish claimed capability from runtime-exposed capability.
  • Note what the model can actually call, read, or mutate.

2. Prove The Boundary Crossing

  • Reproduce one chain from untrusted text to changed planner behavior, changed tool args, or secret exposure.
  • Keep the decisive transcript compact: source chunk, rewritten planner state, final tool invocation.
  • Prefer the smallest transcript that still demonstrates the bug.

3. Report By Boundary

  • State which layer failed: retrieval, summarizer, planner, executor, tool normalization, or output post-processing.
  • Separate instruction drift from actual side effect.

Read This Reference

  • Load references/prompt-injection.md for the checklist, evidence layout, and common prompt-boundary pitfalls.

What To Preserve

  • Original malicious chunk or prompt
  • Intermediate summary or planner drift if it matters
  • Final tool args, file paths, or exposed secret surface
信息
Category 人工智能
Name competition-prompt-injection
版本 v20260903
大小 2.43KB
更新时间 2026-09-05
语言