Login
Download
Skill UI
Browse and discover
18671+
curated skills
All
Development
Artificial Intelligence
Design & Creative
Product & Business
Data Science
Marketing
Soft Skills
Productivity
Engineering
Languages
Search
Test
, found
2335
results
Default
Newest
Most Downloaded
Robustness and Reproducibility Stress Testing
jse-tju-robustness-reproducibility
brycewang-stanford/Awesome-Journal-Skills
282
Stress-test core claims for Journal of Systems Engineering submissions: parameter sensitivity, initial conditions, alternative models and metrics, extreme scenarios, Monte Carlo repetitions, random seeds, software environments, data/code statements, and explicit failure boundaries.
View Details
Systems Engineering Topic Selection
jse-tju-topic-selection
brycewang-stanford/Awesome-Journal-Skills
475
This skill guides researchers in selecting and refining topics for the Journal of Systems Engineering (Tianjin University). It helps define system boundaries, identify feedback loops, and ensure testable questions. It prevents vague topics by requiring specific system structures and verifiable outcomes. Suitable for academic research in systems engineering, operations research, and complex systems.
View Details
MCP Server Release Quality Assurance
mcp-release-qa
github/awesome-copilot
478
This skill verifies an MCP server before release by exercising a real protocol session. It compares runtime capabilities with source code and documentation, tests failure paths, and records reproducible evidence. Use this when shipping or reviewing an MCP server, tool, resource, prompt, catalog, or install path to ensure protocol behavior and transport correctness.
View Details
OpenSquilla Skill Hub Canary Fixture
opensquilla-live-skill-hub-canary
TokenRhythm/opensquilla
486
This is a synthetic, instructional fixture designed solely for verifying the delivery and access of community skills within the OpenSquilla live Skill Hub canary environment. It is crucial for testing platform capability (e.g., resource access validation) and should not be executed for actual functionality or command processing.
View Details
Roslyn Analyzer Development
roslyn-analyzers
github/awesome-copilot
436
Build, review, debug, package, and test Roslyn diagnostic analyzers, code fix providers, and incremental source generators. Covers IOperation analysis, dependency pinning, test harnesses, and NuGet packaging.
View Details
Automated Test Gap Audit
test-gap-audit
github/awesome-copilot
221
This skill performs a read-only audit to identify missing, weak, or stale test coverage in codebases. It analyzes full repositories or specific scopes like PRs, features, and APIs to recommend concrete test cases. It prioritizes high-risk areas and ensures behavior is properly covered without executing tests or performing security reviews.
View Details
WCAG Accessibility Audit and Testing
accessibility-compliance-accessibility-audit
sickn33/agentic-awesome-skills
329
This expert conducts comprehensive audits to ensure digital products comply with WCAG standards. It specializes in identifying accessibility barriers, testing user experiences for users with disabilities, and providing step-by-step remediation guidance. Use it for auditing web or mobile platforms, establishing compliance, and improving overall inclusivity.
View Details
Active Directory Attack Techniques
active-directory-attacks
sickn33/agentic-awesome-skills
75
Comprehensive offensive techniques for Microsoft Active Directory environments including reconnaissance, credential harvesting, Kerberos attacks, lateral movement, privilege escalation, and domain dominance for authorized red team operations and penetration testing.
View Details
AI Agent Evaluation Suites
agent-evals
sickn33/agentic-awesome-skills
180
Build automated evaluation suites for AI agents using golden datasets, rubrics, and regression gates. Covers prompt-level unit evals, tool-calling checks, end-to-end tasks, safety/adversarial tests, and LLM-as-judge scoring to gate deployments on quality.
View Details
AI Agent Behavior Evaluation
agent-evaluation
sickn33/agentic-awesome-skills
492
Evaluate AI agent behavior using versioned cases and explicit verifiers. This skill helps compare agent or prompt changes, reproduce failures, and run regression tests. It ensures reliability on declared task distributions without inferring product readiness from generic scores. Includes methods for uncertainty reporting and safety verification.
View Details
AI Agent Evaluation Reporting
agent-evaluation-reporting
sickn33/agentic-awesome-skills
179
Generate decision-ready reports from AI agent evaluation runs. This skill ensures autonomous, assisted, failed, and timed-out outcomes remain distinct and comparable. It defines metrics with correct denominators, handles latency and cost populations honestly, and maps evidence to decision gates. Ideal for benchmarking, regression testing, and production readiness validation of AI agents.
View Details
Agent Harness Fault Injection
agent-harness-fault-injection
sickn33/agentic-awesome-skills
500
This skill enables deterministic fault injection testing for agent workflows to verify recovery mechanisms, state preservation, and safety boundaries. It generates fault matrices and event timelines to distinguish between recovered, contained, and unrecoverable failures without impacting production systems or real user data.
View Details
Prev
1
2
3
...
168
169
170
171
172
173
174
...
193
194
195
Next
Language
简体中文
English