Login
Download
Skill UI
Browse and discover
15857+
curated skills
All
Development
Artificial Intelligence
Design & Creative
Product & Business
Data Science
Marketing
Soft Skills
Productivity
Engineering
Languages
Search
Model Improvement
, found
3
results
Default
Newest
Most Downloaded
Evolve AI Policy Surfaces for Optimization
harness-evolve
ruvnet/ruflo
412
This skill runs an automated evolution process designed to systematically optimize the seven core policy surfaces of an AI agent harness (e.g., planner, contextBuilder, reviewer). By mutating these policies and scoring the variants in a sandbox, it identifies crucial performance improvements without requiring expensive foundation model retraining. It is ideal for empirical configuration tuning and establishing optimized baselines.
View Details
Run GEPA Learning Cycle For Policy Optimization
harness-learn
ruvnet/ruflo
440
This tool executes a GEPA learning cycle, automating the evolution and optimization of policy prompts against a scored task corpus. Instead of relying on manual prompt iteration, it measures performance improvements, identifying the best working 'harness genome.' It is ideal for measuring sustainable improvements in task performance and supports a free, cost-estimating dry-run mode before committing to actual model calls.
View Details
AI Security Detection Benchmarking
harness-security-bench
ruvnet/ruflo
468
This tool executes the advanced `metaharness/darwin` security benchmark (ADR-155). It rigorously evaluates the security detection capabilities of AI systems against a controlled ground-truth corpus, providing critical metrics like TPR, FPR, and patch-pass rates. It compares the champion model against multiple baselines, offering empirical data crucial for iterative security improvement and drift detection.
View Details
1
Language
简体中文
English