skill-evaluation-graphlisted
Install: claude install-skill MaxLaurieHutchinson/skill-evaluation-graph
# SEG: Skill Evaluation Graph: Procedural Driver for Agent Skill Audit & Optimization
Turn agent skills into reliable, token-efficient, production-grade engineering assets. Evaluates trigger precision, progressive disclosure architecture, behavioral steering, execution determinism, operational safety, and token economics.
---
## Quick Reference Matrix
| Concern | Purpose | Authoritative Resource |
|:---|:---|:---|
| **Canonical Terminology** | Authoritative vocabulary & domain definitions | [references/terminology.md](references/terminology.md) |
| **6-Pillar Audit Rubric** | Objective 1–5 scoring definitions | [references/audit-rubric.md](references/audit-rubric.md) |
| **Anti-Pattern Catalog** | Diagnostic guide for 14 recurring skill defects | [references/anti-patterns.md](references/anti-patterns.md) |
| **Context Engineering** | Multi-tier progressive disclosure patterns | [references/progressive-disclosure-patterns.md](references/progressive-disclosure-patterns.md) |
| **Evaluation Graph & Loop** | Autonomous Evaluator Loop Engine runbook | [references/workflow-graph-and-evaluator-loop.md](references/workflow-graph-and-evaluator-loop.md) |
| **Harness Compatibility** | Cross-platform tool mappings (Antigravity/Claude/Codex) | [references/harness-tool-matrix.md](references/harness-tool-matrix.md) |
| **Behavioral Trial Runner** | Control vs. Treatment evaluation & live trials | [scripts/eval_skill.py](scripts/eval_skill.py) |
| **Capability Claim Audit** | Technical