ln-510-quality-coordinator
Use when coordinating story quality evaluation with mandatory research, worker summaries, agent review, regression evidence, and bounded refinement.
What this skill does
> **Paths:** File paths (`references/`, `../ln-*`) are relative to this skill directory.
**Type:** L2 Coordinator
**Category:** 5XX Quality
# Quality Coordinator
Evaluation-platform coordinator for story quality review.
## Mandatory Read
**MANDATORY READ:** Load `references/evaluation_coordinator_runtime_contract.md`, `references/evaluation_summary_contract.md`, `references/evaluation_research_contract.md`, `references/loop_health_contract.md`
**MANDATORY READ:** Load `references/agent_delegation_pattern.md`
**MANDATORY READ:** Load `references/criteria_validation.md`, `references/gate_levels.md`
Agent review policy: run health check, record skipped reason when no advisor is available, verify every advisor claim before verdict, and treat transport/auth/tool failures as operator evidence rather than quality findings. Load `references/agent_review_workflow.md` only when debugging lifecycle/liveness details outside the evaluation runtime.
## Purpose
- invoke `ln-511-code-quality-checker`
- invoke `ln-512-tech-debt-cleaner`
- invoke `ln-513-regression-checker`
- invoke `ln-514-test-log-analyzer`
- run inline agent review in parallel with read-only evidence gathering
- keep merge, refinement, and verdict sequential
- return normalized quality results
## Inputs
Primary input:
- `storyId`
- `--previous-cycle-focus` (optional, from ln-500): comma-separated blocking categories from prior FAIL cycle
Status filter:
- `To Review`
## Critical Rule
Fast-track paths that skip research are not allowed.
Every quality run must include:
1. official documentation or standards
2. MCP Ref
3. Context7 when a framework or library is involved
4. current web best-practice research
## Runtime Contract
Runtime family:
- `evaluation-runtime`
Identifier:
- `quality-{storyId}`
Phase order:
1. `PHASE_0_CONFIG`
2. `PHASE_1_DISCOVERY`
3. `PHASE_2_READ_ONLY_EVIDENCE`
4. `PHASE_3_CLEANUP`
5. `PHASE_4_AGENT_BARRIER`
6. `PHASE_5_MERGE`
7. `PHASE_7_REFINEMENT`
8. `PHASE_8_VERDICT`
9. `PHASE_9_SELF_CHECK`
## Worker Invocation (MANDATORY)
**Host Skill Invocation:** `Skill(skill: "...", args: "...")` is mandatory delegation.
- Claude: call the Skill tool exactly as shown.
- Codex: if no Skill tool exists, locate the named skill in available skills, read its `SKILL.md`, treat `args` as `$ARGUMENTS`, execute that skill workflow, then return here with its result/artifact.
- Do not inline worker logic or mark the worker complete without executing the target skill.
Use the Skill tool for delegated workers. Do not inline worker logic inside the coordinator.
TodoWrite format (mandatory):
- `Resolve Story and build runtime manifest`
- `Load Story metadata and detect changed files`
- `Run quality checkers and research in parallel`
- `Apply safe tech-debt cleanup`
- `Sync agents and wait for all evidence`
- `Merge and deduplicate all findings`
- `Run bounded refinement loop`
- `Compute quality verdict and score`
- `Verify runtime cleanup and self-check`
Representative invocations:
```text
Skill(skill: "ln-511-code-quality-checker", args: "{storyId}")
Skill(skill: "ln-512-tech-debt-cleaner", args: "{storyId}")
Skill(skill: "ln-513-regression-checker", args: "{storyId}")
Skill(skill: "ln-514-test-log-analyzer", args: "{storyId}")
```
## Workflow
### Phase 0: Config
1. Resolve `storyId`.
2. Build evaluation runtime manifest with `required_research=true`.
3. Start `evaluation-runtime`.
### Phase 1: Discovery
1. Load Story metadata and completed implementation task scope.
2. Detect changed files and project stack.
3. Index semantic graph when available.
### Phase 2: Read-Only Evidence
Parallel work allowed in this phase:
- inline research by the coordinator (per `references/evaluation_research_contract.md`)
- `ln-511-code-quality-checker`
- `ln-513-regression-checker`
- `ln-514-test-log-analyzer`
- external agent launch
Rules:
- research is mandatory
- worker summaries are the only completion signal
- no merge or mutation occurs in this phase
When `previous_cycle_focus` is provided:
- Prioritize evidence collection for the listed blocking categories.
- ln-511 code quality checker should focus on the specified areas first.
- This does not exclude other evidence — it reorders priority.
### Phase 3: Cleanup
1. Run `ln-512-tech-debt-cleaner` only after read-only evidence is collected.
2. Cleanup remains sequential because it mutates files.
3. Record the worker summary and any cleanup evidence.
### Phase 4: Agent Barrier
1. Sync agents through `evaluation-runtime`.
2. Do not cross this barrier until all required agents are resolved or explicitly skipped.
3. Treat `failure_class` from agent results as transport evidence:
- `rate_limited`, `tool_missing`, `auth_missing`, `permission_denial`, and `asked_question` are not quality FAIL findings by themselves.
- `timeout_productive` can continue to merge/review only when output/log/session evidence exists.
- repeated identical worker/agent failure without new artifacts pauses through loop health before another cycle.
### Phase 5: Merge
Merge inputs:
- inline research evidence
- `ln-511` summary
- `ln-512` summary
- `ln-513` summary
- `ln-514` summary
- agent findings
Rules:
- deduplicate before scoring
- unsupported claims are rejected
- security and correctness issues remain high priority
### Phase 6: Refinement
Refinement uses a 2-stage state machine per `references/agents/prompt_templates/iterative_refinement.md` and `references/agents/prompt_templates/refinement_perspectives.md`:
- Stage 1 (parallel): `dry_run_executor`, `new_dev_tester`, `adversarial_reviewer`
- Stage 2 (after merge): `final_sweep`
Rules:
- Stage 1 runs in parallel, Stage 2 after merge
- cleanup evidence required for spawned processes
- no research skipping
### Phase 7: Verdict
Compute normalized quality verdict using:
- code quality
- cleanup result
- agent review
- criteria validation
- linter result
- regression result
- log analysis result
Final verdict values:
- `PASS`
- `CONCERNS`
- `FAIL`
### Phase 8: Self-Check
Required checks:
- [ ] runtime started
- [ ] mandatory research completed
- [ ] all worker summaries recorded
- [ ] all required agents resolved before merge
- [ ] cleanup verified
- [ ] refinement trace recorded when applicable
- [ ] coordinator summary written
## Summary Contract
Write `summary_kind=evaluation-coordinator`.
Recommended payload:
- `status`
- `final_result`
- `report_path`
- `worker_count`
- `agent_count`
- `issues_total`
- `severity_counts`
- `warnings`
- `cleanup_verified`
- `research_completed`
## Definition of Done
- [ ] Evaluation runtime started
- [ ] Mandatory research completed
- [ ] Read-only evidence workers completed
- [ ] Cleanup worker completed or justified
- [ ] Agent barrier resolved
- [ ] Merge completed
- [ ] Refinement executed or explicitly justified
- [ ] Final verdict calculated
- [ ] `evaluation-coordinator` summary written
- [ ] Runtime completed
## Meta-Analysis
Optional reference: load `references/meta_analysis_protocol.md` only when the user asks for post-run meta-analysis or protocol-formatted run reflection.
When requested after the coordinator run, analyze the session per protocol section 7 and include the protocol-formatted output with the final quality verdict.
## References
- Runtime: `references/evaluation_coordinator_runtime_contract.md`, `references/evaluation_summary_contract.md`
- Research: `references/evaluation_research_contract.md`
- Workers: `../ln-511-code-quality-checker/SKILL.md`, `../ln-512-tech-debt-cleaner/SKILL.md`, `../ln-513-regression-checker/SKILL.md`, `../ln-514-test-log-analyzer/SKILL.md`
- Quality criteria: `references/criteria_validation.md`, `references/gate_levels.md`
---
**Version:** 7.0.0
**Last Updated:** 2026-02-09
Related in AI Agents
skill-development
IncludedComprehensive meta-skill for creating, managing, validating, auditing, and distributing Claude Code skills and slash commands (unified in v2.1.3+). Provides skill templates, creation workflows, validation patterns, audit checklists, naming conventions, YAML frontmatter guidance, progressive disclosure examples, and best practices lookup. Use when creating new skills, validating existing skills, auditing skill quality, understanding skill architecture, needing skill templates, learning about YAML frontmatter requirements, progressive disclosure patterns, tool restrictions (allowed-tools), skill composition, skill naming conventions, troubleshooting skill activation issues, creating custom slash commands, configuring command frontmatter, using command arguments ($ARGUMENTS, $1, $2), bash execution in commands, file references in commands, command namespacing, plugin commands, MCP slash commands, Skill tool configuration, or deciding between skills vs slash commands. Delegates to docs-management skill for official documentation.
reprompter
IncludedTransform messy prompts into well-structured, effective prompts — single or multi-agent. Use when: "reprompt", "reprompt this", "clean up this prompt", "structure my prompt", rough text needing XML tags and best practices, "reprompter teams", "repromptception", "run with quality", "smart run", "smart agents", multi-agent tasks, audits, parallel work, anything going to agent teams. Don't use when: simple Q&A, pure chat, immediate execution-only tasks. See "Don't Use When" section for details. Outputs: Structured XML/Markdown prompt, quality score (before/after), optional team brief + per-agent sub-prompts, agent team output files. Success criteria: Single mode quality score ≥ 7/10; Repromptception per-agent prompt quality score 8+/10; all required sections present, actionable and specific.
adaptive-compaction
IncludedAdaptive add-on policy and recovery layer that decides WHEN to compact, prune, snapshot, or fork -- replacing fixed-percent auto-compaction across Claude Code, Codex, and MCP-capable hosts. Trigger on auto-compact timing or damage: "when should I compact", "is it safe to compact now or start a fresh session", "auto-compact fires too early/mid-task", "switching to an unrelated task but the window still has space", "context rot", "answers get worse the longer the session runs", "the agent forgot the plan or my decisions after it summarized", "add a layer on top that manages context without changing the agent", raising autoCompactWindow to give the policy room, or installing/tuning a cross-tool compaction policy or PreCompact hook -- even when "compaction" is never said but the problem is context-window pressure or post-summarization memory loss. Do NOT use to summarize a conversation, build RAG, write a summarization prompt (decides WHEN not HOW), or answer max-context-length trivia.
agent-skill-creator
IncludedCreate cross-platform agent skills from workflow descriptions. Activates when users ask to create an agent, automate a repetitive workflow, create a custom skill, or need advanced agent creation. Triggers on phrases like create agent for, automate workflow, create skill for, every day I have to, daily I need to, turn process into agent, need to automate, create a cross-platform skill, validate this skill, export this skill, migrate this skill. Supports single skills, multi-agent suites, transcript processing, template-based creation, interactive configuration, cross-platform export, and spec validation.
llm-wiki
IncludedUse when building or maintaining a persistent personal knowledge base (second brain) in Obsidian where an LLM incrementally ingests sources, updates entity/concept pages, maintains cross-references, and keeps a synthesis current. Triggers include "second brain", "Obsidian wiki", "personal knowledge management", "ingest this paper/article/book", "build a research wiki", "compound knowledge", "Memex", or whenever the user wants knowledge to accumulate across sessions instead of being re-derived by RAG on every query.
skill-master
IncludedAgent Skills authoring, evaluation, and optimization. Create, edit, validate, benchmark, and improve skills following the agentskills.io specification. Use when designing SKILL.md files, structuring skill folders (references, scripts, assets), ingesting external documentation into skills, running trigger evals, benchmarking skill quality, optimizing descriptions, or performing blind A/B comparisons. Keywords: agentskills.io, SKILL.md, skill authoring, eval, benchmark, trigger optimization.