fix(kyberforge): fix tests/ read gap, evals/ loop, and description nav pointers
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_016z2ZFYHQCex8yZAMVMTZzZ
This commit is contained in:
@@ -29,7 +29,6 @@ If the skill dir is missing, ask for it. If no signals are present, stop: "This
|
||||
Signals can come from anywhere in the conversation or referenced files:
|
||||
- Grill session output (most common predecessor in the factory sequence)
|
||||
- `/skill-audit` findings (PASS/FAIL/SUGGESTION punch list)
|
||||
- Eval assertion failures (grading.json, benchmark.json)
|
||||
- Human feedback (feedback.json, inline in conversation, PR or issue comments)
|
||||
- Session context describing what went wrong
|
||||
|
||||
@@ -41,7 +40,7 @@ Group signals by **root cause**, not symptom. Ask: "What single gap in the skill
|
||||
|
||||
```text
|
||||
Example:
|
||||
- Eval fails because output format is wrong
|
||||
- Session context: output format is wrong on every run
|
||||
- Audit finding: no output template defined
|
||||
- User feedback: "I always have to ask it to format the output"
|
||||
→ Root cause: SKILL.md has no output format specification → one fix: add an output template
|
||||
@@ -72,5 +71,3 @@ If a signal points to a script or reference file, edit that file directly rather
|
||||
## Step 5 — Validate and close
|
||||
|
||||
Run `/skill-audit` on the skill directory. Resolve any FAIL findings before considering the improvement complete.
|
||||
|
||||
If the skill has no `evals/` directory, note it after the audit: "No evals found — consider adding an `evals/` directory with assertion-based test cases to give future improvement cycles quantitative signals to work from."
|
||||
|
||||
Reference in New Issue
Block a user