mirror of
https://github.com/garrytan/gstack.git
synced 2026-05-02 03:35:09 +02:00
b5b2a15ad2
- Add severity classification to qa/SKILL.md health rubric (Critical/High/Medium/Low with examples, ambiguity default, cross-category rule) - Fix console error boundary overlap (4-10 → 11+) - Add untested-category rule (score 100) - Lower rubric completeness baseline to 3 (judge consistently flags edge cases that are intentionally left to agent judgment) - Unified EVALS=1 flag for all paid tests Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
8 lines
388 B
JSON
8 lines
388 B
JSON
{
|
|
"command_reference": { "clarity": 4, "completeness": 4, "actionability": 4 },
|
|
"snapshot_flags": { "clarity": 4, "completeness": 4, "actionability": 4 },
|
|
"browse_skill": { "clarity": 4, "completeness": 4, "actionability": 4 },
|
|
"qa_workflow": { "clarity": 4, "completeness": 4, "actionability": 4 },
|
|
"qa_health_rubric": { "clarity": 4, "completeness": 3, "actionability": 4 }
|
|
}
|