calibrate
SolidCalibrate skills/role cards for leaks/gaps with recall, precision, and confidence-accuracy checks.
Install
Quality Score: 82/100
Skill Content
Details
- Author
- Borda
- Repository
- Borda/AI-Rig
- Created
- 7 months ago
- Last Updated
- 3 days ago
- Language
- Python
- License
- Apache-2.0
Similar Skills
Semantically similar based on skill content — not just same category
calibrate
Reflective end-of-session self-improvement. Scans the current Claude Code session for corrections, preferences, repeated patterns, errors, success patterns, and voice violations, then proposes numbered concrete patches to memory, settings, ceo-only skills, and ceo-only rules. Corporate files route to a separate review queue and are NEVER auto-applied. Use at end of every working session. Light mode for low-token state or quick sweeps. CEO-only - never propagates to execs.
audit
Audit Codex configuration/workflow drift; emit ranked gaps and measurable gates.
forge-calibrate
This skill should be used to measure how reliable the project's code review is — phrases like "forge-calibrate", "calibrate the reviewer", "how consistent is forge-review", "is the review flaky", "test the reviewer", or "review golden set". It runs forge-review's lenses several times over a golden set of small diffs with known findings, scores recall (did it find what's planted) and run-to-run agreement (does it find the same things twice), and names the blind spots. Reports numbers; changes nothing.