night-market-model-and-harness-updates

Featured

Sweep plugins, skills, agents, commands, and hooks after a model release or Claude Code version bump. Use when upstream ships. Do not use for routine edits; use night-market-change-control.

AI & Automation 339 stars 35 forks Updated 2 days ago MIT

Install

View on GitHub

Quality Score: 88/100

Stars 20%
84
Recency 20%
100
Frontmatter 20%
70
Documentation 15%
100
Issue Health 10%
50
License 10%
100
Description 5%
100

Skill Content

# Night Market Model and Harness Updates When Anthropic ships a model or Claude Code ships a version, the pins scattered through this repo rot silently. This skill runs the sweep that finds the rot, researches what actually changed, applies the updates, and records where upstream stood so the next run reports only the new delta. The watermark is the point. Without it every audit restarts from zero and re-derives the same answer by hand. `.claude/upstream-baseline.json` holds the last recorded upstream state, and each run diffs against it. ## When to run | Trigger | Signal | |---------|--------| | Model release | A tier or model ID ships that the ledger does not record | | Harness release | `claude --version` differs from the ledger | | Scheduled check | Monthly, to catch a release nobody noticed | ## The five steps Run them in order. Each one gates the next. ```bash # 1. Detect. Deterministic, no model in the loop. python3 scripts/check_upstream_drift.py # 2. Research what changed (only when step 1 reports drift). # Release notes and model cards are mandatory sources. # 3. Map findings onto asset classes. # 4. Sweep the implicated classes. # 5. Prove, then record the new watermark. python3 scripts/check_upstream_drift.py && \ python3 scripts/check_agent_model_matrix.py ``` Step 5 runs before the ledger is written, never after. Recording a migration that has not passed its proof is the failure the ledger exists to prevent. ## What the detector proves and what i...

Details

Author
athola
Repository
athola/claude-night-market
Created
10 months ago
Last Updated
2 days ago
Language
Python
License
MIT

Integrates with

Bundled in these plugins

Similar Skills

Semantically similar based on skill content — not just same category

AI & Automation Listed

generation-audit

Single entry point for auditing the harness when a new Claude model takes one of its roles (judge tier / build tier / everyday session). Runs two checks in one pass — the runtime-layer cross-check (live system prompt + tool descriptions vs rules / CLAUDE.md / output style / skills / agents: conflict and redundancy) and the dated-pattern scan of every self-authored asset via `/claude-api prompt-audit` — then applies the line fixes under this harness's writing rules and hands asset-level verdicts to rules-stocktake / skill-stocktake / agent-stocktake. Invoke with /generation-audit when a new Claude model takes a harness role. NOT for — rewriting one skill (skill-creator); routine stocktakes (call them directly); rule compliance (skill-comply); whole-config GC (config-gc).

3 Updated today
shimo4228
AI & Automation Listed

generation-audit

Single entry point for auditing the harness when a new Claude model takes one of its roles (judge tier / build tier / everyday session). Runs two checks in one pass — the runtime-layer cross-check (live system prompt + tool descriptions vs rules / CLAUDE.md / output style / skills / agents: conflict and redundancy) and the dated-pattern scan of every self-authored asset via `/claude-api prompt-audit` — then applies the line fixes under this harness's writing rules and hands asset-level verdicts to rules-stocktake / skill-stocktake / agent-stocktake. Invoke with /generation-audit when a new Claude model takes a harness role. NOT for — rewriting one skill (skill-creator); routine stocktakes (call them directly); rule compliance (skill-comply); whole-config GC (config-gc).

0 Updated today
shimo4228
AI & Automation Listed

refreshing-anthropic-guidance

Sweeps the local Claude Code harness — agents, skills, hooks, rules, workflows, CLAUDE.md — against Anthropic's current guidance/models, diffs last-verified state, plans remediation. Use for /refreshing-anthropic-guidance, "is our harness up to date with Anthropic", "model pins current". Distinct from researching-anthropic-guidance and auditing.

0 Updated 1 weeks ago
monte3l