vitaliikapliuk/modelharness
Make every model cheaper or better. Zero-config behavioral harness for Claude Code, with a reproducible 408-run benchmark
ARCHETYPE
Pragmatist
A balanced, no-drama setup: some rules, some tools, nothing extreme.
Copy this rig
# review before running: this installs third-party code $ npx degit vitaliikapliuk/modelharness/agents ./rig-modelharness/agents $ npx degit vitaliikapliuk/modelharness/commands ./rig-modelharness/commands $ npx degit vitaliikapliuk/modelharness/skills ./rig-modelharness/skills $ npx degit vitaliikapliuk/modelharness/hooks ./rig-modelharness/hooks
MCP servers are added to Claude Code at local scope; env vars are shown as YOUR_… placeholders — we never store values. Files are fetched with degit into a separate folder so you can review before merging.
$ npx degit vitaliikapliuk/modelharness/skills/delegation-triggers .claude/skills/delegation-triggers $ npx degit vitaliikapliuk/modelharness/skills/memory-discipline .claude/skills/memory-discipline $ npx degit vitaliikapliuk/modelharness/skills/verification-loop .claude/skills/verification-loop
$ curl -fsSL --create-dirs -o .claude/agents/verifier.md https://raw.githubusercontent.<redacted>.md This rig commits no guardrails. Here is the community baseline instead — the deny/ask rules most often found across all 7,204 rigs:
{
"permissions": {
"deny": [
"Read(./.env)",
"Read(**/.env)",
"Read(~/.ssh/**)",
"Bash(rm -rf *)",
"Read(**/*.pem)",
"Bash(rm -rf /)",
"Bash(git push --force:*)",
"Bash(sudo *)",
"Read(.env)",
"Bash(rm -rf /*)",
"Read(./.env.*)",
"Read(~/.aws/**)",
"Bash(git push --force*)",
"Bash(rm -rf:*)",
"Read(**/*.key)",
"Read(**/.env.*)",
"Bash(sudo:*)",
"Bash(git reset --hard*)",
"Bash(git reset --hard:*)",
"Read(.env.*)"
],
"ask": [
"Bash(git push:*)",
"Bash(git push *)",
"Bash(git commit:*)",
"Bash(rm *)",
"Bash(rm:*)",
"Bash(git rebase *)",
"Bash(wget *)",
"Bash(npm publish:*)",
"Bash(git commit *)",
"Bash(gh pr merge *)"
]
}
} Skills (3)
Subagents (1)
| verifier | Fresh-context verifier. Use PROACTIVELY after completing any multi-step task - verifies work against its specification by running real checks, immune to the implementer's rationalizations. |
Slash commands (3)
/goal/retro/verify
Similar rigs
Remus3/Legion-Wallpaper
Self-auditing AI image restoration pipeline (super-resolution, LaMa inpainting, face repair, metric-plus-vision gate ladder), built by a multi-agent Claude Code system with independent verifiers, git-hook gates and headless background lanes
Automator 5.5k tok ·
leavemagic-cyber/earned-confidence
A behavior contract for AI coding agents — every rule added after a real failure. Ships with a verifier agent that cannot edit what it judges.
Pragmatist 3.6k tok ·
NotRedFox/NotRedFoxs-Claude-skills
Claude Code skills: a docs fact-checker, a code auditor, a bug-memory log, a mod and more.
Automator 577 tok ·
rlaope/quoted-andrej-karpathy
An harness and behavioral guidelines to enforce reliable execution in AI coding agents.
Minimalist 1.5k tok ·