~ / leaderboard / skill / bench
Skill · #2018
bench
Formal benchmark harness. Runs a metric command N times across 2-3 variants (git refs or state-prep shell commands), checks variance, computes delta vs. declared baseline, and emits a reproducible TSV
4rigs use it
0%of all rigs
30.1kavg rig context tax
Get it
$ npx degit ddalcu/mlx-serve/.claude/skills/bench .claude/skills/bench From the most-starred rig that ships it: ddalcu/mlx-serve. Same-named skills in different rigs may differ — review before use.
Often used together with
Rigs using bench (4)
ddalcu/mlx-serve
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Zig backend, Swift frontend macOS app with chat, music, voice, video generation.
Pragmatist 71.4k tok ·
autonomous-ai/openharness
The ultimate harness for coding agents and beyond. All your agents. All your machines. One command center. Start with code, then follow your curiosity and build across disciplines: CAD, circuits, robots, games and music.
Skill Collector 23.3k tok ·
arbazkhan971/godmode
Autonomous AI coding agent — 126 skills, 7 subagents, 5 platforms. Iterative optimization with automatic rollback, failure memory, and parallel multi-agent execution. Plugin for Claude Code, Cursor, Codex, Gemini CLI, OpenCode.
Skill Collector 20.8k tok ·
ocrosby/claude-config
Claude Code configuration
YOLO Cowboy 5.1k tok ·