UiPath/coder_eval
Playwright for coding agents. Test that your skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, A/B experiments, CI gates.
ARCHETYPE
Pragmatist
A balanced, no-drama setup: some rules, some tools, nothing extreme.
Copy this rig
# review before running: this installs third-party code
$ npx degit UiPath/coder_eval/.claude ./rig-coder_eval # inspect, then merge into .claude/ MCP servers are added to Claude Code at local scope; env vars are shown as YOUR_… placeholders — we never store values. Files are fetched with degit into a separate folder so you can review before merging.
$ npx degit UiPath/coder_eval/plugins/coder-eval/skills/analyze .claude/skills/analyze $ npx degit UiPath/coder_eval/plugins/coder-eval/skills/check-skill .claude/skills/check-skill $ npx degit UiPath/coder_eval/plugins/coder-eval/skills/ci .claude/skills/ci $ npx degit UiPath/coder_eval/plugins/coder-eval/skills/init .claude/skills/init $ npx degit UiPath/coder_eval/plugins/coder-eval/skills/lint-tasks .claude/skills/lint-tasks $ npx degit UiPath/coder_eval/plugins/coder-eval/skills/task .claude/skills/task
This rig commits no guardrails. Here is the community baseline instead — the deny/ask rules most often found across all 6,974 rigs:
{
"permissions": {
"deny": [
"Read(./.env)",
"Read(**/.env)",
"Read(~/.ssh/**)",
"Bash(rm -rf *)",
"Read(**/*.pem)",
"Bash(rm -rf /)",
"Bash(git push --force:*)",
"Bash(sudo *)",
"Read(~/.aws/**)",
"Bash(rm -rf /*)",
"Read(./.env.*)",
"Read(.env)",
"Bash(git push --force*)",
"Bash(rm -rf:*)",
"Read(**/.env.*)",
"Read(**/*.key)",
"Bash(sudo:*)",
"Bash(git reset --hard*)",
"Bash(git reset --hard:*)",
"Read(.env.*)"
],
"ask": [
"Bash(git push:*)",
"Bash(git push *)",
"Bash(git commit:*)",
"Bash(rm *)",
"Bash(rm:*)",
"Bash(npm publish:*)",
"Bash(wget *)",
"Bash(git rebase *)",
"Bash(gh pr merge *)",
"Bash(git commit *)"
]
}
} Skills (6)
Slash commands (6)
/coder-eval-code-review-full/coder-eval-code-review-wf/coder-eval-code-review/coder-eval-create-plan/coder-eval-implement-plan/coder-eval-review
Similar rigs
bentleypark/claude-code-mobile-spine
Claude Code plugin marketplace bundling mobile-spine — scaffolds a meta-repo coordinating 4 subagents (api/pm/android/ios) for mobile teams whose Android, iOS, and Backend live in 3 separate repos.
Minimalist 84 tok ·
yongkyung-oh/agent-bootstrap
A minimal workflow bootstrap for AI coding agents — three markdown files that solve context loss between sessions. No framework, no dependencies.
Minimalist 1.7k tok ·
oter/bruh
Claude Code plugin for coordinating work across sessions and machines: coordinator, clankers, clerks, workflows
Pragmatist 5.2k tok ·
sravan27/context-os
Cut Claude Code token usage by 40.9% and ship a CI gate for coding-agent cost leaks. Stdlib Python hook + GitHub Action for Claude Code, Codex, Cursor, and agentic coding teams.
Pragmatist 680 tok ·