๐Ÿงช
AI EngineeringTypeScript

Bare Eval

by yonatangross

Bare Eval is an AI Engineering skill for Claude Code, published by yonatangross in orchestkit.

225 stars24 forkson yonatangross/orchestkitAdded 2026/09/03Repository updated 2026/09/02
agent-orchestrationagentsai-agentsai-developmentanthropicclaude-codeclaude-code-pluginclaude-plugindeveloper-toolsfastapilanggraphllmmcpragreactsecuritytestingtypescript
Install in seconds
Install Bare Eval
Copy Bare Eval into your Claude Code skills folder. Run the command in your terminal, or review the source on GitHub before installing.
terminal
npx degit https://github.com/yonatangross/orchestkit/tree/main/src/skills/bare-eval ~/.claude/skills/bare-eval

Requires Node.js. Downloads this skill only โ€” not the rest of the repository โ€” into your Claude Code skills folder.

Without Node.js

git clone https://github.com/yonatangross/orchestkit.git

Clones the whole repository, then copy the skillโ€™s own directory into your skills folder yourself.

In this catalog

Source file
src/skills/bare-eval/SKILL.md in yonatangross/orchestkit
Installs to
~/.claude/skills/bare-eval
Collection
One of 25 skills cataloged from this repository
Category
AI Engineering โ€” 2793 skills

What Bare Eval does

Bare Eval runs `claude -p --bare` for isolated evaluation, trigger testing, and LLM grading without plugin or hook interference. Use it when benchmarking prompts, scoring outputs, or checking skill matches in a clean non-interactive call.

Bare Eval is cataloged under AI Engineering on DirSkills. Bare Eval comes from a repository tagged agent-orchestration, agents, ai-agents, ai-development and anthropic.

Documentation

README

Bare Eval โ€” Isolated Evaluation Calls

Run claude -p --bare for fast, clean eval/grading without plugin overhead.

CC 2.1.81 required. The --bare flag skips hooks, LSP, plugin sync, and skill directory walks.

When to Use

  • Grading skill outputs against assertions
  • Trigger classification (which skill matches a prompt)
  • Description optimization iterations
  • Any scripted -p call that doesn't need plugins

This is the opening of the README. Read the full README on GitHub.

Frequently asked about Bare Eval

  • What else does yonatangross publish alongside Bare Eval?

    Bare Eval is one of 25 skills that DirSkills catalogs from yonatangross/orchestkit, the repository it ships in. Its siblings there include AI UI Generation, API Design and Accessibility. Each one is a separate skill with its own page in this directory, installs the same way Bare Eval does, and is maintained by yonatangross in that same repository. The rest of the collection is listed on the yonatangross/orchestkit page.

  • How does Bare Eval compare to other AI Engineering skills?

    Bare Eval ranks #2529 by stars among the 2793 AI Engineering skills in this catalog. The most-starred ones next to it are Architecture Decision Records, AI-First Engineering and Agentic OS. DirSkills ranks by the star count of the repository each skill ships in, so that order reflects how popular those repositories are rather than any review of Bare Eval against them. Open each page to compare what they document and how they install.

More from yonatangross/orchestkit

Bare Eval is one of 25 skills cataloged on DirSkills from yonatangross/orchestkit.

See all 25 skills โ†’
๐ŸŽจ
7h ago

AI UI Generation

AI UI Generation covers prompt patterns, tool selection, review checklists, token injection, and CI gates for AI-produced UI. Use it when generating components or full-stack UI with json-render, v0.app, Stitch, Bolt Cloud, or Cursor.
AI Engineering
22524
๐Ÿ”Œ
7h ago

API Design

API Design covers REST and GraphQL contract design, versioning schemes, RFC 9457 Problem Details, and OpenAPI specs. Use it when defining endpoint shapes, deprecation windows, or standardizing error responses across services.
AI Engineering
22524
โ™ฟ
7h ago

Accessibility

Accessibility provides patterns for WCAG 2.2 AA compliance, keyboard focus management, React Aria components, and honoring user preferences. Use it when building screen-reader support, keyboard navigation, dialogs, forms, or reduced-motion and contrast handling.
Frontend
22524
๐Ÿ“Š
7h ago

Activation Audit

Activation Audit checks real spawn telemetry to see which OrchestKit sub-agents are being used, which never fire, and how often generic agents replace specialists. Use it before pruning the catalog or after changing agent spawn paths.
AI Engineering
22524
๐Ÿค–
7h ago

Agent Orchestration

Agent Orchestration provides patterns for agent loops, multi-agent coordination, framework selection, and multi-scenario workflows. Use it when building autonomous loops, routing work across agents, or comparing orchestration frameworks.
AI Engineering
22524
๐Ÿ“Š
7h ago

Analytics

Analytics queries local OrchestKit data to report agent usage, skill frequency, hook timing, team activity, session replays, cost estimates, and trends. Use it when reviewing performance or understanding usage patterns from privacy-safe local files.
Data
22524