Documentation
README
Agent Evaluation
This is a Hermes-native agent-evaluation workflow skill.
Why This Exists
agent-evaluation gives OMH a way to improve executor choice empirically, not by vibes, while preserving executor-neutral product language across Codex, Claude Code, Hermes, and generic runtimes.
Do Not Use When
- The user needs current runtime readiness only; use
executor-runtime-readiness. - The user already selected an executor and wants implementation; use the coding handoff or delivery workflow.
- The user asks for workflow learning from a single failed route; use
workflow-learning. - The ask is to find and fix runtime, memory, cost, or rendering hotspots rather than score executor or model output quality; use
ultraperf.
Examples
Good example:
This is the opening of the README. Read the full README on GitHub.