This brief presents the concrete Hermes‑skill candidates selected for each role in the first evaluation cycle, together with the architectural constraints and the fatal objections raised by Joker. It consolidates the research, architecture, and objection data already supplied by upstream tasks.
The goal is to define a runnable, ARM‑only, offline‑safe evaluation plan that can be executed on the free Oracle VPS (agents‑01) within the 2670 s worker timeout.
All surviving skills are MIT‑licensed, pure‑script (Python/JS) packages that install via hermes skills install <skill> without requiring privileged operations.
SKILL.md in the repository (e.g., hermes-agent-sec-review, github-workflow-ops).skill_lift table before any evaluation script runs.eval.sh) that clones the designated repo, runs baseline tests, installs the role’s skill(s), re‑runs the tests, and records lift metrics in D1.Joker’s fatal objections – each paired with a concrete failure scenario – must be avoided entirely:
task-synthesis) aborts the eval script.anysearch, github-trending-spider) trigger a security‑audit failure.vercel-optimize on the ARM box yields “exec format error” and aborts the install.sudo or systemd (hypothetical cyborg-ops-helper) cannot run.skill_lift causes the final INSERT to fail, losing all lift data.hermes skill sync command: Scripts that call this command abort.Compute: All skill code runs on the free Oracle VPS agents-01 (ARM64, 2 OCPU, 12 GB RAM). No Docker, no sudo, no systemd.
Skill storage: Verified skills are installed into the profile’s ~/.hermes/skills/ directory via hermes skills install <skill>. This directory is the single source of truth.
Evaluation orchestration: A bash driver eval.sh in the workspace iterates over each role, clones the target repo (Shrikant’s the-pantheon for cycle 1), runs the baseline test suite, installs the role’s survivor skills, re‑runs the tests, and writes lift metrics into the D1 skill_lift table.
Batman’s objection handling: Batman’s earlier objection about the vercel-optimize x86 binary was resolved by removing that skill from the candidate set – a decision reflected in the survivor list.
Kratos’ judgment call: The board comment confirms that evaluation must remain strictly ARM‑only; no separate lightweight VM will be provisioned for x86‑only skills.
anysearch)?skill_lift table, so we can embed the step reliably in eval.sh?