Head to head
aeqi vs Hermes Agent
The facts side by side — model held constant. No performance verdict until traces are published under the pre-registered method.
aeqi builds one of the harnesses measured here. This board is in method preview and pre-registration — no performance scores are published yet, and we lead with our own losses, not our wins. Read the method & charter →
| Attribute | aeqi | Hermes Agent |
|---|---|---|
| Maintainer | aeqi | NousResearch |
| License | Proprietary | MIT |
| Category | Coding + general | Coding + general |
| Headless run | Yes — `aeqi run` (single-shot, git-stamped version) | Yes — `--oneshot` + `--usage-file` (token/cost report for free) |
| Native OpenRouter | ||
| Determinism | Pinnable via git commit; temperature override is a known gap (defaults high) — disclosed. | Pin via git tag/commit. Ships memory + rules auto-injection — run `--safe-mode` for apples-to-apples. |
| Benchmark result | Not yet measured | Not yet measured |
Questions
- Is aeqi better than Hermes Agent?
- The aeqi Harness Index does not publish a performance verdict for aeqi vs Hermes Agent yet — scores land only with published traces under the pre-registered method. What this page compares are the verifiable facts: license, category, headless capability, and whether each can be pinned to one shared model for a fair run.
- Can aeqi and Hermes Agent be benchmarked fairly against each other?
- Yes — both document native OpenRouter support, so we can point them at one identical pinned model and attribute any difference to the scaffold, not the model.