Skip to content
HarnessRouter

HarnessRouter

Released
0 bookmarks
Visit

Summary

One API for agent harnesses — run Codex, Claude Code or Hermes behind your product with shared sessions, streaming, files and cost controls, and swap harness with a config change.

Screenshots

Description

HarnessRouter sits between your product and the coding-agent harnesses that actually do the work. Instead of integrating Codex, Claude Code and Hermes separately — each with its own session model, streaming format, file handling and failure modes — you call one API and pick the harness with a configuration value.

Why a router

A harness is not a model. It is the loop around the model: the sandbox, the tool calls, the retries, the artefacts it leaves behind. Products that embed agents end up rebuilding that loop per vendor, and then rebuilding it again when a better harness ships. HarnessRouter's answer is the Unified Harness Protocol, a single request/response shape that covers session lifecycle, streaming, file upload, workspace persistence, cancellation and failure.

Supported harnesses

Codex (code, apps, images), Claude Code (code, files, documents) and Hermes, an open-source autonomous agent that runs against any frontier model. A fourth, Pi, is listed as coming soon.

Two ways to run it
  • Community Edition — Apache 2.0, ships as a single Docker container with local SQLite state and a management console. Self-hosted, free, and the reference implementation of the protocol. Source at github.com/HarnessRouter/harnessrouter.
  • HarnessRouter Cloud — managed serverless execution that provisions isolated sandboxes on demand and handles retries, timeouts, permissions and cost caps. A free tier includes 500 launch credits, with Developer, Production and Scale plans above it.
Benchmarking

The platform ships built-in benchmarking so you can compare harness/model pairs on the same task by success rate, latency and cost. The vendor cites a case where the cheapest passing configuration cost 99.8% less than the most expensive one on identical work — the kind of spread that makes a router worth having independently of vendor lock-in.

Who uses it

Y Combinator-backed, with Epsilla, Hibo, Readily, Spira AI and Stanford Medicine named as users on the site.

Reviews

Similar App Suggestions

An Apache-2.0 TypeScript framework for building AI agents — workflows, memory, RAG and evals — with a local studio and an agentic software factory on top.

Coding & DevelopmentFreemium

AI design engineer that generates distinctive UI designs and production code inside your own repo, Figma and design system.

Coding & DevelopmentFreemium

App: oMLX

Jun Kim

New

Menu-bar LLM inference server for Apple Silicon, with continuous batching and tiered KV caching that keeps local models fast enough for real coding work.

Coding & DevelopmentFree

Open-source local inference server that profiles your hardware, picks models that fit, and points your coding agent at them — free, private and offline.

Coding & DevelopmentFree

Open-source AI coding agent for VS Code, JetBrains and the terminal — 500+ models at provider cost, task-specific modes, parallel agents in isolated git worktrees.

Coding & DevelopmentFreemium

Covered in the Weekly