Skip to main content

Clarity Harness · Open Source

Compound your context.

Every agent run makes the next one better. Incubate your agents like new hires — check their work, give verdicts, bank every correction — until they earn autopilot. Then share your agents and harnesses across the team. Your subjective taste is the advantage. The harness compounds it, right in the flow of work.

1,437 builders already in line · free & open source · no license keys

The Clarity Harness — mission control for your agents.

The pain

Agents made you fast.
They didn't make you sure.

01

You bought agents. You got drafts.

Claude Code, Cursor, your own stack — throughput 10x’d overnight. But every artifact is a draft until someone you trust says it isn’t. You didn’t remove the bottleneck. You moved it.

02

You’re the eval now.

There are evals for models. There are none for your agents’ actual work-product — the stories, specs, and PRs they ship into your backlog. So your senior people read everything. Or worse: they stop reading.

03

Your corrections don’t compound.

You catch the same mistake every week. The fix dies in a Slack thread. Tomorrow’s run starts with zero memory of being wrong. All that judgment, spent — none of it banked.

Speed without compounding is just expensive slop.

The promise

Every verdict compounds.

The harness closes the loop: capture the run, deploy the agent, review with a verdict, learn from the correction. Today's annotation is tomorrow's default. Week one feels like review. Week six feels like leverage.

Clarity Loop

Capture

Every run recorded, receipts attached

Deploy

The agent runs your workflow

Review

You give the verdict in minutes

Learn

Corrections banked into the next run

Continuously improving

Proof · every run shows receipts

Every artifact shows its chain.

There are agent runs, and one question that matters: did the input produce the expected output — and how does it get better next time? The harness makes that question answerable.

01

Quoted from here

Every requirement links back to the exact transcript line it came from. No mystery provenance.

02

The chain that got you here

See which agents worked together and in what order — story writer with requirements synthesizer, on this run, from this input.

03

The diff from before

What did this version add? What changed from the last one? The full artifact, one click away.

04

Your verdict

Approve it — or annotate it: “this requirement was too narrow, specify broader.” Save. Regenerate. The harness learns your judgment.

Proof · we run on it

This is how we're evolving our work.

If you're a product manager, you're going to have a story writer at the top — and a stack of agents underneath it doing the work you used to do by hand. The screenshots on this page are our production instance, not a demo: 8 agents, 129 runs, 155 verdicts in the queue — the harness running the company that built it.

◆ agent

Story Writer

at the top

Turns captured requirements into stories, drafted and waiting for your verdict — not your keystrokes.

◆ agent

Clarity Orchestrator

routing the work

Orchestrator calls the agent, agent calls the skill. Point your coding agent at it and the whole stack composes.

◆ agent

Interview Agent

capturing input

Sits on your meetings and conversations, turning what people actually said into structured evidence.

◆ agent

Requirements Synthesizer

from transcript to spec

Works with the story writer to pull requirements straight out of transcripts — quoted, sourced, traceable.

One philosophy, every scale

Individual. Domain. Company-wide.

Individual

Your own harness

Every individual has their own harness — your agents, your runs, your judgment, your data. In the flow of work: writing stories, making prototypes, agents doing stuff.

Domain

Domain-specific harnesses

Privacy, payroll, compliance — wherever tacit knowledge lives, a domain harness turns the expert’s judgment into something every run can use.

Company-wide

The company brain

Department and company-wide harnesses share one brain across your people and agents. That’s the layer we build with design partners today.

See the company brain →

What ships day one

Everything we run internally. All of it.

The full harness UI

Mission control, run views, judgment queue — the exact interface in the screenshots above.

Run capture with full provenance

Every run recorded: source input, agent chain, tool calls, decisions taken and not taken.

The verdict system

Approve, accept with edit, reject, defer — with annotations that persist as structured data, not Slack threads.

Versioned datasets

Your corrections become a dataset with lineage — dev, staging, prod — that regenerates your agents better.

MCP serve

Point Claude Code or any coding agent at your harness — orchestrator calls the agent, agent calls the skill.

Our internal setup guide

The same playbook we used to stand up the harness we run Clarity on. First harness running the same afternoon.

Built in the open

A community for builders who believe agents deserve better than vibes.

The harness ships free and open source because the thesis is bigger than a tool. This is the Clarity / Epistemic Me bet — and we're building a community of AI builders who share it:

Subjectivity is the missing primitive

Autonomy doesn’t come from bigger models. It comes from knowing whose goals, whose judgment, and whose context a run answers to.

Alignment is layered

Individual, team, and organization goals aren’t the same thing — a real harness reconciles all three instead of pretending one prompt covers it.

Closed loops beat bigger prompts

True autonomous agents are grown, not prompted: run → verdict → learn → run again, with the human judgment banked every cycle.

If that's your thesis too, don't just star the repo — come build the loop with us.

Open source · early access

Start compounding before everyone else.

Early access gets the repo first, plus the setup guide we use internally — first harness running the same afternoon. Open source, no license keys.

1,437 builders already in line