AI research lab · agent harnesses, memory, measurement

The harness decides what an agent sees, remembers, and costs.

Kleos is a research lab. We do R&D on the layer between a person and a model — the part that chooses what goes into the context window, what survives between sessions, and how much work it takes to get an answer.

We publish what we learn, and we ship the products it produces.

What we work on

01

Harness efficiency

Two agents can finish the same task and spend wildly different amounts getting there. We measure what an agent actually consumes — tokens, calls, latency, local compute — and treat that as a first-class result rather than a footnote.

02

Memory for agents

A decision made in one tool is invisible to the next. We work on memory that persists across sessions and across harnesses, stays inspectable, and knows when old evidence has stopped applying.

03

Work that compounds

The test is not a benchmark task, it is a working day. Software, spreadsheets, research, communication — done repeatedly, so the system gets better at your recurring work instead of starting over.

What we ship

2 products

0xCopilot

desktop agent · open source

A local-first desktop agent. Give it an outcome and it plans the work, moves through your files and connected tools, and stops for your approval before anything important leaves the building.

copilot.kleosresearch.xyz ↗

Kaleidoscope

memory system · native binary

Scattered fragments of work, composed into one bounded view an agent can be handed. Runs on your machine, with adapters for 0xCopilot, Claude Code, Codex and Cursor.

memory.kleosresearch.xyz ↗

Research

5 items
All research →