Overview
The harness landscape
An agent harness is everything wrapped around a model to make it useful and reliable: the loop that structures its actions, the tools it can call, the memory and context it sees, the sandbox its code runs in, and the evals that prove any of it works. This page maps the 69 startups and 114 open-source projects in the directory across the layers of that stack.
A note on taxonomy: the layers overlap by design — Letta is a memory company and an orchestration framework; FastMCP is a protocol tool and an integration layer. Most entries live in two or three categories; the primary tag is listed first on each profile.
Protocols & Interop
3 startups · 13 projectsThe shared languages of the harness: how agents discover tools (MCP), talk to each other (A2A, ACP) and drive user interfaces (AG-UI). Protocols are what keep the layer from fragmenting into silos.
Orchestration & Frameworks
22 startups · 49 projectsThe core of the stack: frameworks and SDKs that structure the agent loop — plans, tools, state, handoffs, guardrails and control flow. From graph-based runtimes (LangGraph) to Big Tech SDKs (OpenAI Agents, ADK, Strands) and minimal cores (smolagents, PocketFlow).
Memory & Context
9 startups · 14 projectsEverything past the context window: persistent memory, knowledge graphs and context engineering that let agents remember across sessions and ground decisions in real data.
Data & Retrieval
5 startups · 12 projectsIngestion, crawling, search and retrieval — feeding agents grounded context: RAG frameworks, web crawlers and search APIs purpose-built for LLM consumption.
Tools & Integrations
14 startups · 23 projectsThe hands of the agent: browser automation, tool registries, integration platforms and MCP servers that connect agents to the outside world.
Runtimes & Sandboxes
14 startups · 13 projectsWhere agent code actually executes: secure sandboxes, code-interpreter clouds and serverless agent runtimes that make tool calls safe and fast.
Evals & Observability
13 startups · 15 projectsThe measurement half of harness engineering: tracing, testing, red-teaming and continuous evaluation of agent behavior in dev and production.
Safety & Control
7 startups · 5 projectsGuardrails, permissions and red-teaming — keeping powerful, tool-using agents inside the lines their operators drew.
Infra & Compute
10 startups · 3 projectsThe substrate below the harness: durable execution, queues and workflows that survive crashes, plus gateways that route and proxy model traffic.
Coding Agents
12 startups · 30 projectsThe flagship vertical: agents that write, review and ship software — from autonomous SWE agents to IDE companions and terminal-first harnesses.
Vertical Agents
7 startups · 0 projectsHarnesses pointed at a domain: customer experience, finance, back-office work and the 'AI workforce' platforms shipping agents to end customers.
Community & Events
0 startups · 4 projectsThe meetups, conferences and hackathons where harness engineering know-how spreads — tracked in the Events directory.
Where it all gets discussed
The community layer: 20 upcoming events and 30 tracked news items.