Langfuse (ClickHouse)Acquired
Open-source LLM engineering platform: tracing, prompts, evals.
Langfuse is an open-core platform for observing and improving LLM apps — trace-level debugging, prompt management, datasets and evals. Self-hostable, with a cloud offering. Acquired by ClickHouse in January 2026 as part of its push into AI observability; the open-source core remains actively maintained.
Open-source project
In the news
DeepSeek ships Harness Desktop; the 'everything is a plugin' harness passes 240k stars
DeepSeek releases a macOS and Windows Desktop app for DeepSeek Harness, its open-source agent harness where everything — tools, sub-agents, even the loop — is a plugin. The repo passed 240k stars within three months of its August launch, making it the fastest-growing harness in the field's history.
xAI open-sources Grok Build, its coding agent harness
xAI open-sources Grok Build — a Rust coding agent harness and fullscreen TUI around Grok models — unusually early for a first-party tool. It joins Claude Code, Codex and Gemini CLI as the fourth major first-party harness, now with a fourth open one.
ClickHouse acquires Langfuse
ClickHouse acquires the open-source LLM observability platform Langfuse as part of its $400M Series D at a $15B valuation — the first big consolidation in agent observability. The open-source core remains maintained; the deal signals that agent tracing is becoming a database-scale problem.
Anthropic launches the Claude Agent SDK
Anthropic open-sources the SDK behind Claude Code — the agent loop, tool use, subagents, memory and hooks — letting developers build their own Claude-powered agents on production rails.
Related startups
LangChain
The de-facto standard toolkit for building LLM applications and agents.
LangChain makes the LangChain framework and LangGraph, the low-level orchestration standard for stateful, controllable agents, plus the LangSmith platform for tracing, evaluating and monitoring them in production.
Braintrust
Enterprise-grade evals, data and AI gateway.
Braintrust turns prompts and agent traces into testable datasets and CI, plus an inference gateway — helping teams measure whether harness changes actually improve agent quality.
Arize AI
Observability and evaluation for AI — makers of OSS Phoenix.
Arize traces and evaluates LLM and agent systems in production; their open-source Phoenix tracer is a common instrumentation choice for agent harnesses that need to be debuggable.