Modal
Serverless cloud built for AI and agentic compute.
Modal lets teams run inference, fine-tuning and arbitrary agent workloads as serverless Python without managing infrastructure — a common substrate for harness backends at scale.
Related startups
Temporal
Durable execution for long-running agent workflows.
Temporal's durable execution engine is widely adopted to make agent pipelines reliable: state survives crashes, retries are automatic, and workflows run for minutes or months. A key piece of serious harness infrastructure.
Inngest
Durable functions and workflows for AI agents.
Inngest runs agent logic as durable steps — queueing, retries, concurrency control and human-in-the-loop waits — without operating your own workflow infrastructure.
BerriAI (LiteLLM)
The gateway between your harness and every model provider.
BerriAI builds LiteLLM — the open-source proxy that normalizes 100+ LLM providers behind the OpenAI API with routing, budgets and observability. A quiet standard inside production agent stacks.