v0.8 · YC F25

See every prompt. Trace every call.

Driftlane is the trace-first observability platform for distributed AI agents. Capture every prompt, replay every failure, and ship to production without bolting on five separate dashboards.

No credit card. SDK in 10 lines. Free forever for solo developers.

trace · run_id=trc_8f3c2a · status=ok 2.41s
Trusted by AI teams at
Cortexa Lattice AI Northwind Helio Labs Parsec
The problem

Why AI engineers can't sleep at night

Three failure modes show up in every production agent stack — and traditional APM tools were never built to catch them.

Pain · 01

Silent failures

Your agent returned a "completed" status — but the answer was wrong, the tool call was skipped, or the user got a polite hallucination. Nothing throws an exception. Your alerting stays green. The complaint email shows up four days later.

Pain · 02

Cost explosions

A single ten-cent agent run quietly turns into a four-dollar loop. Your monthly LLM spend is a black box. By the time finance asks why the bill 4x'd, you cannot point to a single trace, customer, or prompt change as the cause.

Pain · 03

Hallucinations in prod

Your evals passed in CI. Real traffic does things your fixtures never imagined. Without a way to replay actual production traces against new prompts or models, every release is a coin flip and every regression is a fire drill.

The mechanism

Trace-first observability for agents

One SDK, one data model, three primitives. Everything else — costs, evals, alerts — falls out of the trace.

01

Capture

One import wraps every LLM call, tool call, and retrieval. Driftlane records inputs, outputs, latency, and cost — with PII redaction before bytes leave your process.

02

Trace

Spans roll up into a single tree per agent run. See the full reasoning path, branching decisions, retries, and which tool returned the bad payload — searchable across millions of runs.

03

Replay

Pick any production trace and re-run it against a new prompt, a new model, or a patched tool. Diff outputs side by side, score with evals, and promote the winner — no synthetic dataset needed.

Before / After

Stop stitching dashboards. Start shipping.

What a normal day looks like before Driftlane — and what it looks like once your agents are instrumented.

Before Driftlane

Logs scattered across five tools

  • Prompts buried in CloudWatch, costs in OpenAI billing, traces in a Notion doc.
  • Customer reports a bad answer — you cannot find the actual run that produced it.
  • Prompt changes ship with no way to A/B against historical traffic.
  • On-call engineers stare at JSON dumps at 2 a.m. with no span hierarchy.
  • Token bills surprise you on the 1st of every month.
With Driftlane

One unified trace per run

  • Every prompt, tool call, and token in one searchable timeline.
  • Filter by user, tenant, or release tag — find the exact failing trace in seconds.
  • Replay last week's traffic against your new prompt before merging.
  • Cost dashboards by agent, tenant, model, and feature flag.
  • Slack alerts when latency, error rate, or spend cross your guardrails.
Features

Everything you need to run agents in production

Six primitives that replace the half-dozen tools you would otherwise stitch together.

Structured traces

OpenTelemetry-compatible spans roll up into a single tree per run. Search, filter, and pivot across millions of agent executions in under a second.

Cost attribution

Slice token spend by agent, customer, feature flag, or release tag. Get alerted before a runaway loop burns through your monthly budget.

Trace replay

Re-run any production trace against a new prompt, model, or tool. Diff the outputs, score with evals, and promote the winner without a synthetic dataset.

Built-in evals

Helpfulness, faithfulness, toxicity, JSON-validity — ship with a battery of LLM-as-judge scorers, or plug in your own Python function as a custom eval.

PII redaction

Mask emails, names, and custom regex patterns inside the SDK before payloads leave your process. Self-host the control plane for VPC-isolated workloads.

Live alerts

Slack, PagerDuty, and webhook alerts when latency p95, error rate, or spend per tenant crosses your thresholds. Tied directly to the offending trace.

Pricing

Honest pricing. No seat tax.

Pay for the spans you ingest. Invite the whole team for free. Cancel any time.

Solo
$0 / forever

For solo developers shipping their first agent. No card required.

Get started
  • 10,000 spans / month
  • 3-day trace retention
  • All framework adapters
  • Community Discord support
Scale
$499 / month

For teams running agents at production volume across multiple environments.

Book a demo
  • 2,000,000 spans / month
  • 30-day trace retention
  • Self-hosted control plane
  • SSO, audit logs, RBAC
  • Dedicated support channel
Customers

What teams ship with Driftlane

Builders running real agents in production — not demos.

Reduced debug time 70%

"We used to chase agent bugs across CloudWatch, Sentry, and three Postgres tables. Now it is one trace. My on-call engineers actually sleep again."

MR
Mira Reyes
Staff Engineer, Brightline Ops
Caught 3 prod bugs week 1

"Within seven days of instrumenting our support agent, Driftlane surfaced two tool regressions and a silent rate-limit retry storm. The replay feature paid for itself before the trial ended."

JC
Jonas Carver
Founding Engineer, Quill AI
42% lower token spend

"Cost attribution by tenant let us spot two enterprise accounts looping on a bad retrieval prompt. One config change cut our OpenAI bill almost in half."

AT
Aditi Thakur
CTO, Northwind Robotics
Ship 4x more often

"Replaying yesterday's traffic against a new prompt before merging completely changed our release cadence. We went from one prompt change a week to four a day, with fewer regressions."

SL
Sam Linde
Tech Lead, Helio Labs
Onboarded in 8 minutes

"SDK in, env var set, first trace on the dashboard before my coffee finished brewing. The dev experience is the closest thing to Stripe I have seen in AI tooling."

PK
Priya Kapoor
Solo founder, Loopdesk
99.99% trace delivery

"We process millions of agent runs a month. The Driftlane SDK never blocks our main loop, and the disk-buffered replay saved us during a noisy AWS week. Zero dropped traces."

DV
Diego Vela
Platform Lead, Parsec Systems
Who builds Driftlane

Built by engineers from Stripe and Anthropic

Hi, I'm Eitan Brook — co-founder & CTO.

I spent four years at Stripe building the observability stack behind Radar, then two years at Anthropic on the team that wrote the first internal tooling for agent evals. Driftlane is the product I kept wanting at both companies and could never find on the market.

We are a small, distributed engineering team — most of us have shipped production AI systems before starting Driftlane. We backed by Y Combinator (F25), General Catalyst, and a handful of operators from OpenAI, Vercel, and Datadog.

YC F25 SOC 2 Type II HIPAA-ready GDPR
No risk

Free 14-day trial. No credit card. Cancel any time.

If Driftlane does not earn its place in your stack by day 14, your account quietly downgrades to the free Solo tier. No clawback emails, no friction.

No credit card SOC 2 Type II Self-host on Scale Cancel any time
FAQ

Questions engineers actually ask

Do I have to rewrite my agent code to use Driftlane?

No. The Driftlane SDK is a single import that wraps your existing LLM and tool calls. Most teams instrument their first agent in under ten minutes — typically two lines in your entrypoint plus an environment variable for your project key. You keep your framework of choice, whether that is LangGraph, CrewAI, Vercel AI SDK, or hand-rolled.

Which LLM providers and frameworks does Driftlane support?

Out of the box we trace OpenAI, Anthropic, Google, Mistral, Groq, Together, and any provider that follows the OpenAI-compatible chat schema. On the framework side we have first-class adapters for LangChain, LangGraph, LlamaIndex, CrewAI, Vercel AI SDK, Pydantic AI, and the Anthropic Agent SDK. Anything else can be instrumented with our generic span API.

How does Driftlane handle sensitive data inside prompts?

Redaction happens in the SDK before the payload ever leaves your machine. You can configure regex masks, key allowlists, or run our local PII model. For regulated workloads we offer a self-hosted control plane that keeps trace bodies inside your VPC while metadata streams to our cloud for the dashboard view.

Will Driftlane add latency to my production agents?

Tracing is fully async and batched. The SDK adds under one millisecond per span on a warm process and never blocks your main loop — even on cold start the overhead stays below five milliseconds. If our ingest is unreachable, spans buffer to disk and replay automatically once connectivity returns.

Can I run evals and replay traces against new model versions?

Yes. Every captured trace is replayable. Pick any subset — failed traces, expensive traces, traces from a specific customer — and re-run them against a new prompt, a new model, or an updated tool. Diff the outputs side by side, score with built-in or custom evals, and promote the winning configuration.

How does pricing scale once I move past the free Solo tier?

Team is a flat forty-nine dollars per month and includes one hundred thousand spans plus seven-day retention. Scale is four hundred ninety-nine per month for two million spans and thirty-day retention. Overage is metered per ten thousand spans and never surprises you — we pause ingest at your configured cap and email you first.

Is there a self-hosted or on-prem option?

Scale tier includes a self-hosted control plane image you can deploy into your own Kubernetes cluster. It ships as a single Helm chart with Postgres and ClickHouse backends. Enterprise customers get a hardened build with SSO, audit logs, and BYOK encryption — happy to walk you through it on a call.

Who is Driftlane actually for?

Teams building production agents — not chatbots. If your system makes multiple LLM calls, hits external tools, branches on results, and needs to be debuggable at 2 a.m., Driftlane is built for you. Today our customers run customer-support agents, coding agents, research agents, and internal ops automations at companies from two-person YC startups to publicly traded fintechs.

Ship with confidence

Trace your first agent in ten minutes.

Free Solo tier, no card, no sales call. Upgrade when your traffic does.