Startup profile for Braintrust: latest funding, latest known valuation, total disclosed funding, investors, status, sources, and why the company is interesting. Part of the Venture Capital Tracker startup directory.

Startup profile · funding coverage

Braintrust funding, valuation and investors

AI evaluation, observability, and prompt tooling for production LLM products.

Keep track of Braintrust

Save this profile to your VCT watchlist for a quick return.

View watchlist

Funding, valuation & investors

Answer-first snapshot

Latest funding

Series A

$36M · October 2024

Latest known valuation

Not publicly disclosed

Total disclosed equity funding

$116M

Excludes debt, grants, acquisitions, secondaries, and IPO proceeds.

Current status

private

series b · San Francisco, California

Investors in latest funding

Lead: Andreessen Horowitz

Sources: latest funding Last verified: 2026-07-25

Overview

Braintrust provides an AI evaluation and observability platform — helping product teams test prompts, score model outputs, and monitor LLM applications in production.

Why Braintrust is interesting

Used by Stripe, Notion, and Vercel for eval loops — a16z-led A with Greylock participation as evals become table stakes for shipping AI.

Product & use cases

LLM eval infrastructure — datasets, automated scoring, prompt versioning, and production observability for AI product teams.

  • Automated eval scoring before LLM feature releases
  • Prompt regression testing in CI/CD pipelines
  • Production monitoring of model output quality

Key facts

  • Series A (Oct 2024): $36M led by a16z for AI evals and LLM devtools (Braintrust blog)
  • Series B: $80M with Greylock re-upping alongside a16z (company blog)
  • Co-investors not in VCT directory: ICONIQ (Series B lead), Databricks Ventures, Datadog
  • Source: https://www.braintrust.dev/blog/announcing-series-a

Funding history (newest first)

Investors in our directory

Funds linked from Braintrust's profile — open a fund page for stage focus and related deal articles.

Competitive landscape

Edge: Used by Stripe, Notion, and Vercel for eval loops — a16z-led A with Greylock participation as evals become table stakes for shipping AI.

Eval tooling is table stakes for production AI. Braintrust wins on developer UX and references (Stripe, Notion, Vercel).

  • Weights & Biases direct

    ML experiment tracking expanding into LLM evals.

  • LangSmith direct

    LangChain's LLM observability platform.

  • Humanloop direct

    Prompt management and eval tooling.

Notable stories

  • Reference customers include Stripe, Notion, and Vercel.
  • ICONIQ led Series B; Databricks Ventures and Datadog among strategics.

Industries

AI & Machine Learning Developer Tools Enterprise SaaS

Related funding articles

Venture Capital Tracker pieces that cover Braintrust's financing or category context.

FAQs about Braintrust

Practical answers founders, operators, and investors typically search for.

Braintrust is dev infrastructure for shipping LLM features: eval datasets, automated scoring, prompt versioning, and production observability used by teams at Stripe, Notion, and Vercel.
Andreessen Horowitz (see /fund/andreessen-horowitz) led the $36M Series A in October 2024. Greylock Partners (see /fund/greylock-partners) participated in the $80M Series B alongside a16z.
Series B ($80M). ICONIQ led the B round per company disclosures; Braintrust remains private.
Evals and observability became table stakes for production AI — Braintrust sits in the same tooling layer as CI/CD but for non-deterministic model outputs.
No — Braintrust remains private.
Headquarters not publicly disclosed in sources we index.
Founding year not publicly disclosed in sources we index.
Braintrust has not publicly disclosed profitability.
See fundraising rounds and highlights on this profile for disclosed amounts.
ai-ml, developer-tools, enterprise-saas

By Venture Capital Tracker

Last updated:

Editorial note: AI tools assisted with research, structure, or drafting. Venture Capital Tracker retains human editorial responsibility for factual accuracy, relevance, and source quality before publication.