Startup profile for Inferact: latest funding, latest known valuation, total disclosed funding, investors, status, sources, and why the company is interesting. Part of the Venture Capital Tracker startup directory.

Startup profile · funding coverage

Inferact funding, valuation and investors

Company behind vLLM — open-source LLM inference engine used by Meta, Google, and Character.ai in production.

Keep track of Inferact

Save this profile to your VCT watchlist for a quick return.

View watchlist

Funding, valuation & investors

Answer-first snapshot

Latest funding

Seed

$150M · January 2026

Latest known valuation

$800M

Seed · January 2026

Total disclosed equity funding

$150M

Excludes debt, grants, acquisitions, secondaries, and IPO proceeds.

Current status

private

seed · San Francisco, California

Sources: latest funding Last verified: 2026-07-25

Overview

Inferact is the company behind vLLM, the open-source LLM inference engine used in production by Meta, Google, and Character.ai. It raised one of the largest seed rounds of 2026 to commercialize inference infrastructure.

Why Inferact is interesting

One of largest seed rounds of 2026; explicit mission to fund vLLM OSS while building inference products.

Product & use cases

Inferact is the company behind vLLM — the dominant open-source LLM inference engine — building commercial inference products while funding OSS development.

  • Production LLM inference at scale
  • Open-source vLLM maintenance and optimization
  • Universal inference layer across hardware and model architectures

Key facts

  • Seed (2026): $150M at $800M valuation co-led by a16z and Lightspeed for vLLM commercialization — TechCrunch
  • vLLM is the dominant open-source LLM inference engine used by Meta, Google, and Character.ai in production

Funding history (newest first)

Investors in our directory

Funds linked from Inferact's profile — open a fund page for stage focus and related deal articles.

Competitive landscape

Edge: vLLM runs on 400K+ GPUs concurrently with 2,000+ contributors — Meta, Google, Character.ai use in production per a16z.

Inference cost is the bottleneck for production LLM apps. Inferact's explicit mission is funding vLLM OSS — commercial layer must complement, not compete with, existing providers using vLLM.

  • TensorRT-LLM (NVIDIA) direct

    Hardware-vendor inference stack; vLLM hardware-agnostic OSS.

  • TGI (Hugging Face) direct

    Open-source inference server; vLLM larger production footprint.

  • SGLang/RadixArk direct

    Berkeley-lab sibling project also commercializing (2026).

Notable stories

  • One of largest seed rounds in history — $150M at $800M valuation (TechCrunch, Jan 2026).
  • Founded by vLLM maintainers including Simon Mo and Kaichao You — incubated at UC Berkeley Ion Stoica lab.

Industries

AI & Machine Learning Infrastructure & Cloud Developer Tools

Related funding articles

Venture Capital Tracker pieces that cover Inferact's financing or category context.

FAQs about Inferact

Practical answers founders, operators, and investors typically search for.

Company behind vLLM open-source inference engine — funds OSS development and builds commercial inference products.
a16z (/fund/andreessen-horowitz) and Lightspeed (/fund/lightspeed-venture-partners-nyc) co-led $150M seed (Jan 2026). Sequoia (/fund/sequoia), Altimeter (/fund/altimeter-capital), Redpoint (/fund/redpoint-ventures) participated.
Dominant open-source LLM inference engine — 400K+ GPUs, 2,000+ contributors, used by Meta and Google in production.
vLLM is hardware-agnostic OSS; TensorRT tied to NVIDIA stack.
a16z says Inferact will work with existing providers — most already use vLLM under the hood.
Seed — $150M at $800M valuation (Jan 2026), one of largest seeds on record.
vLLM creators including CEO Simon Mo and co-founder Kaichao You.
No — launched January 2026, remains private.

By Venture Capital Tracker

Last updated:

Editorial note: AI tools assisted with research, structure, or drafting. Venture Capital Tracker retains human editorial responsibility for factual accuracy, relevance, and source quality before publication.