Skip to content
Tech CEO Daily
StartupsFunding

Firecrawl raises $75 million to build a paid knowledge library for AI agents

The web-data startup’s Series B, led by Smash Capital, funds Alexandria, a platform that plans to pay content owners when AI agents use their data.

TC

By Tech CEO Daily Staff, Newsroom

· 2 min read

A vast modern library with tall bookshelves and glowing digital screens between the stacks
AI-generated image for illustration. Not a photograph of the events described.

The news

Firecrawl, a startup whose tools extract and structure data from websites for AI applications, said on September 22 that it raised a $75 million Series B. Smash Capital led the round, with Altos Ventures, Nexus Venture Partners, Y Combinator, Freestyle and Offline Ventures participating. No valuation was disclosed.

The funding comes with a new product called Alexandria. Firecrawl describes it as a single interface that combines official data providers, custom connectors, its own indexes and live web access, so AI agents can find sources and retrieve information. The company said it will spend much of the money buying knowledge from individuals and organizations, improving search, adding data sources and building a self-service system to pay content creators and data providers.

Firecrawl said it has 1.5 million users. Dealroom reported it serves more than 150,000 companies, naming Shopify, Apple, Lovable and Canva, and that the startup was founded in 2024 by Caleb Peffer, Eric Ciarla and Nicolas Silberstein Camara. The round comes about a year after its Series A, which Dealroom said was led by Nexus and Y Combinator.

The company also cited an internal test in which agents using Alexandria scored 21% higher on answer quality than those using built-in web tools across 845 tasks. That is a company-run benchmark and has not been independently verified.

The numbers

Series B
$75M
Lead investor
Smash Capital
Users (company-stated)
1.5M

Why CEOs should care

AI agents are only as good as the information they can reach, and the open web is getting harder and legally riskier to scrape as publishers block crawlers and sue. Firecrawl’s pivot toward paying for data mirrors a broader shift: access to reliable, licensed information is becoming a cost line in AI projects. Companies building agents should budget for it and check the provenance of the data their vendors supply.

For publishers and companies sitting on specialized data, a new set of buyers is emerging. Marketplaces like the one Firecrawl describes could turn documentation, research and archives into licensing revenue, though terms, pricing and enforcement are still unproven.

The bigger picture

Capital is concentrating in the “picks and shovels” layer for agents: data access, orchestration and evaluation. Weekly funding roundups show similar bets on AI training data and agent infrastructure.

Sources

TC
Tech CEO Daily Staff

Newsroom

Reporting and analysis from the Tech CEO Daily newsroom. Each story is researched from primary sources — company announcements, regulatory filings and official advisories — and fact-checked before publication.

Spotted an error? Request a correction. Read our editorial standards and AI policy.

The Daily Brief

The technology briefing for people running businesses.

Weekdays at 6 a.m. ET. Free.

More in Startups