arlen/benchOpen benchmarks for agentic consumers
INDEPENDENT · CC-BY-4.0
UPDATED 14 JUN 2026 · BERKELEY, CA
Vendor Profile · Illustrative prototype

Firecrawl

Firecrawl on the arlen/bench leaderboards — agent web-search accuracy, extraction fidelity, freshness, cost per verified-correct answer, and agent-readiness, scored against golden truth.

Answer firstillustrative

Firecrawl #2 of 4 for extraction fidelity at 0.66, and is not yet agent-ready in live onboarding trials.

extraction fidelity
0.66
web extraction
agent-ready
No
onboarding harness
§ B

Web Extraction — Firecrawl

snapshot web_extraction-2026-q2
Metric FirecrawlLeaderboard bestDirection
fidelity 0-10.66Exa 0.74higher is better
phrase recall0.58Exa 0.67higher is better
boilerplate excl.0.64Exa 0.76higher is better
cost/correct $$0.0294Exa $0.0047lower is better
coverage %97.3Jina 99.3higher is better
§ C

Compare Firecrawl

Firecrawl vs Exa Firecrawl vs SerpAPI Firecrawl vs Brave Firecrawl vs Tavily Firecrawl vs Jina Firecrawl vs Serper

Illustrative prototype. No verified vendor run has been published yet; every figure here is a placeholder and must not be cited as a measured result. Numbers are replaced when a snapshot’s first full run lands.