Web Data and Document Extraction
Use Firecrawl with CrewAI
Firecrawl connects to CrewAI through a built in by the framework vendor path (crewai_tools FirecrawlScrapeWebsiteTool (crewai[tools])). This page records what that connection looks like and links the evidence behind it.
Pairing facts
Why this pairing
A web data API that searches, scrapes, and crawls sites into clean markdown and structured output for AI agents.
Best for: Agents and pipelines that need clean markdown or structured JSON from many pages without running browsers themselves.
Not ideal for: Fully local, no-account workflows. The practical path is the hosted API, and the open-source stack is a larger self-host than a single library install.
Configuration
export FIRECRAWL_API_KEY="<your_firecrawl_api_key>"Source: https://docs.crewai.com/v1.15.22/en/tools/web-scraping/firecrawlscrapewebsitetool. Examples use placeholders only.
Evidence
The CrewAI docs page documents FirecrawlScrapeWebsiteTool, requires FIRECRAWL_API_KEY and installs the Firecrawl SDK with the crewai[tools] package.
https://docs.crewai.com/v1.15.22/en/tools/web-scraping/firecrawlscrapewebsitetool
Compatibility here means the cited page documents the pairing. It does not mean AgentsUse benchmarked the combination or that every version works unchanged. Pin versions, run the tool on a small job first, and review the first outputs before widening access.
Watch for
- The hosted API needs an account and a key, and usage is metered by the vendor. Check current pricing before building a high-volume job on it.
- The open-source license is AGPL-3.0. Review what that means for your deployment before self-hosting or embedding it.
- No independently checked benchmark is published by AgentsUse for this tool yet. Coverage and latency figures on the vendor site are vendor claims.