---
source_url: "https://docs.crewai.com/v1.14.7/en/tools/web-scraping/firecrawlscrapewebsitetool"
title: Firecrawl Scrape Website - CrewAI
mirrored_at: 2026-08-13T03:40:43.438Z
host: docs.crewai.com
cited_in_42a: true
mirror_canonical: "https://index.42a.ai/docs.crewai.com/v1.14.7/en/tools/web-scraping/firecrawlscrapewebsitetool"
---

> **Original source:** https://docs.crewai.com/v1.14.7/en/tools/web-scraping/firecrawlscrapewebsitetool

## `FirecrawlScrapeWebsiteTool`

## Description

[Firecrawl](https://firecrawl.dev/) is a platform for crawling and convert any website into clean markdown or structured data.

## Installation

-   Get an API key from [firecrawl.dev](https://firecrawl.dev/) and set it in environment variables (`FIRECRAWL_API_KEY`).
-   Install the [Firecrawl SDK](https://github.com/mendableai/firecrawl) along with `crewai[tools]` package:

## Example

Utilize the FirecrawlScrapeWebsiteTool as follows to allow your agent to load websites:

Code

## Arguments

-   `api_key`: Optional. Specifies Firecrawl API key. Defaults is the `FIRECRAWL_API_KEY` environment variable.
-   `url`: The URL to scrape.
-   `page_options`: Optional.
    -   `onlyMainContent`: Optional. Only return the main content of the page excluding headers, navs, footers, etc.
    -   `includeHtml`: Optional. Include the raw HTML content of the page. Will output a html key in the response.
-   `extractor_options`: Optional. Options for LLM-based extraction of structured information from the page content
    -   `mode`: The extraction mode to use, currently supports ‘llm-extraction’
    -   `extractionPrompt`: Optional. A prompt describing what information to extract from the page
    -   `extractionSchema`: Optional. The schema for the data to be extracted
-   `timeout`: Optional. Timeout in milliseconds for the request

Was this page helpful?