Firecrawl
for LuumenAI

Crawl sites, extract structured data, and watch job status from the terminal

Connect Firecrawl and Luumen can read your crawl and extract jobs directly: what is running, what finished, which URLs failed or were blocked by robots, and how many credits the team has burned this month. When you need new data, Luumen can start a crawl, batch scrape a list of URLs, or run a structured extraction as an approved step. Authentication is by API key.

The Firecrawl toolbox

28 tools: 19 read, 9 write. Reads answer instantly. Writes require approval by default. Everything is logged.

  • ReadGet batch scrape statusRetrieves the current status and results of a batch scrape job using the job ID.
  • ReadGet errors from batch scrape jobRetrieve error details from a batch scrape job, including failed URLs and URLs blocked by robots.txt.
  • ReadGet crawl job statusRetrieve the status and results of a Firecrawl crawl job.
  • ReadGet errors from a crawl jobRetrieve errors from a Firecrawl crawl job.
  • ReadGet all active crawl jobsRetrieve all active crawl jobs for the authenticated team.
  • ReadPreview crawl parametersPreview crawl parameters before starting a crawl by generating optimal configuration from natural language instructions.
  • ReadGet team credit usageGet current team credit usage information.
  • ReadGet historical team credit usageRetrieve historical team credit usage on a monthly basis.
  • ReadGet extract job statusRetrieve the status and results of a previously submitted extract job.
  • ReadGet agent job statusGet the status and results of an agent job.
  • ReadGet deep research statusRetrieves the status and results of a deep research job by its ID.
  • ReadGet the status of a crawl jobRetrieves the current status, progress, and details of a web crawl job, using the job ID obtained when the crawl was initiated.
  • ReadGet LLMs.txt generation job statusGet the status and results of an LLMs.txt generation job.
  • ReadMap multiple URLsMaps a website by discovering URLs from a starting base URL, with options to customize the crawl via search query, subdomain inclusion, sitemap handling, and result limits; search effectiveness is…
  • ReadGet team queue statusRetrieve metrics about the team's scrape queue.
  • ReadScrape URLScrapes a publicly accessible URL, optionally performing pre-scrape browser actions or extracting structured JSON using an LLM, to retrieve content in specified formats.
  • ReadSearchPerforms a web search for a query, scrapes content from the top search results using Firecrawl, and returns details in specified formats.
  • ReadGet team token usageRetrieve the current team's token usage and balance information for Firecrawl's Extract feature.
  • ReadGet historical team token usageRetrieve historical team token usage on a monthly basis.
  • WriteCancel an agent jobCancel an in-progress agent job by its ID. Approval by default
  • WriteBatch scrape multiple URLsScrape multiple URLs in batch with concurrent processing. Approval by default
  • WriteCancel a batch scrape jobCancel a running batch scrape job using its unique identifier. Approval by default
  • WriteStart a web crawlInitiates a Firecrawl web crawl from a given URL, applying various filtering and content extraction rules, and polls until the job is complete; ensure the URL is accessible and any regex patterns… Approval by default
  • WriteCancel a crawl jobCancels an active or queued web crawl job using its ID; attempting to cancel completed, failed, or previously canceled jobs will not change their state. Approval by default
  • WriteStart a web crawl (v2) [NEW][NEW v2 API] Initiates a Firecrawl v2 web crawl with enhanced features over v1: natural language prompts for automatic crawler configuration, crawlEntireDomain for sibling/parent page discovery… Approval by default
  • WriteExtract structured dataExtracts structured data from web pages by initiating an extraction job and polling for completion; requires a natural language `prompt` or a JSON `schema` (one must be provided). Approval by default
  • WriteGenerate LLMs.txt for a websiteInitiates an async job to generate an LLMs.txt file for a website, converting web content into LLM-friendly format. Approval by default
  • WriteStart an agent jobStart an agent job for agentic web extraction with multi-page navigation and interaction capabilities. Approval by default

One prompt, start to finish

What a governed Firecrawl run looks like inside Luumen.

Questions

How does LuumenAI connect to Firecrawl?

Authorize once with API token. Luumen lists the scopes each action needs before you approve the connection, and credentials never appear in the chat.

Can LuumenAI change things in Firecrawl on its own?

Read actions answer immediately. Anything that writes — cancel an agent job, batch scrape multiple urls, cancel a batch scrape job, start a web crawl, and more — is shown as a plan and requires approval by default, including the 3 actions classified as destructive. Administrators configure that per tool, so you decide exactly which actions can ever run unattended.

Who gets access to the integration?

You decide. Actions are granted per agent, skill, and team, and per environment — production is not staging. Read access can be broad while writes stay narrow.

Is there an audit trail?

Every call to Firecrawl — read or write, approved or declined — is recorded with the actor, the input, and the result, and can be linked to the ticket or change record.

Put Firecrawl to work with Luumen

Connect in minutes. Every action scoped, approved, and audited from day one.