FI

firecrawl-core-workflow-a

Use Firecrawl to scrape single pages or crawl entire sites for content ingestion.

Install

mkdir -p .claude/skills/firecrawl-core-workflow-a && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/8270" && unzip -o skill.zip -d .claude/skills/firecrawl-core-workflow-a && rm skill.zip

Installs to .claude/skills/firecrawl-core-workflow-a

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Execute Firecrawl primary workflow: scrape and crawl websites into LLM-ready
76 charsno explicit “when” trigger
Beginner

Key capabilities

  • →Scrape single web pages into markdown
  • →Crawl entire domains with depth limits
  • →Poll for async crawl job completion
  • →Filter URLs using include and exclude patterns
  • →Process and store crawled content as markdown files

How it works

It utilizes scrapeUrl for single-page extraction and crawlUrl for multi-page discovery, converting rendered web content into structured markdown.

Inputs & outputs

You give it
Target URL and scraping configuration
You get back
Clean, LLM-ready markdown content

When to use firecrawl-core-workflow-a

  • →Scrape website documentation into markdown
  • →Build content ingestion pipelines
  • →Crawl entire domains for data training
  • →Extract structured markdown from web pages

About this skill

Firecrawl Scrape and Crawl Workflow

Overview

Use scrape for one URL and crawl for recursive discovery. Keep the authorization basis, scope, cost ceiling, result completeness, and storage decision explicit.

Prerequisites

  • The target repository or integration path and the requested operator outcome.
  • The source authorization, data classification, and environment policy.
  • Current Firecrawl documentation, credentials only when needed, and an owner for approvals.

Current Contract

The current Node client is Firecrawl from the firecrawl package; Python uses Firecrawl from firecrawl-py. Use scrape for a single URL, crawl as the waiter, startCrawl for asynchronous submission, getCrawlStatus for status and pagination, and cancelCrawl for cancellation. REST uses the /v2 routes and Bearer authentication.

Authentication

For authenticated Cloud operations, inject FIRECRAWL_API_KEY from an approved secret manager. REST requests use Authorization: Bearer with the key. Never print, commit, transmit, or place a key in a URL. Keyless access is suitable only where the current documentation explicitly allows it and the workload accepts its limits; production workflows should make identity and team ownership explicit.

Instructions

  1. Confirm the target is authorized, normalize its origin, and document allowed paths, excluded paths, subdomain/external-link policy, desired formats, and freshness.
  2. Create the client from FIRECRAWL_API_KEY for authenticated work. Keep the key in a secret manager and never place it in source, arguments, logs, or generated examples.
  3. For one page, call scrape with only the required formats and options. Validate the returned document, metadata.sourceURL, metadata.statusCode, and required content fields.
  4. For a site, set an explicit crawl limit and path/depth policy. Use crawl when blocking is acceptable or startCrawl when another worker owns status and cancellation.
  5. Retrieve every required result page. Preserve the next cursor or URL until pagination is complete, and record partial completion separately from terminal success.
  6. Deduplicate by canonical source URL and content hash, validate output quality, and write only approved fields to the downstream store.
  7. Emit counts, creditsUsed when returned, cache state, failures, policy version, and rollback/cancellation outcome without logging page bodies.

Tool Discipline

Use Read, Glob, and Grep to inspect code, configuration, tests, and evidence. Use Write/Edit only for approved implementation or documentation changes. Do not call Firecrawl, rotate keys, change account settings, scrape a target, or deploy merely because this skill was invoked.

Approval Boundaries

Require approval before authenticated-page scraping, external-link traversal, increasing scope or limit, enabling sensitive headers/actions, or retaining raw HTML, screenshots, or personal data.

Output

Return normalized scope, endpoint and SDK method, job ID where applicable, pagination completion, accepted/rejected counts, content-quality checks, retention decision, and a redacted run receipt.

Error Handling

  • Captured origin error page: quarantine by metadata.statusCode and do not index it as successful content.
  • Async job exceeds its deadline: cancel when safe, persist the cursor and receipt, and require an explicit resume decision.
  • Output fails schema or quality checks: retain only permitted evidence and route the page to review.

Examples

  • "Scrape this release note" selects one v2 scrape and validates the returned document.
  • "Crawl our docs" requires path rules, an explicit limit, pagination ownership, and a cancellation plan.

Resources

Read official Firecrawl evidence before relying on an endpoint, SDK method, plan limit, price, retention option, or self-hosted release.

When not to use it

  • →When the target site requires complex user-authenticated sessions
  • →When real-time browser interaction beyond basic rendering is required

Prerequisites

@mendable/firecrawl-js packageFIRECRAWL_API_KEY environment variable

Limitations

  • →Crawl results may be limited by URL filter strictness
  • →Requires increased waitFor duration for sites with heavy JS rendering

How it compares

This workflow automates the conversion of raw HTML into clean markdown while handling JavaScript rendering, which is more reliable than manual scraping scripts.

Compared to similar skills

firecrawl-core-workflow-a side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
firecrawl-core-workflow-a (this skill)02moReviewBeginner
chrome-devtools418moReviewIntermediate
markdown-to-html168moReviewBeginner
nuxt-content37moNo flagsIntermediate

Try saying

Example prompts that trigger this skill in your AI assistant.

More by jeremylongshore

View all by jeremylongshore →

analyzing-logs

jeremylongshore

Analyze application logs to detect performance issues, identify error patterns, and improve stability by extracting key insights.

14123

ollama-setup

jeremylongshore

Configure auto-configure Ollama when user needs local LLM deployment, free AI alternatives, or wants to eliminate hosted API costs. Trigger phrases: "install ollama", "local AI", "free LLM", "self-hosted AI", "replace OpenAI", "no API costs". Use when appropriate context detected. Trigger with relevant phrases based on skill purpose.

1167

backtesting-trading-strategies

jeremylongshore

Backtest crypto and traditional trading strategies against historical data. Calculates performance metrics (Sharpe, Sortino, max drawdown), generates equity curves, and optimizes strategy parameters. Use when user wants to test a trading strategy, validate signals, or compare approaches. Trigger with phrases like "backtest strategy", "test trading strategy", "historical performance", "simulate trades", "optimize parameters", or "validate signals".

1071

generating-database-seed-data

jeremylongshore

Process this skill enables AI assistant to generate realistic test data and database seed scripts for development and testing environments. it uses faker libraries to create realistic data, maintains relational integrity, and allows configurable data volumes. u... Use when working with databases or data models. Trigger with phrases like 'database', 'query', or 'schema'.

1033

cursor-codebase-indexing

jeremylongshore

Execute set up and optimize Cursor codebase indexing. Triggers on "cursor index setup", "codebase indexing", "index codebase", "cursor semantic search". Use when working with cursor codebase indexing functionality. Trigger with phrases like "cursor codebase indexing", "cursor indexing", "cursor".

885

testing-mobile-apps

jeremylongshore

Execute mobile app testing on iOS and Android devices/simulators. Use when performing specialized testing. Trigger with phrases like "test mobile app", "run iOS tests", or "validate Android functionality".

810

You might also like

chrome-devtools

mrgoonie

Browser automation, debugging, and performance analysis using Puppeteer CLI scripts. Use for automating browsers, taking screenshots, analyzing performance, monitoring network traffic, web scraping, form automation, and JavaScript debugging.

41157

markdown-to-html

github

Convert Markdown files to HTML similar to `marked.js`, `pandoc`, `gomarkdown/markdown`, or similar tools; or writing custom script to convert markdown to html and/or working on web template systems like `jekyll/jekyll`, `gohugoio/hugo`, or similar web templating systems that utilize markdown documents, converting them to html. Use when asked to "convert markdown to html", "transform md to html", "render markdown", "generate html from markdown", or when working with .md files and/or web a templating system that converts markdown to HTML output. Supports CLI and Node.js workflows with GFM, CommonMark, and standard Markdown flavors.

1662

nuxt-content

onmax

Use when working with Nuxt Content v3 - provides collections (local/remote/API sources), queryCollection API, MDC rendering, database configuration, NuxtStudio integration, hooks, i18n patterns, and LLMs integration

327

firecrawl-reliability-patterns

jeremylongshore

Implement FireCrawl reliability patterns including circuit breakers, idempotency, and graceful degradation. Use when building fault-tolerant FireCrawl integrations, implementing retry strategies, or adding resilience to production FireCrawl services. Trigger with phrases like "firecrawl reliability", "firecrawl circuit breaker", "firecrawl idempotent", "firecrawl resilience", "firecrawl fallback", "firecrawl bulkhead".

36

fireflies-webhooks-events

jeremylongshore

Implement Fireflies.ai webhook signature validation and event handling. Use when setting up webhook endpoints, implementing signature verification, or handling Fireflies.ai event notifications securely. Trigger with phrases like "fireflies webhook", "fireflies events", "fireflies webhook signature", "handle fireflies events", "fireflies notifications".

08

firecrawl-hello-world

jeremylongshore

Create a minimal working FireCrawl example. Use when starting a new FireCrawl integration, testing your setup, or learning basic FireCrawl API patterns. Trigger with phrases like "firecrawl hello world", "firecrawl example", "firecrawl quick start", "simple firecrawl code".

15

Search skills

Search the agent skills registry