firecrawl-known-pitfalls
A checklist of common Firecrawl mistakes to avoid in production code to prevent credit waste and data issues.
Install
mkdir -p .claude/skills/firecrawl-known-pitfalls && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/8057" && unzip -o skill.zip -d .claude/skills/firecrawl-known-pitfalls && rm skill.zipInstalls to .claude/skills/firecrawl-known-pitfalls
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Identify and avoid Firecrawl anti-patterns and common integration mistakes.Key capabilities
- →Identify unbounded crawl risks
- →Enforce output format specifications
- →Implement wait times for JavaScript-heavy pages
- →Optimize polling with exponential backoff
- →Validate extracted content with Zod
How it works
The skill provides code patterns to replace inefficient scraping methods with optimized, error-checked, and resource-aware implementations.
Inputs & outputs
When to use firecrawl-known-pitfalls
- →Reviewing Firecrawl implementation code
- →Preventing accidental credit depletion
- →Setting up scraping best practices
- →Onboarding new team members
About this skill
Firecrawl Integration Pitfall Review
Overview
Find high-impact failure modes before production. Tie every finding to current Firecrawl behavior and repository evidence instead of applying a generic checklist.
Prerequisites
- The target repository or integration path and the requested operator outcome.
- The source authorization, data classification, and environment policy.
- Current Firecrawl documentation, credentials only when needed, and an owner for approvals.
Current Contract
Common current hazards include legacy scrapeUrl/crawlUrl method names, implicit crawl scope, confusing API success with origin status, dropping pagination, retrying 4xx errors, assuming one credit per request, trusting scraped instructions, enabling cache on sensitive data, skipping webhook HMAC verification, and exposing the self-host quickstart.
Authentication
For authenticated Cloud operations, inject FIRECRAWL_API_KEY from an approved secret manager. REST requests use Authorization: Bearer with the key. Never print, commit, transmit, or place a key in a URL. Keyless access is suitable only where the current documentation explicitly allows it and the workload accepts its limits; production workflows should make identity and team ownership explicit.
Instructions
- Locate Firecrawl dependencies, imports, wrapper clients, REST paths, config, queues, webhooks, stores, tests, and deployment files.
- Flag legacy v0/v1 routes or FirecrawlApp-style methods unless they are intentionally isolated behind the documented feature-frozen compatibility surface.
- Find every crawl, batch, search, agent, and browser call. Require explicit scope, credit/time limits, cancellation ownership, and environment policy.
- Verify SDK versus REST response handling, metadata.statusCode checks, complete pagination, terminal-state handling, and idempotent webhook processing.
- Compare retry logic with the official error catalog. Reject retries for auth, credits, restrictions, validation, and other non-retryable failures.
- Trace scraped or extracted content into logs, prompts, stores, and actions. Require sanitization, provenance, prompt-injection boundaries, retention, and deletion.
- Rank findings by exploitability and impact, propose minimal fixes and regression tests, and verify the corrected paths.
Tool Discipline
Use Read, Glob, and Grep to inspect code, configuration, tests, and evidence. Use Write/Edit only for approved implementation or documentation changes. Do not call Firecrawl, rotate keys, change account settings, scrape a target, or deploy merely because this skill was invoked.
Approval Boundaries
Require approval before changing dependencies, expanding crawl scope, enabling raw formats/actions, weakening cache or retention safeguards, or auto-fixing production code.
Output
Return evidence-linked findings, severity, affected paths, current contract, recommended patch, tests, approvals, and residual risk. Report clean controls as evidence, not as a blanket assurance.
Error Handling
- Current docs disagree with installed types: pin both versions and resolve the discrepancy before editing.
- The target policy is missing: report the missing authority rather than assuming scraping is allowed.
- A fix would change behavior broadly: isolate it behind a canary and rollback boundary.
Examples
- "Review our Firecrawl wrapper" checks v2 methods, pagination, retries, provenance, and data boundaries.
- "All requests return 200" still fails if captured documents have metadata.statusCode errors.
Resources
Read official Firecrawl evidence before relying on an endpoint, SDK method, plan limit, price, retention option, or self-hosted release.
When not to use it
- →When scraping small, static sites where performance is not a concern
- →When using deprecated or incorrect package names
Limitations
- →Requires explicit configuration of limits and formats
- →Batch scraping is necessary for multiple URLs to avoid sequential latency
How it compares
This approach replaces manual, error-prone scraping scripts with a checklist-driven implementation that prevents common production failures.
Compared to similar skills
firecrawl-known-pitfalls side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| firecrawl-known-pitfalls (this skill) | 0 | 2mo | Review | Intermediate |
| schema-markup | 10 | 8mo | No flags | Intermediate |
| nextjs-developer | 328 | 4mo | No flags | Advanced |
| deepwiki-rs | 25 | 11mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
schema-markup
davila7
When the user wants to add, fix, or optimize schema markup and structured data on their site. Also use when the user mentions "schema markup," "structured data," "JSON-LD," "rich snippets," "schema.org," "FAQ schema," "product schema," "review schema," or "breadcrumb schema." For broader SEO issues, see seo-audit.
nextjs-developer
zenobi-us
Expert Next.js developer mastering Next.js 14+ with App Router and full-stack features. Specializes in server components, server actions, performance optimization, and production deployment with focus on building fast, SEO-friendly applications.
deepwiki-rs
sopaco
AI-powered Rust documentation generation engine for comprehensive codebase analysis, C4 architecture diagrams, and automated technical documentation. Use when Claude needs to analyze source code, understand software architecture, generate technical specs, or create professional documentation from any programming language.
writing-registry-meta
siriwatknp
Use this skill when writing meta file for MUI Treasury registry.
markdown-to-html
github
Convert Markdown files to HTML similar to `marked.js`, `pandoc`, `gomarkdown/markdown`, or similar tools; or writing custom script to convert markdown to html and/or working on web template systems like `jekyll/jekyll`, `gohugoio/hugo`, or similar web templating systems that utilize markdown documents, converting them to html. Use when asked to "convert markdown to html", "transform md to html", "render markdown", "generate html from markdown", or when working with .md files and/or web a templating system that converts markdown to HTML output. Supports CLI and Node.js workflows with GFM, CommonMark, and standard Markdown flavors.
coding-standards
affaan-m
适用于TypeScript、JavaScript、React和Node.js开发的通用编码标准、最佳实践和模式。