firecrawl-incident-runbook
Standardized triage, mitigation, and investigation procedures for Firecrawl integration outages and errors.
Install
mkdir -p .claude/skills/firecrawl-incident-runbook && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/2395" && unzip -o skill.zip -d .claude/skills/firecrawl-incident-runbook && rm skill.zipInstalls to .claude/skills/firecrawl-incident-runbook
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Execute Firecrawl incident response procedures with triage, mitigation,Key capabilities
- →Test Firecrawl API health and check credit balance
- →Verify API key validity and rotate if needed
- →Implement emergency rate limiting with delays
- →Enable graceful degradation with cached content fallback
- →Collect debug bundles with credits and application logs
- →Generate postmortem reports with timeline and root cause
How it works
The skill starts with quick triage steps to test API health and credit balance, then uses a decision tree to identify the error type. It provides immediate actions for authentication failures, credit exhaustion, rate limiting, and Firecrawl outages, and includes steps for post-incident evidence collection and reporting
Inputs & outputs
When to use firecrawl-incident-runbook
- →Troubleshoot API service outages
- →Triage crawl job failures
- →Check credit balances
- →Run incident post-mortems
About this skill
Firecrawl Incident Response
Overview
Stabilize the affected workflow, preserve privacy-safe evidence, and restore service through a tested degraded mode or rollback. Separate Firecrawl service health, target-origin behavior, and internal pipeline failures.
Prerequisites
- The target repository or integration path and the requested operator outcome.
- The source authorization, data classification, and environment policy.
- Current Firecrawl documentation, credentials only when needed, and an owner for approvals.
Current Contract
Firecrawl's error catalog defines retryability; async operations expose job/status/cancellation surfaces; queue status helps distinguish capacity pressure; status polling remains the recovery path when webhooks are delayed or exhausted. A crawl webhook set does not include a crawl.failed event in the current documented event list.
Authentication
For authenticated Cloud operations, inject FIRECRAWL_API_KEY from an approved secret manager. REST requests use Authorization: Bearer with the key. Never print, commit, transmit, or place a key in a URL. Keyless access is suitable only where the current documentation explicitly allows it and the workload accepts its limits; production workflows should make identity and team ownership explicit.
Instructions
- Declare incident owner, severity, affected operation/environment, start time, customer impact, data risk, approved communication channel, and next update time.
- Pause or bound producers before investigating if retries, crawl scope, or pay-as-you-go could amplify cost or target load.
- Check internal deployments and dependencies, Firecrawl service evidence, credentials, credits, team restrictions, queue/concurrency, job state, webhook delivery, and target-origin status.
- Classify the failure with the official error catalog. Retry only documented retryable classes, honor Retry-After, and cap attempts.
- Choose a reversible mitigation: reduce concurrency, narrow limits, switch to polling, serve last known approved content, disable an expensive option, or roll back the application release.
- Verify recovery with a synthetic or approved canary and confirm queue drain, error rate, output quality, data integrity, and spend stabilization.
- Communicate resolution, retain redacted evidence, and create owned corrective actions for detection, prevention, runbook, and rollback gaps.
Tool Discipline
Use Read, Glob, and Grep to inspect code, configuration, tests, and evidence. Use Write/Edit only for approved implementation or documentation changes. Do not call Firecrawl, rotate keys, change account settings, scrape a target, or deploy merely because this skill was invoked.
Approval Boundaries
Require incident-command approval before key rotation, plan or pay-as-you-go changes, traffic failover, target-scope changes, disabling security/retention controls, or vendor disclosure.
Output
Return the timeline, impact, classification, mitigations, approvals, canary and recovery evidence, residual risk, next update, and post-incident actions.
Error Handling
- Service state is ambiguous: hold producers and collect bounded evidence rather than mass retrying.
- Mitigation changes data quality or freshness: label degraded output and obtain product-owner acceptance.
- Potential credential or content exposure: invoke the security incident path and preserve evidence without broad collection.
Examples
- "Crawls stopped completing" checks deployment, queue, job status, errors, and target status before retrying.
- "Webhooks stopped" switches to bounded status polling while signature and delivery failures are investigated.
Resources
Read official Firecrawl evidence before relying on an endpoint, SDK method, plan limit, price, retention option, or self-hosted release.
When not to use it
- →When the issue is not related to Firecrawl integration failures
- →When the problem is not an API outage, credential issue, or credit exhaustion
- →When the problem is not a crawl job failure or webhook delivery problem
Prerequisites
Limitations
- →Cannot reach Firecrawl API if there is a network or DNS issue
- →Scrapes may return empty if the target site changed or bot detection is active
- →Crawl jobs may not complete if there is a queue backup
How it compares
This skill provides a structured, automated runbook for Firecrawl incidents, including specific bash and TypeScript commands for triage and mitigation, unlike a general incident response process.
Compared to similar skills
firecrawl-incident-runbook side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| firecrawl-incident-runbook (this skill) | 1 | 2mo | Review | Intermediate |
| openevidence-incident-runbook | 0 | 2mo | Review | Intermediate |
| perplexity-incident-runbook | 0 | 2mo | Review | Intermediate |
| managing-filestack | 0 | 6mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
openevidence-incident-runbook
jeremylongshore
Execute OpenEvidence incident response procedures with triage, mitigation, and postmortem. Use when responding to OpenEvidence-related outages, investigating errors, or running post-incident reviews for clinical AI integration failures. Trigger with phrases like "openevidence incident", "openevidence outage", "openevidence down", "openevidence emergency", "clinical ai broken".
perplexity-incident-runbook
jeremylongshore
Execute Perplexity incident response procedures with triage, mitigation, and postmortem. Use when responding to Perplexity-related outages, investigating errors, or running post-incident reviews for Perplexity integration failures. Trigger with phrases like "perplexity incident", "perplexity outage", "perplexity down", "perplexity on-call", "perplexity emergency", "perplexity broken".
managing-filestack
cloudthinker-ai
|
telegram-bot-builder
davila7
Expert in building Telegram bots that solve real problems - from simple automation to complex AI-powered bots. Covers bot architecture, the Telegram Bot API, user experience, monetization strategies, and scaling bots to thousands of users. Use when: telegram bot, bot api, telegram automation, chat bot telegram, tg bot.
mcp-integration
anthropics
This skill should be used when the user asks to "add MCP server", "integrate MCP", "configure MCP in plugin", "use .mcp.json", "set up Model Context Protocol", "connect external service", mentions "${CLAUDE_PLUGIN_ROOT} with MCP", or discusses MCP server types (SSE, stdio, HTTP, WebSocket). Provides comprehensive guidance for integrating Model Context Protocol servers into Claude Code plugins for external tool and service integration.
n8n-workflow-patterns
czlonkowski
Proven workflow architectural patterns from real n8n workflows. Use when building new workflows, designing workflow structure, choosing workflow patterns, planning workflow architecture, or asking about webhook processing, HTTP API integration, database operations, AI agent workflows, or scheduled tasks.