firecrawl-load-scale
Strategies and scripts to scale Firecrawl scraping pipelines while staying within plan rate limits.
Install
mkdir -p .claude/skills/firecrawl-load-scale && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/4743" && unzip -o skill.zip -d .claude/skills/firecrawl-load-scale && rm skill.zipInstalls to .claude/skills/firecrawl-load-scale
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Load test and scale Firecrawl scraping pipelines with concurrency controlKey capabilities
- →Measure scraping throughput
- →Implement batch scraping for efficiency
- →Manage concurrent requests with p-queue
- →Scale async crawl jobs
- →Estimate capacity and credit runway
How it works
It utilizes batching and queue management to maximize throughput while staying within the rate limits defined by the user's Firecrawl plan.
Inputs & outputs
When to use firecrawl-load-scale
- →Implementing batch scraping
- →Optimizing concurrent scrape performance
- →Planning capacity for large crawl jobs
- →Testing scraping throughput
About this skill
Firecrawl Capacity and Load Validation
Overview
Find the safe operating envelope with synthetic or explicitly authorized targets. A load test must not become an uncontrolled scrape campaign or consume unapproved credits.
Prerequisites
- The target repository or integration path and the requested operator outcome.
- The source authorization, data classification, and environment policy.
- Current Firecrawl documentation, credentials only when needed, and an owner for approvals.
Current Contract
Firecrawl separates per-team requests-per-minute limits from concurrent browser capacity. Work beyond browser capacity can queue, queue time counts against request timeout, and queue status exposes availability. Crawl and batch calls also accept maxConcurrency, while a crawl delay forces concurrency to one.
Authentication
For authenticated Cloud operations, inject FIRECRAWL_API_KEY from an approved secret manager. REST requests use Authorization: Bearer with the key. Never print, commit, transmit, or place a key in a URL. Keyless access is suitable only where the current documentation explicitly allows it and the workload accepts its limits; production workflows should make identity and team ownership explicit.
Instructions
- Define the approved target set, environment, maximum requests/pages/credits, concurrency steps, duration, stop thresholds, and owner. Prefer a controlled synthetic origin.
- Measure a single-worker baseline for submission latency, completion latency, throughput, queue time, error classes, origin status, output size, quality, and credit usage.
- Increase producer concurrency in small steps below the current plan/team limits. Observe Firecrawl queue status and the application's own backlog separately.
- For crawl or batch, test maxConcurrency and explicit limits; do not assume client request concurrency equals page-processing concurrency.
- Exercise 429 rate pressure, concurrency queuing, Retry-After handling, timeout, cancellation, partial pagination, and backpressure using synthetic responses before live load.
- Stop on error, spend, target-load, latency, queue-age, or quality thresholds. Drain or cancel work according to the test plan.
- Report the sustainable envelope with headroom, bottleneck evidence, configuration, cost, and rollback; never publish captured bodies.
Tool Discipline
Use Read, Glob, and Grep to inspect code, configuration, tests, and evidence. Use Write/Edit only for approved implementation or documentation changes. Do not call Firecrawl, rotate keys, change account settings, scrape a target, or deploy merely because this skill was invoked.
Approval Boundaries
Require approval before live load, increasing credits or plan capacity, using third-party targets, changing target delay/concurrency, or extending the test window.
Output
Return the load profile, target authorization, baseline and stepped metrics, queue behavior, throttle/error counts, credits, sustainable envelope, stop event, cleanup, and capacity recommendation.
Error Handling
- Queue grows without stable throughput: stop producers and drain before testing another step.
- Target-origin failures rise: treat target protection as a stop condition, not a reason to add proxies.
- Credit telemetry is delayed: hold the next step until async usage settles.
Examples
- "Can we run 100 workers?" derives the answer from current team limits and a stepped test.
- "Stress a competitor's site" is refused because target authorization is absent.
Resources
Read official Firecrawl evidence before relying on an endpoint, SDK method, plan limit, price, retention option, or self-hosted release.
When not to use it
- →When scraping a single URL
- →When the target site has strict rate limits that conflict with high concurrency
Prerequisites
Limitations
- →Throughput is constrained by the specific Firecrawl plan tier
- →Batch scraping requires splitting large lists into chunks
How it compares
This method uses programmatic concurrency control instead of basic sequential loops to optimize performance.
Compared to similar skills
firecrawl-load-scale side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| firecrawl-load-scale (this skill) | 1 | 2mo | Review | Advanced |
| chrome-devtools | 41 | 8mo | Review | Intermediate |
| playwright-browser-automation | 29 | 9mo | Review | Intermediate |
| bullmq-specialist | 25 | 8mo | No flags | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
chrome-devtools
mrgoonie
Browser automation, debugging, and performance analysis using Puppeteer CLI scripts. Use for automating browsers, taking screenshots, analyzing performance, monitoring network traffic, web scraping, form automation, and JavaScript debugging.
playwright-browser-automation
lackeyjb
Complete browser automation with Playwright. Auto-detects dev servers, writes clean test scripts to /tmp. Test pages, fill forms, take screenshots, check responsive design, validate UX, test login flows, check links, automate any browser task. Use when user wants to test websites, automate browser interactions, validate web functionality, or perform any browser-based testing.
bullmq-specialist
davila7
BullMQ expert for Redis-backed job queues, background processing, and reliable async execution in Node.js/TypeScript applications. Use when: bullmq, bull queue, redis queue, background job, job queue.
perf-lighthouse
tech-leads-club
Run Lighthouse audits locally via CLI or Node API, parse and interpret reports, set performance budgets. Use when measuring site performance, understanding Lighthouse scores, setting up budgets, or integrating audits into CI. Triggers on: lighthouse, run lighthouse, lighthouse score, performance audit, performance budget.
documenso-local-dev-loop
jeremylongshore
Set up local development environment and testing workflow for Documenso. Use when configuring dev environment, setting up test workflows, or establishing rapid iteration patterns with Documenso. Trigger with phrases like "documenso local dev", "documenso development", "test documenso locally", "documenso dev environment".
replit-load-scale
jeremylongshore
Implement Replit load testing, auto-scaling, and capacity planning strategies. Use when running performance tests, configuring horizontal scaling, or planning capacity for Replit integrations. Trigger with phrases like "replit load test", "replit scale", "replit performance test", "replit capacity", "replit k6", "replit benchmark".