EX

Provides benchmarking, k6 load testing, and scaling patterns for high-throughput Exa applications.

Install

mkdir -p .claude/skills/exa-load-scale && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/8846" && unzip -o skill.zip -d .claude/skills/exa-load-scale && rm skill.zip

Installs to .claude/skills/exa-load-scale

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Implement Exa load testing, capacity planning, and scaling strategies.
70 charsno explicit “when” trigger
Advanced

Key capabilities

  • →Conduct k6 load testing against application wrappers
  • →Implement request queuing to respect 10 QPS limits
  • →Configure LRU caching for search results
  • →Estimate capacity based on daily search volume
  • →Generate performance benchmark reports

How it works

The skill provides load testing scripts, request queueing patterns to stay under rate limits, and caching strategies to optimize throughput.

Inputs & outputs

You give it
Search traffic parameters and API configuration
You get back
Performance benchmarks and capacity estimates

When to use exa-load-scale

  • →Conduct load testing with k6
  • →Plan capacity for search traffic
  • →Optimize throughput under rate limits
  • →Implement search result caching

About this skill

Exa Load and Capacity Verification

Overview

Validate Exa capacity with synthetic workload models, endpoint-specific budgets, bounded queues, and stop conditions. Treat credentials, queries, retrieved content, generated output, spend, and destructive state as separately governed boundaries.

Prerequisites

  • The target repository, environment, Exa team, product surface, and accountable owner.
  • The workload's data classification, latency and freshness promise, cost ceiling, and retention policy.
  • Current first-party documentation plus credentials only for a narrowly approved live check.

Current Contract

Published QPS differs by endpoint and enterprise limits can differ by team. Search, Contents, and Answer have distinct ceilings; asynchronous Agent or Batch is not equivalent to synchronous load. Enterprise Batch is beta, must be enabled, and requires the current first-party Exa-Beta header documented for that feature.

Authentication

For normal REST work, inject EXA_API_KEY from an approved server-side secret manager and send it only as Authorization: Bearer to the configured first-party Exa API host. Team Management service keys, hosted MCP OAuth or enterprise managed authorization, and payment-protocol calls are separate trust models. Never print, commit, place in a URL, or expose a credential to an untrusted client.

Instructions

  1. Obtain vendor and service-owner approval for the environment, volume, duration, and spend.
  2. Model endpoint mix, result sizes, content modes, freshness, and asynchronous work.
  3. Use synthetic non-sensitive inputs and shared limiters across workers.
  4. Ramp gradually with hard stop conditions for errors, latency, cost, and backlog.
  5. Measure 429 separately from 503, partial crawl status, and downstream saturation.
  6. Drain queues, reconcile runs or batches, and delete test artifacts after the evidence window.

Tool Discipline

Use Read, Glob, and Grep to inspect repository code, configuration, fixtures, and evidence. Use Write and Edit only for approved implementation or documentation changes. Do not call Exa, run paid research, create or alter a Monitor, Webset, Agent run, Batch, team, member, API key, budget, webhook, or deployment merely because this skill was invoked.

Approval Boundaries

Require an accountable owner before live queries involving sensitive intent, production credentials, spend or rate-limit changes, forced live crawling, generated summaries, external delivery, deployment, member or key changes, schedule creation, or destructive cancellation, stopping, deletion, or revocation. Read-only repository inspection and synthetic offline validation do not authorize live vendor actions.

Failure Modes

  • Do not load-test production or third-party target sites without authorization.
  • More workers can violate a shared team or network limit.
  • A completed Batch can contain failed item rows that need separate accounting.

Output

Return the operation scope, environment, team and product surface, authorization class, contract and policy decisions, deterministic validation results, content-free identifiers, status and cost counts, risks, cleanup or rollback state, and a concise pass or fail receipt. Exclude credentials, raw queries, prompts, presigned URLs, retrieved content, generated output, and customer-derived data unless separately approved.

Example

  • Ramp synthetic Search to the approved ceiling while Contents runs in its own pool, then stop on throttle or cost thresholds.
  • Finish with request or resource IDs, assertion counts, cost and terminal state, rollback or deletion status, and the decision owner; never reproduce secrets or retrieved content.

Validation

Rerun the smallest relevant deterministic test, compare actual behavior with the requested outcome and current first-party contract, verify sensitive fields are absent from evidence, and confirm deadlines, terminal state, downstream retention, and rollback before reporting success.

References

Review the dated first-party evidence map before relying on any endpoint, parameter, search type, price, limit, beta, compliance, identity, retry, or lifecycle claim.

When not to use it

  • →When the application does not use Redis for caching
  • →When the environment lacks k6

Prerequisites

k6 load testing tool installedTest environment Exa API keyRedis for result caching

Limitations

  • →Default rate limit is 10 QPS
  • →Deep search type has higher latency than instant search

How it compares

It provides specific capacity planning and rate-limiting strategies tailored to Exa's 10 QPS default limit rather than generic load testing.

Compared to similar skills

exa-load-scale side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
exa-load-scale (this skill)02moReviewAdvanced
chrome-devtools418moReviewIntermediate
code-coverage-with-gcov156moReviewIntermediate
angular-best-practices215moNo flagsAdvanced

Try saying

Example prompts that trigger this skill in your AI assistant.

More by jeremylongshore

View all by jeremylongshore →

analyzing-logs

jeremylongshore

Analyze application logs to detect performance issues, identify error patterns, and improve stability by extracting key insights.

14123

ollama-setup

jeremylongshore

Configure auto-configure Ollama when user needs local LLM deployment, free AI alternatives, or wants to eliminate hosted API costs. Trigger phrases: "install ollama", "local AI", "free LLM", "self-hosted AI", "replace OpenAI", "no API costs". Use when appropriate context detected. Trigger with relevant phrases based on skill purpose.

1167

backtesting-trading-strategies

jeremylongshore

Backtest crypto and traditional trading strategies against historical data. Calculates performance metrics (Sharpe, Sortino, max drawdown), generates equity curves, and optimizes strategy parameters. Use when user wants to test a trading strategy, validate signals, or compare approaches. Trigger with phrases like "backtest strategy", "test trading strategy", "historical performance", "simulate trades", "optimize parameters", or "validate signals".

1071

generating-database-seed-data

jeremylongshore

Process this skill enables AI assistant to generate realistic test data and database seed scripts for development and testing environments. it uses faker libraries to create realistic data, maintains relational integrity, and allows configurable data volumes. u... Use when working with databases or data models. Trigger with phrases like 'database', 'query', or 'schema'.

1033

cursor-codebase-indexing

jeremylongshore

Execute set up and optimize Cursor codebase indexing. Triggers on "cursor index setup", "codebase indexing", "index codebase", "cursor semantic search". Use when working with cursor codebase indexing functionality. Trigger with phrases like "cursor codebase indexing", "cursor indexing", "cursor".

885

testing-mobile-apps

jeremylongshore

Execute mobile app testing on iOS and Android devices/simulators. Use when performing specialized testing. Trigger with phrases like "test mobile app", "run iOS tests", or "validate Android functionality".

810

Search skills

Search the agent skills registry