exa-performance-tuning
Optimizes Exa API latency by selecting search types based on requirements and implementing caching.
Install
mkdir -p .claude/skills/exa-performance-tuning && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/9311" && unzip -o skill.zip -d .claude/skills/exa-performance-tuning && rm skill.zipInstalls to .claude/skills/exa-performance-tuning
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Optimize Exa API performance with search type selection, caching, andKey capabilities
- →Select an Exa search type based on latency requirements
- →Minimize content retrieval to reduce latency
- →Cache search results using an LRU cache
- →Parallelize independent Exa search queries
- →Implement a two-phase search for selective content retrieval
- →Normalize queries to increase cache hit rates
How it works
The skill optimizes Exa API performance by selecting appropriate search types, reducing result content, caching responses, and executing queries in parallel.
Inputs & outputs
When to use exa-performance-tuning
- →Improve search response times
- →Implement request caching
- →Optimize production search workloads
About this skill
Exa Latency and Retrieval Performance Tuning
Overview
Tune Exa search type, content mode, freshness, result count, and concurrency against measured latency and retrieval quality. Treat credentials, queries, retrieved content, generated output, spend, and destructive state as separately governed boundaries.
Prerequisites
- The target repository, environment, Exa team, product surface, and accountable owner.
- The workload's data classification, latency and freshness promise, cost ceiling, and retention policy.
- Current first-party documentation plus credentials only for a narrowly approved live check.
Current Contract
Search types trade latency and synthesis depth; content extraction, outputSchema, forced livecrawl, summaries, and subpages add work. Highlights are usually more token-efficient than full text. Published latency values are guidance, not a service-specific SLO.
Authentication
For normal REST work, inject EXA_API_KEY from an approved server-side secret manager and send it only as Authorization: Bearer to the configured first-party Exa API host. Team Management service keys, hosted MCP OAuth or enterprise managed authorization, and payment-protocol calls are separate trust models. Never print, commit, place in a URL, or expose a credential to an untrusted client.
Instructions
- Build a consented benchmark set with explicit relevance and freshness judgments.
- Measure the current adapter end to end, including queue and downstream processing.
- Vary one dimension at a time: type, result count, content mode, freshness, or subpages.
- Track p50, p95, errors, content size, relevance, freshness, and cost together.
- Choose separate profiles for interactive, background, and deep-research paths.
- Canary the winning profile and retain rollback thresholds.
Tool Discipline
Use Read, Glob, and Grep to inspect repository code, configuration, fixtures, and evidence. Use Write and Edit only for approved implementation or documentation changes. Do not call Exa, run paid research, create or alter a Monitor, Webset, Agent run, Batch, team, member, API key, budget, webhook, or deployment merely because this skill was invoked.
Approval Boundaries
Require an accountable owner before live queries involving sensitive intent, production credentials, spend or rate-limit changes, forced live crawling, generated summaries, external delivery, deployment, member or key changes, schedule creation, or destructive cancellation, stopping, deletion, or revocation. Read-only repository inspection and synthetic offline validation do not authorize live vendor actions.
Failure Modes
- Faster results can be less relevant or stale.
- Forced livecrawl and deep synthesis stack latency rather than replacing it.
- Vendor request latency alone omits queues, parsing, reranking, and model consumption.
Output
Return the operation scope, environment, team and product surface, authorization class, contract and policy decisions, deterministic validation results, content-free identifiers, status and cost counts, risks, cleanup or rollback state, and a concise pass or fail receipt. Exclude credentials, raw queries, prompts, presigned URLs, retrieved content, generated output, and customer-derived data unless separately approved.
Example
- Compare fast plus highlights against auto plus capped text on the same query set and promote only if the quality floor holds.
- Finish with request or resource IDs, assertion counts, cost and terminal state, rollback or deletion status, and the decision owner; never reproduce secrets or retrieved content.
Validation
Rerun the smallest relevant deterministic test, compare actual behavior with the requested outcome and current first-party contract, verify sensitive fields are absent from evidence, and confirm deadlines, terminal state, downstream retention, and rollback before reporting success.
References
Review the dated first-party evidence map before relying on any endpoint, parameter, search type, price, limit, beta, compliance, identity, retry, or lifecycle claim.
When not to use it
- →When maximum coverage is required, as this may increase latency
- →When complex research questions require deep-reasoning search type
Limitations
- →Neural search on complex queries can take 3 seconds or more
- →Large pages or slow sources can cause content retrieval timeouts
- →Unique queries each time can lead to a high cache miss rate
How it compares
This skill provides specific strategies like two-phase search and query normalization to improve Exa API performance, rather than just making basic API calls.
Compared to similar skills
exa-performance-tuning side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| exa-performance-tuning (this skill) | 3 | 2mo | Review | Intermediate |
| mcp-builder | 136 | 5mo | Review | Advanced |
| stripe-integration | 48 | 4mo | No flags | Advanced |
| copilot-sdk | 7 | 5mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
mcp-builder
anthropics
Guide for creating high-quality MCP (Model Context Protocol) servers that enable LLMs to interact with external services through well-designed tools. Use when building MCP servers to integrate external APIs or services, whether in Python (FastMCP) or Node/TypeScript (MCP SDK).
stripe-integration
wshobson
Implement Stripe payment processing for robust, PCI-compliant payment flows including checkout, subscriptions, and webhooks. Use when integrating Stripe payments, building subscription systems, or implementing secure checkout flows.
copilot-sdk
github
Build agentic applications with GitHub Copilot SDK. Use when embedding AI agents in apps, creating custom tools, implementing streaming responses, managing sessions, connecting to MCP servers, or creating custom agents. Triggers on Copilot SDK, GitHub SDK, agentic app, embed Copilot, programmable agent, MCP server, custom agent.
openrouter-hello-world
jeremylongshore
Create your first OpenRouter API request with a simple example. Use when learning OpenRouter or testing your setup. Trigger with phrases like 'openrouter hello world', 'openrouter first request', 'openrouter quickstart', 'test openrouter'.
telegram-dev
2025Emma
Telegram 生态开发全栈指南 - 涵盖 Bot API、Mini Apps (Web Apps)、MTProto 客户端开发。包括消息处理、支付、内联模式、Webhook、认证、存储、传感器 API 等完整开发资源。
mistral-upgrade-migration
jeremylongshore
Analyze, plan, and execute Mistral AI SDK upgrades with breaking change detection. Use when upgrading Mistral SDK versions, detecting deprecations, or migrating to new API versions. Trigger with phrases like "upgrade mistral", "mistral migration", "mistral breaking changes", "update mistral SDK", "analyze mistral version".