exa-observability
Provides instrumentation for tracking Exa API metrics, search latency, and health monitoring.
Install
mkdir -p .claude/skills/exa-observability && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/4829" && unzip -o skill.zip -d .claude/skills/exa-observability && rm skill.zipInstalls to .claude/skills/exa-observability
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Set up monitoring, metrics, and alerting for Exa search integrations.Key capabilities
- →Instrument Exa client for metrics emission
- →Track search latency by type
- →Monitor cache hit and miss rates
- →Configure Prometheus alert rules
How it works
The skill provides code patterns to instrument the Exa client, emit metrics for latency and errors, and define alert rules for production monitoring.
Inputs & outputs
When to use exa-observability
- →Monitor search API latency
- →Configure dashboard for search metrics
- →Set up alerts for API failures
- →Track cache hit/miss rates
About this skill
Exa Operational Telemetry
Overview
Instrument Exa calls with content-free metrics, traces, cost, and asynchronous lifecycle signals that support diagnosis without logging retrieved text. Treat credentials, queries, retrieved content, generated output, spend, and destructive state as separately governed boundaries.
Prerequisites
- The target repository, environment, Exa team, product surface, and accountable owner.
- The workload's data classification, latency and freshness promise, cost ceiling, and retention policy.
- Current first-party documentation plus credentials only for a narrowly approved live check.
Current Contract
Exa responses expose request IDs and often costDollars; errors expose status and tags; Contents exposes per-URL statuses. Agent, Monitor, Webset, and Batch operations add IDs, states, queues, events, and terminal outcomes that must be reconciled separately.
Authentication
For normal REST work, inject EXA_API_KEY from an approved server-side secret manager and send it only as Authorization: Bearer to the configured first-party Exa API host. Team Management service keys, hosted MCP OAuth or enterprise managed authorization, and payment-protocol calls are separate trust models. Never print, commit, place in a URL, or expose a credential to an untrusted client.
Instructions
- Define service-level objectives for each endpoint and asynchronous product.
- Emit endpoint, status, tag, latency, attempt, result count, and request ID safely.
- Measure Contents status mix and freshness mode rather than logging page content.
- Track run age, terminal state, webhook lag, queue depth, and reconciliation drift.
- Attribute cost by environment, workload, key, product, and owner.
- Alert on error-class shifts, stalled state, throttling, overload, credit risk, and missing events.
Tool Discipline
Use Read, Glob, and Grep to inspect repository code, configuration, fixtures, and evidence. Use Write and Edit only for approved implementation or documentation changes. Do not call Exa, run paid research, create or alter a Monitor, Webset, Agent run, Batch, team, member, API key, budget, webhook, or deployment merely because this skill was invoked.
Approval Boundaries
Require an accountable owner before live queries involving sensitive intent, production credentials, spend or rate-limit changes, forced live crawling, generated summaries, external delivery, deployment, member or key changes, schedule creation, or destructive cancellation, stopping, deletion, or revocation. Read-only repository inspection and synthetic offline validation do not authorize live vendor actions.
Failure Modes
- Raw queries, URLs, highlights, summaries, and outputs are not safe default labels.
- HTTP success can hide partial Contents failures.
- Webhook delivery metrics without run reconciliation can report false completion.
Output
Return the operation scope, environment, team and product surface, authorization class, contract and policy decisions, deterministic validation results, content-free identifiers, status and cost counts, risks, cleanup or rollback state, and a concise pass or fail receipt. Exclude credentials, raw queries, prompts, presigned URLs, retrieved content, generated output, and customer-derived data unless separately approved.
Example
- A trace records request ID, Search type, content mode, result count, latency, cost, and redacted policy outcome, never the query or text.
- Finish with request or resource IDs, assertion counts, cost and terminal state, rollback or deletion status, and the decision owner; never reproduce secrets or retrieved content.
Validation
Rerun the smallest relevant deterministic test, compare actual behavior with the requested outcome and current first-party contract, verify sensitive fields are absent from evidence, and confirm deadlines, terminal state, downstream retention, and rollback before reporting success.
References
Review the dated first-party evidence map before relying on any endpoint, parameter, search type, price, limit, beta, compliance, identity, retry, or lifecycle claim.
When not to use it
- →When the application lacks a metrics backend
Prerequisites
Limitations
- →Requires integration with a metrics backend
- →Alerting depends on external systems like PagerDuty or Slack
How it compares
It provides specific metrics and alert thresholds tailored to Exa's performance characteristics, such as search type latency and cache hit rates.
Compared to similar skills
exa-observability side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| exa-observability (this skill) | 1 | 2mo | Review | Intermediate |
| distributed-tracing | 5 | 4mo | No flags | Intermediate |
| service-mesh-observability | 5 | 4mo | No flags | Advanced |
| observability-engineer | 12 | 5mo | No flags | Advanced |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
distributed-tracing
wshobson
Implement distributed tracing with Jaeger and Tempo to track requests across microservices and identify performance bottlenecks. Use when debugging microservices, analyzing request flows, or implementing observability for distributed systems.
service-mesh-observability
wshobson
Implement comprehensive observability for service meshes including distributed tracing, metrics, and visualization. Use when setting up mesh monitoring, debugging latency issues, or implementing SLOs for service communication.
observability-engineer
sickn33
Build production-ready monitoring, logging, and tracing systems. Implements comprehensive observability strategies, SLI/SLO management, and incident response workflows. Use PROACTIVELY for monitoring infrastructure, performance optimization, or production reliability.
prometheus-configuration
wshobson
Set up Prometheus for comprehensive metric collection, storage, and monitoring of infrastructure and applications. Use when implementing metrics collection, setting up monitoring infrastructure, or configuring alerting systems.
langfuse
davila7
Expert in Langfuse - the open-source LLM observability platform. Covers tracing, prompt management, evaluation, datasets, and integration with LangChain, LlamaIndex, and OpenAI. Essential for debugging, monitoring, and improving LLM applications in production. Use when: langfuse, llm observability, llm tracing, prompt management, llm evaluation.
slo-implementation
wshobson
Define and implement Service Level Indicators (SLIs) and Service Level Objectives (SLOs) with error budgets and alerting. Use when establishing reliability targets, implementing SRE practices, or measuring service performance.