instantly-incident-runbook
Offers triage, mitigation, and recovery protocols for campaign health crises and API outages.
Install
mkdir -p .claude/skills/instantly-incident-runbook && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/7125" && unzip -o skill.zip -d .claude/skills/instantly-incident-runbook && rm skill.zipInstalls to .claude/skills/instantly-incident-runbook
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Execute Instantly.ai incident response procedures with triage, mitigation,Key capabilities
- →Triage campaign health and account vitals
- →Mitigate broken account issues
- →Handle bounce protect triggers
- →Diagnose webhook delivery failures
- →Execute recovery steps for degraded warmup
How it works
It uses diagnostic scripts to query the Instantly API for campaign statuses, account vitals, and webhook health. It then provides structured recovery steps based on the identified severity level.
Inputs & outputs
When to use instantly-incident-runbook
- →Triage campaign failures
- →Execute incident mitigation steps
- →Perform post-incident analysis
About this skill
Instantly Incident Response
Overview
Contain a production incident while preserving evidence, tenant boundaries, and reversible recovery actions. Record assumptions, evidence, approval state, and rollback ownership so another operator can reproduce the result.
Prerequisites
- The target repository, Instantly workspace, environment, and accountable owner
- Current security, privacy, compliance, capacity, and change-control requirements
- An approved API v2 key only when a bounded live verification is necessary
Tool Discipline
Use Read, Glob, and Grep to inspect code, configuration, and evidence. Use WebFetch only for current first-party Instantly documentation and package metadata. Use Write or Edit only when implementation was requested and exact target files are known; never write credentials, lead data, email content, or unrestricted environment output.
Current Contract
- Campaign pause, account pause, webhook resume, and key revocation are state-changing actions.
- Sending-status, account vitals, background jobs, and webhook aggregates provide documented evidence surfaces.
- Workspace-wide 429s require coordinated throttling across every key.
Authentication
Use an API v2 key as Authorization: Bearer <key> against https://api.instantly.ai/api/v2. Grant only the endpoint-specific scopes needed, inject the key from an approved server-side secret manager, and never print, persist, commit, or place it in a URL. Treat key creation, rotation, revocation, member changes, workspace delegation, and production access as owner-approved actions.
Instructions
- Declare severity, affected workspace, owner, time window, and customer impact.
- Preserve redacted request IDs, statuses, deploy SHA, campaign state, and aggregate health.
- Classify identity, scope, quota, sending, account, job, webhook, or provider failure.
- Propose the smallest containment action with blast radius and rollback.
- Execute pause, resume, revoke, or rollback only after incident-owner approval unless a pre-approved runbook applies.
- Verify recovery, monitor recurrence, and record follow-up ownership.
Approval Boundaries
Do not create, rotate, reveal, or revoke keys; invite or remove members; delegate across workspaces; connect sending accounts; create or activate campaigns; import or delete leads; change suppression or retention; register, patch, resume, or delete webhooks; alter plans or paid capacity; transmit diagnostics; or perform another production mutation without explicit approval from the accountable owner. Keep diagnosis read-only unless implementation was requested.
Output
Return the workspace-safe scope, files and contracts inspected, exact API v2 routes and required scopes, evidence collected, validation result, sensitive fields redacted, remaining risk, accountable owner, approval state, and rollback or next action.
Error Handling
| Condition | Response |
|---|---|
401 | Stop and verify that the bearer key exists, is current, and was not revoked. |
403 | Stop and compare the operation with its exact required scope; do not broaden to all:all by default. |
429 | Coordinate the workspace-wide budget, honor endpoint overrides, and bound retries. |
| Schema or tenant mismatch | Fail closed, preserve redacted evidence, and do not retry a mutation. |
Examples
Use a compact handoff that makes scope, mutation authority, and evidence reviewable.
Input:
incident=INC-123; severity=P2; action=diagnose-only
Expected handoff:
cause=scope-regression; containment=proposed; approval=required
Resources
When not to use it
- →When the Instantly API is completely unreachable
- →When dashboard access is sufficient for resolution
Limitations
- →Can't reach API during Instantly outage
- →Can't pause accounts due to 403 scope errors
- →Runbook script rate-limited by too many diagnostic calls
How it compares
This runbook automates the triage of campaign and account health metrics compared to manual dashboard inspection.
Compared to similar skills
instantly-incident-runbook side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| instantly-incident-runbook (this skill) | 1 | 2mo | Review | Advanced |
| home-assistant-manager | 9 | 10mo | Review | Advanced |
| firecrawl-incident-runbook | 1 | 2mo | Review | Intermediate |
| netalertx-plugin-run-development | 1 | 8mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
home-assistant-manager
komal-SkyNET
Expert-level Home Assistant configuration management with efficient deployment workflows (git and rapid scp iteration), remote CLI access via SSH and hass-cli, automation verification protocols, log analysis, reload vs restart optimization, and comprehensive Lovelace dashboard management for tablet-optimized UIs. Includes template patterns, card types, debugging strategies, and real-world examples.
firecrawl-incident-runbook
jeremylongshore
Execute FireCrawl incident response procedures with triage, mitigation, and postmortem. Use when responding to FireCrawl-related outages, investigating errors, or running post-incident reviews for FireCrawl integration failures. Trigger with phrases like "firecrawl incident", "firecrawl outage", "firecrawl down", "firecrawl on-call", "firecrawl emergency", "firecrawl broken".
netalertx-plugin-run-development
netalertx
Create and run NetAlertX plugins. Use this when asked to create plugin, run plugin, test plugin, plugin development, or execute plugin script.
k8s-browser
rohitg00
Browser automation for Kubernetes dashboards and web UIs. Use when interacting with Kubernetes Dashboard, Grafana, ArgoCD UI, or other web interfaces. Requires MCP_BROWSER_ENABLED=true.
lokalise-incident-runbook
jeremylongshore
Execute Lokalise incident response procedures with triage, mitigation, and postmortem. Use when responding to Lokalise-related outages, investigating errors, or running post-incident reviews for Lokalise integration failures. Trigger with phrases like "lokalise incident", "lokalise outage", "lokalise down", "lokalise on-call", "lokalise emergency", "translations broken".
anomaly-detection
dadbodgeoff
Rule-based anomaly detection for production systems with configurable thresholds, cooldown periods to prevent alert storms, and error pattern tracking for repeated failures.