MI

mistral-incident-runbook

Covers triage, mitigation, and postmortem procedures for Mistral AI service outages.

Install

mkdir -p .claude/skills/mistral-incident-runbook && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/7107" && unzip -o skill.zip -d .claude/skills/mistral-incident-runbook && rm skill.zip

Installs to .claude/skills/mistral-incident-runbook

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Execute Mistral AI incident response procedures with triage, mitigation,
72 charsno explicit “when” trigger
Advanced

Key capabilities

  • →Classify incident severity
  • →Perform API health triage
  • →Execute mitigation steps for specific error codes
  • →Collect diagnostic evidence
  • →Conduct post-incident reviews

How it works

The skill provides a structured runbook including a triage script to check API health, a decision tree for error mitigation, and templates for incident communication and postmortems.

Inputs & outputs

You give it
Incident report or error observation
You get back
Mitigation status and evidence bundle

When to use mistral-incident-runbook

  • →Triage service outages
  • →Execute incident response
  • →Conduct postmortems
  • →Check API health

About this skill

Mistral Incident Response

Overview

Restore safety before throughput. Classify, stop amplification, protect evidence, reconcile ambiguous work, and use the smallest reversible containment action.

Prerequisites

  • An incident commander, severity policy, inventory, and current runbooks.
  • Content-free telemetry, deployment controls, rotation, and admin/status evidence.
  • Communication, evidence-retention, and post-incident owners.

Current Contract

Provider, account, app, and data incidents differ. Workspace caps can suspend access; Files, Batch, Conversations, Agents, and Workflows may outlive a failed request and need reconciliation.

Authentication

Never paste keys into incident channels. On exposure, revoke through approved admin, rotate consumers, and verify old-key denial without displaying values.

Instructions

  1. Declare scope, severity, commander, UTC timeline, environments, and impact.
  2. Stop retry storms and risky work while preserving queue/idempotency evidence.
  3. Classify credential, provider, capacity, spend, data, model, deploy, or state failure.
  4. Choose bounded containment: disable, shed, pause, rollback, rotate, or isolate.
  5. Reconcile files, jobs, runs, conversations, and app actions before replay.
  6. Recover through synthetic canary, monitor convergence, communicate, and assign actions.

Tool Discipline

Use Read, Glob, and Grep to inspect code, locks, configuration, tests, and evidence. Use Write and Edit only for approved repository changes. Invocation alone does not authorize network calls, paid usage, uploads, stateful resources, admin mutations, deployments, or deletion.

Approval Boundaries

The commander must approve revocation, traffic, queue or job changes, deletion, limits or spend changes, rollback, and customer communication. Record the approver and exact containment scope.

Error Handling

  • Blind replay can duplicate paid/stateful work.
  • Purging queues can destroy reconciliation evidence.
  • Provider recovery does not prove app backlog convergence.

Output

Return scope and timeline, classification, containment approvals, affected state, reconciliation, recovery, residual risk, communications, and follow-ups. Keep unresolved ambiguity visible with a named owner.

Examples

  • Disable batch submission while reconciling accepted jobs.
  • Rotate an exposed key and prove old-key denial plus new-key canary.

Validation

Tabletop key exposure, 429, outage, cap reached, content leak, stuck stream, and ambiguous state; verify rollback/comms.

Resources

  • Current first-party evidence map — recheck dated sources before relying on mutable endpoints, models, limits, prices, preview status, or retention.
  • Record live account observations as environment-specific evidence, not universal Mistral guarantees.

When not to use it

  • →When the incident is unrelated to Mistral AI
  • →When the service is fully operational

Prerequisites

MISTRAL_API_KEYkubectl accesscurl

Limitations

  • →Metrics triage requires Prometheus availability
  • →Evidence collection excludes API keys

How it compares

It provides a standardized, repeatable process for incident response that includes specific diagnostic commands and communication templates.

Compared to similar skills

mistral-incident-runbook side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
mistral-incident-runbook (this skill)12moReviewAdvanced
observability-engineer125moNo flagsAdvanced
mlops-engineer35moNo flagsAdvanced
senior-devops79moReviewAdvanced

Try saying

Example prompts that trigger this skill in your AI assistant.

More by jeremylongshore

View all by jeremylongshore →

analyzing-logs

jeremylongshore

Analyze application logs to detect performance issues, identify error patterns, and improve stability by extracting key insights.

14123

ollama-setup

jeremylongshore

Configure auto-configure Ollama when user needs local LLM deployment, free AI alternatives, or wants to eliminate hosted API costs. Trigger phrases: "install ollama", "local AI", "free LLM", "self-hosted AI", "replace OpenAI", "no API costs". Use when appropriate context detected. Trigger with relevant phrases based on skill purpose.

1167

backtesting-trading-strategies

jeremylongshore

Backtest crypto and traditional trading strategies against historical data. Calculates performance metrics (Sharpe, Sortino, max drawdown), generates equity curves, and optimizes strategy parameters. Use when user wants to test a trading strategy, validate signals, or compare approaches. Trigger with phrases like "backtest strategy", "test trading strategy", "historical performance", "simulate trades", "optimize parameters", or "validate signals".

1071

generating-database-seed-data

jeremylongshore

Process this skill enables AI assistant to generate realistic test data and database seed scripts for development and testing environments. it uses faker libraries to create realistic data, maintains relational integrity, and allows configurable data volumes. u... Use when working with databases or data models. Trigger with phrases like 'database', 'query', or 'schema'.

1033

cursor-codebase-indexing

jeremylongshore

Execute set up and optimize Cursor codebase indexing. Triggers on "cursor index setup", "codebase indexing", "index codebase", "cursor semantic search". Use when working with cursor codebase indexing functionality. Trigger with phrases like "cursor codebase indexing", "cursor indexing", "cursor".

885

testing-mobile-apps

jeremylongshore

Execute mobile app testing on iOS and Android devices/simulators. Use when performing specialized testing. Trigger with phrases like "test mobile app", "run iOS tests", or "validate Android functionality".

810

You might also like

observability-engineer

sickn33

Build production-ready monitoring, logging, and tracing systems. Implements comprehensive observability strategies, SLI/SLO management, and incident response workflows. Use PROACTIVELY for monitoring infrastructure, performance optimization, or production reliability.

1242

mlops-engineer

sickn33

Build comprehensive ML pipelines, experiment tracking, and model registries with MLflow, Kubeflow, and modern MLOps tools. Implements automated training, deployment, and monitoring across cloud platforms. Use PROACTIVELY for ML infrastructure, experiment management, or pipeline automation.

333

senior-devops

davila7

Comprehensive DevOps skill for CI/CD, infrastructure automation, containerization, and cloud platforms (AWS, GCP, Azure). Includes pipeline setup, infrastructure as code, deployment automation, and monitoring. Use when setting up pipelines, deploying applications, managing infrastructure, implementing monitoring, or optimizing deployment processes.

720

server-management

davila7

Server management principles and decision-making. Process management, monitoring strategy, and scaling decisions. Teaches thinking, not commands.

113

debug-cluster

openshift

Provides systematic debugging approaches for HyperShift hosted-cluster issues. Auto-applies when debugging cluster problems, investigating stuck deletions, or troubleshooting control plane issues.

25

domain-cloud-native

actionbook

Use when building cloud-native apps. Keywords: kubernetes, k8s, docker, container, grpc, tonic, microservice, service mesh, observability, tracing, metrics, health check, cloud, deployment, 云原生, 微服务, 容器

13

Search skills

Search the agent skills registry