groq-common-errors
Provides diagnostic procedures for Groq API errors, focusing on status codes and structured error responses.
Install
mkdir -p .claude/skills/groq-common-errors && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/5419" && unzip -o skill.zip -d .claude/skills/groq-common-errors && rm skill.zipInstalls to .claude/skills/groq-common-errors
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Diagnose and fix Groq API errors with real error codes and solutions.Key capabilities
- →Capture failing HTTP status and body
- →Confirm API key validity
- →Confirm model existence and deprecation status
- →Map HTTP status to a fix
- →Handle SDK-level typed exception classes
- →Diagnose Groq API errors
How it works
The skill analyzes Groq API error responses by examining HTTP status codes and structured error bodies. It provides steps to verify API keys, check model availability, and map specific error types to documented solutions.
Inputs & outputs
When to use groq-common-errors
- →Validate API keys and models
- →Debug rate limit exceeded errors
- →Interpret structured error payloads
- →Verify service connection health
About this skill
Groq Common Errors
Overview
Comprehensive reference for Groq API error codes, their root causes, and proven fixes. Groq returns standard HTTP status codes with structured error bodies and rate-limit headers. This skill walks the diagnosis from raw error string to fix, then hands off to the full per-status reference for depth.
Every Groq error body follows one shape — read the code and type first:
{
"error": {
"message": "Rate limit reached for model `llama-3.3-70b-versatile`...",
"type": "tokens",
"code": "rate_limit_exceeded"
}
}
Prerequisites
GROQ_API_KEYexported in the environment (keys start withgsk_).curlandjqavailable for the diagnostic probes below.- For SDK-level handling:
groq-sdk(TypeScript) orgroq(Python) installed.
Instructions
-
Capture the failing status and body. Read the raw error response — the HTTP status plus the
code/typefields determine the whole diagnosis path. -
Confirm the key works before assuming anything deeper:
set -euo pipefail # Verify API key is valid — expect a model count, not an auth error curl -s https://api.groq.com/openai/v1/models \ -H "Authorization: Bearer $GROQ_API_KEY" | jq '.data | length' -
Confirm the model still exists. Many 400s are deprecated model IDs — list the live models and Grep your codebase for any stale ID:
curl -s https://api.groq.com/openai/v1/models \ -H "Authorization: Bearer $GROQ_API_KEY" | jq '.data[].id' | sort -
Map the status to a fix using the table below, then drill into references/error-reference.md for the exact error string, causes, and copy-paste fix.
-
For SDK integrations, branch on the typed exception classes — see references/sdk-error-handling.md.
Output
A diagnosis that names the error class, the root cause, and the concrete fix — for example: "429 on TPM: token budget exhausted; add the single-retry handleRateLimit wrapper and honor retry-after," or "400: mixtral-8x7b-32768 is deprecated; switch to llama-3.3-70b-versatile." When run against real code, the output is the edited call site plus a verification curl that returns 200.
Error Handling
Map the HTTP status to its cause; full error strings, rate-limit headers, and fixes live in references/error-reference.md.
| Status | Meaning | First move |
|---|---|---|
| 401 | Invalid / missing key | Confirm GROQ_API_KEY starts with gsk_; test with /models |
| 429 | RPM / TPM / RPD limit hit | Read retry-after; back off and single-retry |
| 400 | Deprecated model or bad params | List live models; replace stale IDs |
| 413 | Request over context window | Trim prompt (Llama models cap at 128K tokens) |
| 500 / 503 | Groq-side outage or overload | Retry with backoff; fall back model; check status page |
When the failure is transient (429/500/503), retry with backoff and honor retry-after rather than hammering. When it is structural (401/400/413), fix the request — retrying will not help.
Examples
Minimal end-to-end probe that isolates auth vs. model vs. payload problems:
# A 200 here means key + model + payload are all valid; a non-200 status
# tells you which layer failed.
curl -s -o /dev/null -w "%{http_code}" \
https://api.groq.com/openai/v1/chat/completions \
-H "Authorization: Bearer $GROQ_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"llama-3.1-8b-instant","messages":[{"role":"user","content":"ping"}],"max_tokens":5}'
- Rate-limit retry wrapper (TypeScript, honors
retry-after): see the 429 section of references/error-reference.md. - Typed SDK exception branching (TypeScript + Python): references/sdk-error-handling.md.
Resources
- Groq Error Codes
- Groq Rate Limits
- Groq Model Deprecations
- Groq Status Page
- Full per-status reference: references/error-reference.md
- SDK error handling + escalation path: references/sdk-error-handling.md
- For comprehensive debugging, see the
groq-debug-bundleskill.
When not to use it
- →When debugging non-Groq API errors
- →When the issue is not related to API communication
- →When Groq integration is not active
Prerequisites
Limitations
- →Retrying will not help for structural errors like 401, 400, or 413
- →Llama models cap at 128K tokens for request over context window
- →Transient errors (429/500/503) require backoff and honoring retry-after
How it compares
This skill provides a systematic diagnostic path for Groq API errors, including specific commands and SDK handling, which is different from general error troubleshooting.
Compared to similar skills
groq-common-errors side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| groq-common-errors (this skill) | 1 | 25d | Review | Intermediate |
| apollo-common-errors | 1 | 25d | Caution | Intermediate |
| mcp-builder | 136 | 3mo | Review | Advanced |
| telegram-bot-builder | 106 | 6mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by jeremylongshore
View all by jeremylongshore →You might also like
apollo-common-errors
jeremylongshore
Diagnose and fix common Apollo.io API errors. Use when encountering Apollo API errors, debugging integration issues, or troubleshooting failed requests. Trigger with phrases like "apollo error", "apollo api error", "debug apollo", "apollo 401", "apollo 429", "apollo troubleshoot".
mcp-builder
anthropics
Guide for creating high-quality MCP (Model Context Protocol) servers that enable LLMs to interact with external services through well-designed tools. Use when building MCP servers to integrate external APIs or services, whether in Python (FastMCP) or Node/TypeScript (MCP SDK).
telegram-bot-builder
davila7
Expert in building Telegram bots that solve real problems - from simple automation to complex AI-powered bots. Covers bot architecture, the Telegram Bot API, user experience, monetization strategies, and scaling bots to thousands of users. Use when: telegram bot, bot api, telegram automation, chat bot telegram, tg bot.
stripe-integration
wshobson
Implement Stripe payment processing for robust, PCI-compliant payment flows including checkout, subscriptions, and webhooks. Use when integrating Stripe payments, building subscription systems, or implementing secure checkout flows.
langchain-architecture
wshobson
Design LLM applications using the LangChain framework with agents, memory, and tool integration patterns. Use when building LangChain applications, implementing AI agents, or creating complex LLM workflows.
copilot-sdk
github
Build agentic applications with GitHub Copilot SDK. Use when embedding AI agents in apps, creating custom tools, implementing streaming responses, managing sessions, connecting to MCP servers, or creating custom agents. Triggers on Copilot SDK, GitHub SDK, agentic app, embed Copilot, programmable agent, MCP server, custom agent.