TO

Estimates and tracks token consumption to prevent context overflow errors during LLM operations.

Install

mkdir -p .claude/skills/token-budget && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/1401" && unzip -o skill.zip -d .claude/skills/token-budget && rm skill.zip

Installs to .claude/skills/token-budget

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Manages token budget estimation and tracking to prevent context overflow
72 charsno explicit “when” trigger
Intermediate

Key capabilities

  • Estimate token usage based on content type
  • Track cumulative context consumption
  • Categorize files by size for loading strategy
  • Trigger alerts at budget thresholds
  • Recommend compression and summarization

How it works

It applies heuristic rules to estimate token counts based on line numbers and content types, then monitors these against defined percentage thresholds.

Inputs & outputs

You give it
File list or content
You get back
Token usage estimate and budget status

When to use token-budget

  • Estimate tokens for prompt engineering
  • Monitor context usage in long tasks
  • Prevent model context window overflow

About this skill

Token Budget Skill

<role> You are a token-efficient agent. Your job is to maximize output quality while minimizing token consumption.

Core principle: Every token counts. Load only what you need, when you need it. </role>


Token Estimation

Quick Estimates

Content TypeTokens/LineNotes
Code~4-6Depends on verbosity
Markdown~3-4Less dense than code
JSON/YAML~5-7Structured, repetitive
Comments~3-4Natural language

Rule of thumb: tokens ≈ lines × 4

File Size Categories

CategoryLinesEst. TokensAction
Small<50<200Load freely
Medium50-200200-800Consider outline first
Large200-500800-2000Use search + snippets
Huge500+2000+Never load fully

Budget Thresholds

Based on PROJECT_RULES.md context quality thresholds:

UsageQualityBudget Status
0-30%PEAK✅ Proceed freely
30-50%GOOD⚠️ Be selective
50-70%DEGRADING🔶 Compress & summarize
70%+POOR🛑 State dump required

Budget Tracking Protocol

Before Each Task

  1. Estimate current usage:

    • Count files in context
    • Estimate tokens per file
    • Calculate approximate %
  2. Check budget status:

    Current: ~X,000 tokens (~Y%)
    Budget: [PEAK|GOOD|DEGRADING|POOR]
    
  3. Adjust strategy:

    • PEAK: Proceed normally
    • GOOD: Prefer search-first
    • DEGRADING: Use outlines only
    • POOR: Trigger state dump

During Execution

Track cumulative context:

## Token Tracker

| Phase | Files Loaded | Est. Tokens | Cumulative |
|-------|--------------|-------------|------------|
| Start | 0 | 0 | 0 |
| Task 1 | 2 | ~400 | ~400 |
| Task 2 | 3 | ~600 | ~1000 |

Optimization Strategies

1. Progressive Loading

Level 1: Outline only (function signatures)
Level 2: + Key functions (based on task)
Level 3: + Related code (if needed)
Level 4: Full file (only if essential)

2. Just-In-Time Loading

  • Load file only when task requires it
  • Unload mentally after task complete
  • Don't preload "just in case"

3. Search Before Load

Always use context-fetch skill first:

  1. Search for relevant terms
  2. Identify candidate files
  3. Load only needed sections

4. Summarize & Compress

After understanding a file:

  • Document key insights in STATE.md
  • Reference summary instead of re-reading
  • Use "I've analyzed X, it does Y" pattern

Budget Alerts

At 50% Budget

⚠️ TOKEN BUDGET: 50%
Switching to efficiency mode:
- Outlines only for new files
- Summarizing instead of loading
- Recommending compression

At 70% Budget

🛑 TOKEN BUDGET: 70%
Quality degradation likely. Recommend:
1. Create state snapshot
2. Run /pause
3. Continue in fresh session

Integration

This skill integrates with:

  • context-fetch — Search before loading
  • context-health-monitor — Quality tracking
  • context-compressor — Compression strategies
  • /pause and /resume — Session handoff

Anti-Patterns

Loading files "for context" — Search first ❌ Re-reading same file — Summarize once ❌ Full file when snippet suffices — Target load ❌ Ignoring budget warnings — Quality will degrade


Part of GSD v1.6 Token Optimization. See PROJECT_RULES.md for efficiency rules.

When not to use it

  • When working in environments without context limits

Limitations

  • Estimates are heuristic and not exact token counts
  • Relies on manual tracking during execution
  • Effectiveness depends on adherence to loading strategies

How it compares

It provides a structured protocol for proactive context management rather than reacting only after hitting a hard limit.

Compared to similar skills

token-budget side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
token-budget (this skill)44moNo flagsIntermediate
sequential-thinking1369moNo flagsIntermediate
ai-wrapper-product56moNo flagsIntermediate
initializing-memory22moReviewBeginner

Try saying

Example prompts that trigger this skill in your AI assistant.

Search skills

Search the agent skills registry