OR

orchestrating-test-execution

Manages parallel test distribution, worker allocation, and result aggregation.

Install

mkdir -p .claude/skills/orchestrating-test-execution && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/5993" && unzip -o skill.zip -d .claude/skills/orchestrating-test-execution && rm skill.zip

Installs to .claude/skills/orchestrating-test-execution

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Test coordinate parallel test execution across multiple environments
68 charsno explicit “when” trigger
Advanced

Key capabilities

  • Coordinate parallel test execution
  • Manage test splitting across workers
  • Aggregate test results
  • Implement intelligent retry logic
  • Configure CI pipeline for parallel tests
  • Identify slowest tests for optimization

How it works

The skill analyzes test suites, classifies tests into execution tiers, configures parallel execution for each tier, and aggregates results into a single report.

Inputs & outputs

You give it
Test files, CI pipeline configuration
You get back
CI pipeline configuration file, merged test result report, execution timeline

When to use orchestrating-test-execution

  • Parallelizing unit tests
  • Sharding end-to-end browser tests
  • Managing multi-environment test runs
  • Optimizing CI test performance

About this skill

Test Orchestrator

Overview

Coordinate parallel test execution across multiple test suites, frameworks, and environments. Manages test splitting, worker allocation, result aggregation, and intelligent retry strategies.

Prerequisites

  • Test runner with parallel execution support (Jest, Vitest, pytest-xdist, Playwright, or JUnit 5)
  • CI/CD platform configured (GitHub Actions, GitLab CI, CircleCI, or Jenkins)
  • Test suite with consistent pass rates (flaky tests identified and tagged)
  • Sufficient CI runner resources for parallel worker count
  • Test result reporting tool (JUnit XML, Allure, or equivalent)

Instructions

  1. Analyze the existing test suite using Grep and Glob to catalog all test files, their framework, approximate run time, and dependency requirements.
  2. Classify tests into execution tiers:
    • Tier 1 (Fast): Unit tests with no I/O -- target under 30 seconds total.
    • Tier 2 (Medium): Integration tests requiring local services -- target under 3 minutes.
    • Tier 3 (Slow): E2E and browser tests -- target under 10 minutes.
  3. Configure parallel execution for each tier:
    • Split unit tests across N workers using jest --shard=i/N or pytest -n auto.
    • Shard E2E tests by test file using Playwright --shard=i/N or Cypress parallelization.
    • Assign heavier integration tests to dedicated workers with more resources.
  4. Create a CI pipeline configuration that runs tiers in parallel:
    • Tier 1 and Tier 2 run concurrently on separate jobs.
    • Tier 3 runs after a fast pre-check gate passes.
    • Each tier reports results to a unified collection step.
  5. Implement intelligent retry logic for flaky tests:
    • Tag known flaky tests with @flaky or equivalent marker.
    • Retry failed tests up to 2 times before marking as failed.
    • Track flaky test frequency in a log file for triage.
  6. Aggregate results from all parallel workers into a single report:
    • Merge JUnit XML files from each shard.
    • Calculate total pass/fail/skip counts and execution time.
    • Identify the slowest tests for optimization targets.
  7. Write the orchestration configuration to the project's CI config file and validate it with a dry run.

Output

  • CI pipeline configuration file (.github/workflows/test.yml, .gitlab-ci.yml, or equivalent)
  • Test sharding configuration with worker count and split strategy
  • Merged test result report in JUnit XML or JSON format
  • Execution timeline showing parallel job durations and bottlenecks
  • Flaky test inventory with retry counts and failure patterns

Error Handling

ErrorCauseSolution
Shard produces zero testsUneven test distribution or incorrect shard indexVerify shard count matches actual test file count; use file-based splitting
Worker out of memoryToo many parallel processes on one runnerReduce --maxWorkers or -n count; increase runner memory; use --workerIdleMemoryLimit
Test ordering dependencyTests pass in isolation but fail in specific shard orderAdd --randomize flag; fix shared state leaks; enforce test independence
Result aggregation mismatchMissing shard results due to job timeoutSet job-level timeouts higher than test timeouts; add result upload as a separate step
CI cache miss slowing startupDependencies not cached between parallel jobsConfigure dependency caching per lockfile hash; use a shared setup job

Examples

GitHub Actions matrix strategy for Jest sharding:

jobs:
  test:
    strategy:
      matrix:
        shard: [1, 2, 3, 4]
    steps:
      - run: npx jest --shard=${{ matrix.shard }}/4 --ci --reporters=jest-junit
      - uses: actions/upload-artifact@v4
        with:
          name: results-${{ matrix.shard }}
          path: junit.xml
  merge:
    needs: test
    steps:
      - uses: actions/download-artifact@v4
      - run: npx junit-merge -d results-* -o merged-results.xml

pytest-xdist parallel execution:

pytest -n auto --dist worksteal -q --junitxml=results.xml

Playwright sharded execution:

npx playwright test --shard=1/3 --reporter=junit

Resources

Prerequisites

Test runner with parallel execution supportCI/CD platform configuredTest suite with consistent pass ratesSufficient CI runner resources for parallel worker count

Limitations

  • Shard produces zero tests due to uneven distribution
  • Worker out of memory from too many parallel processes
  • Test ordering dependency causing failures

How it compares

This skill orchestrates parallel test execution across multiple frameworks and environments, managing test splitting and result aggregation, which differs from running tests sequentially or manually configuring individual test jobs.

Compared to similar skills

orchestrating-test-execution side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
orchestrating-test-execution (this skill)125dReviewAdvanced
e2e-testing-patterns82moNo flagsIntermediate
testing-workflow169moReviewIntermediate
perf-lighthouse135moReviewIntermediate

Try saying

Example prompts that trigger this skill in your AI assistant.

More by jeremylongshore

View all by jeremylongshore

analyzing-logs

jeremylongshore

Analyze application logs to detect performance issues, identify error patterns, and improve stability by extracting key insights.

14123

ollama-setup

jeremylongshore

Configure auto-configure Ollama when user needs local LLM deployment, free AI alternatives, or wants to eliminate hosted API costs. Trigger phrases: "install ollama", "local AI", "free LLM", "self-hosted AI", "replace OpenAI", "no API costs". Use when appropriate context detected. Trigger with relevant phrases based on skill purpose.

1167

backtesting-trading-strategies

jeremylongshore

Backtest crypto and traditional trading strategies against historical data. Calculates performance metrics (Sharpe, Sortino, max drawdown), generates equity curves, and optimizes strategy parameters. Use when user wants to test a trading strategy, validate signals, or compare approaches. Trigger with phrases like "backtest strategy", "test trading strategy", "historical performance", "simulate trades", "optimize parameters", or "validate signals".

1071

generating-database-seed-data

jeremylongshore

Process this skill enables AI assistant to generate realistic test data and database seed scripts for development and testing environments. it uses faker libraries to create realistic data, maintains relational integrity, and allows configurable data volumes. u... Use when working with databases or data models. Trigger with phrases like 'database', 'query', or 'schema'.

1033

cursor-codebase-indexing

jeremylongshore

Execute set up and optimize Cursor codebase indexing. Triggers on "cursor index setup", "codebase indexing", "index codebase", "cursor semantic search". Use when working with cursor codebase indexing functionality. Trigger with phrases like "cursor codebase indexing", "cursor indexing", "cursor".

885

testing-mobile-apps

jeremylongshore

Execute mobile app testing on iOS and Android devices/simulators. Use when performing specialized testing. Trigger with phrases like "test mobile app", "run iOS tests", or "validate Android functionality".

810

You might also like

Search skills

Search the agent skills registry