CO

code-change-verification

Runs the required test, lint, and formatting stack for the OpenAI Agents repository.

Install

mkdir -p .claude/skills/code-change-verification && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/943" && unzip -o skill.zip -d .claude/skills/code-change-verification && rm skill.zip

Installs to .claude/skills/code-change-verification

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Run the mandatory verification stack when changes affect runtime code, tests, or build/test behavior in the OpenAI Agents Python repository.
140 chars✓ has a “when” trigger
Beginner

Key capabilities

  • →Run automated formatting on codebase
  • →Execute linting to enforce quality standards
  • →Perform type checking to ensure code safety
  • →Run full test suite in parallel
  • →Provide heartbeat updates during verification

How it works

The skill executes a sequential pipeline of formatting, linting, type checking, and testing, utilizing parallel execution for the latter three steps with fail-fast behavior.

Inputs & outputs

You give it
Modified runtime code or test files
You get back
Verification report confirming code quality and test success

When to use code-change-verification

  • →Run project-wide linting
  • →Execute full test suite
  • →Verify type checking for code changes

About this skill

Code Change Verification

Overview

Ensure work is only marked complete after formatting, linting, type checking, and tests pass. Use this skill when changes affect runtime code, tests, or build/test configuration. You can skip it for docs-only or repository metadata unless a user asks for the full stack. This is a post-review final gate: when $implementation-final-review applies, do not invoke the broad stack until its clean-review condition applies to the stable task diff.

Quick start

  1. Keep this skill at ./.agents/skills/code-change-verification so it loads automatically for the repository.
  2. Codex on macOS/Linux: /usr/bin/env -u OPENAI_API_KEY OPENAI_AGENTS_TEST_IN_CODEX_SANDBOX=1 UV_DEFAULT_INDEX=https://pypi.org/simple bash .agents/skills/code-change-verification/scripts/run.sh.
  3. Other macOS/Linux environments: env UV_DEFAULT_INDEX=https://pypi.org/simple bash .agents/skills/code-change-verification/scripts/run.sh.
  4. Windows: powershell -ExecutionPolicy Bypass -File .agents/skills/code-change-verification/scripts/run.ps1.
  5. On macOS/Linux, the script runs make format, make lint, make typecheck, and make tests sequentially and stops at the first failure. Parallelism inside each Make target, including pytest workers, is unchanged.
  6. The Bash script streams each command's output directly. The Windows wrapper retains parallel lint, typecheck, and test steps with periodic heartbeat updates.
  7. If any command fails, fix the issue, rerun the script, and report the failing output.
  8. Confirm completion only when all commands succeed with no remaining issues.

Start condition and host capacity

  • During iterative review, use only focused tests and a narrowly targeted static check when the changed typing boundary requires one. Defer repository-wide make typecheck and the rest of this complete stack until review is clean.
  • Immediately before starting the complete stack, use available read-only task or process evidence to check whether another repository-wide test, typecheck, build, examples runner, or integration command is already active on the same host.
  • When concrete contention is visible, continue useful non-heavy work such as review, remediation, evidence preparation, or focused checks, then check again later. Do not create or wait on a repository lock, host-wide mutex, or sentinel file.
  • Start automatically once review is clean, the diff is stable, and observable host capacity is available. Do not require a user-triggered finalize message. If host telemetry is unavailable, do not block solely because capacity cannot be measured.

Codex execution policy

Repository verification and all child processes must remain in the normal Codex workspace sandbox. Never request elevated sandbox permissions for the verification wrapper, and never retry the wrapper with broader host access after a failure.

On macOS, tests marked requires_native_macos_sandbox need to start their own sandbox-exec process. The Codex command sets OPENAI_AGENTS_TEST_IN_CODEX_SANDBOX=1, which skips only that marker before nested sandbox creation. All other tests remain enabled. Ordinary local and CI runs do not set this variable and therefore keep the marked tests enabled.

The marked tests run separately on a disposable GitHub-hosted macOS runner. If that trusted runner is unavailable, report the missing native-macOS coverage; do not compensate by weakening the Codex sandbox boundary.

Environment setup

The verification scripts assume repository dependencies are already installed. Do not run make sync as part of every verification pass; use it for a fresh checkout, after dependency files change, or when dependency resolution fails before the checks start.

On Linux, some Python packages with native extensions may require system packages such as libffi-dev, Python development headers, or build tools. If verification cannot start because one of these packages is missing, treat it as a local environment setup issue. Install the missing dependency when possible, or report the failing command and missing dependency in the PR test plan before rerunning verification in a prepared environment.

Manual workflow

  • For a fresh checkout, or if dependencies are not installed or have changed, run make sync first to install dev requirements via uv.
  • Run from the repository root with make format first, then make lint, make typecheck, and make tests.
  • Do not skip steps; stop and fix issues immediately when a command fails.
  • Run the manual steps sequentially and stop at the first failure. Keep the parallelism provided by each Make target.
  • Re-run the full stack after applying fixes so the commands execute in the required order.

Resources

scripts/run.sh

  • Runs make format, make lint, make typecheck, and make tests sequentially from the repository root. It streams output, preserves the first failure or cancellation status, and cleans up the active step's process group before continuing or exiting.

scripts/run.ps1

  • Windows-friendly wrapper that runs the same sequence with make format first and the remaining steps in parallel with fail-fast semantics, plus periodic heartbeat updates while work is still running. Use from PowerShell with execution policy bypass if required by your environment.

When not to use it

  • →Do not use for documentation-only changes
  • →Do not use for repository metadata updates

Prerequisites

Repository dependencies installedmake utility

Limitations

  • →Requires specific system packages for native extensions
  • →Verification must be re-run in full after applying fixes

How it compares

It replaces manual, piecemeal verification with a standardized, automated pipeline that ensures all quality gates are met before completion.

Compared to similar skills

code-change-verification side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
code-change-verification (this skill)46moReviewBeginner
verification-loop06moReviewIntermediate
scholar-verify04moNo flagsIntermediate
python-testing-patterns774moReviewIntermediate

Try saying

Example prompts that trigger this skill in your AI assistant.

skill-installer

openai

Install Codex skills into $CODEX_HOME/skills from a curated list or a GitHub repo path. Use when a user asks to list installable skills, install a curated skill, or install a skill from another repo (including private repos).

29141

figma-implement-design

openai

Translate Figma nodes into production-ready code with 1:1 visual fidelity using the Figma MCP workflow (design context, screenshots, assets, and project-convention translation). Trigger when the user provides Figma URLs or node IDs, or asks to implement designs or components that must match Figma specs. Requires a working Figma MCP server connection.

2460

figma

openai

Use the Figma MCP server to fetch design context, screenshots, variables, and assets from Figma, and to translate Figma nodes into production code. Trigger when a task involves Figma URLs, node IDs, design-to-code implementation, or Figma MCP setup and troubleshooting.

2266

gh-fix-ci

openai

Use when a user asks to debug or fix failing GitHub PR checks that run in GitHub Actions; use `gh` to inspect checks and logs, summarize failure context, draft a fix plan, and implement only after explicit approval. Treat external providers (for example Buildkite) as out of scope and report only the details URL.

1234

transcribe

openai

Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.

1148

gh-address-comments

openai

Help address review/issue comments on the open GitHub PR for the current branch using gh CLI; verify gh auth first and prompt the user to authenticate if not logged in.

1059

Search skills

Search the agent skills registry