bdistill-behavioral-xray
Evaluates and reports on AI model behavior through automated red-teaming and diagnostic probing.
Install
mkdir -p .claude/skills/bdistill-behavioral-xray && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/13759" && unzip -o skill.zip -d .claude/skills/bdistill-behavioral-xray && rm skill.zipInstalls to .claude/skills/bdistill-behavioral-xray
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
X-ray any AI model's behavioral patterns — refusal boundaries, hallucination tendencies, reasoning style, formatting defaults. No API key needed.Key capabilities
- →Probe AI model for tool use behavior
- →Map AI model refusal boundaries
- →Analyze AI model formatting defaults
- →Evaluate AI model reasoning style
- →Assess AI model persona and tone
- →Measure AI model hallucination resistance
How it works
The skill runs 30 probe questions across six behavioral dimensions, auto-tags responses with metadata, and compiles the results into an HTML report with radar charts and insights.
Inputs & outputs
When to use bdistill-behavioral-xray
- →Test AI model for hallucinations
- →Map model refusal boundaries
- →Compare AI model behavior
- →Audit model compliance
About this skill
Behavioral X-Ray
Systematically probe an AI model's behavioral patterns and generate a visual report. The AI agent probes itself — no API key or external setup needed.
Overview
bdistill's Behavioral X-Ray runs 30 carefully designed probe questions across 6 dimensions, auto-tags each response with behavioral metadata, and compiles results into a styled HTML report with radar charts and actionable insights.
Use it to understand your model before building with it, compare models for task selection, or track behavioral drift over time.
When to Use This Skill
- Use when you want to understand how your AI model actually behaves (not how it claims to)
- Use when choosing between models for a specific task
- Use when debugging unexpected refusals, hallucinations, or formatting issues
- Use for compliance auditing — documenting model behavior at deployment boundaries
- Use for red team assessments — systematic boundary mapping across safety dimensions
How It Works
Step 1: Install
pip install bdistill
claude mcp add bdistill -- bdistill-mcp # Claude Code
For other tools, add bdistill-mcp as an MCP server in your project config.
Step 2: Run the probe
In Claude Code:
/xray # Full behavioral probe (30 questions)
/xray --dimensions refusal # Probe just one dimension
/xray-report # Generate report from completed probe
In any tool with MCP:
"X-ray your behavioral patterns"
"Test your refusal boundaries"
"Generate a behavioral report"
Probe Dimensions
| Dimension | What it measures |
|---|---|
| tool_use | When does it call tools vs. answer from knowledge? |
| refusal | Where does it draw safety boundaries? Does it over-refuse? |
| formatting | Lists vs. prose? Code blocks? Length calibration? |
| reasoning | Does it show chain-of-thought? Handle trick questions? |
| persona | Identity, tone matching, composure under hostility |
| grounding | Hallucination resistance, fabrication traps, knowledge limits |
Output
A styled HTML report showing:
- Refusal rate, hedge rate, chain-of-thought usage
- Per-dimension breakdown with bar charts
- Notable response examples with behavioral tags
- Actionable insights (e.g., "you already show CoT 85% of the time, no need to prompt for it")
Best Practices
- Answer probe questions honestly — the value is in authentic behavioral data
- Run probes on the same model periodically to track behavioral drift
- Compare reports across models to make informed selection decisions
- Use adversarial knowledge extraction (
/distill --adversarial) alongside behavioral probes for complete model profiling
Related Skills
@bdistill-knowledge-extraction- Extract structured domain knowledge from any AI model
When not to use it
- →When the goal is to train an AI model
- →When the task is not related to evaluating an AI model's behavioral patterns
Limitations
- →The skill measures behavioral patterns, not internal mechanisms.
- →The output is a report, not a direct modification of the AI model.
- →The probe questions are intended to reveal specific behavioral dimensions.
How it compares
This skill systematically probes and visualizes an AI model's behavioral patterns, providing a structured evaluation that differs from subjective assessment or ad-hoc testing.
Compared to similar skills
bdistill-behavioral-xray side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| bdistill-behavioral-xray (this skill) | 0 | 4mo | Review | Intermediate |
| backtesting-trading-strategies | 10 | 27d | Review | Intermediate |
| data-quality-frameworks | 6 | 2mo | Review | Intermediate |
| trulens-running-evaluations | 1 | 3mo | No flags | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by wegonbeok45
View all by wegonbeok45 →You might also like
backtesting-trading-strategies
jeremylongshore
Backtest crypto and traditional trading strategies against historical data. Calculates performance metrics (Sharpe, Sortino, max drawdown), generates equity curves, and optimizes strategy parameters. Use when user wants to test a trading strategy, validate signals, or compare approaches. Trigger with phrases like "backtest strategy", "test trading strategy", "historical performance", "simulate trades", "optimize parameters", or "validate signals".
data-quality-frameworks
wshobson
Implement data quality validation with Great Expectations, dbt tests, and data contracts. Use when building data quality pipelines, implementing validation rules, or establishing data contracts.
trulens-running-evaluations
truera
Execute TruLens evaluations and view results
test-reporting-analytics
proffesor-for-testing
Advanced test reporting, quality dashboards, predictive analytics, trend analysis, and executive reporting for QE metrics. Use when communicating quality status, tracking trends, or making data-driven decisions.
extract-test-set
tradingstrategy-ai
Extract raw price dataframe for a test case
detect-metrics
tidymodels
Detect and list all metric functions in the yardstick package. Use when a user asks to find, list, or identify all metrics in the package.