RE

Queries the Reactome database to perform biological pathway analysis and study molecular interactions.

Install

mkdir -p .claude/skills/reactome-database && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/4423" && unzip -o skill.zip -d .claude/skills/reactome-database && rm skill.zip

Installs to .claude/skills/reactome-database

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

Query Reactome REST API for pathway analysis, enrichment, gene-pathway mapping, disease pathways, molecular interactions, expression analysis, for systems biology studies.
171 charsno explicit “when” trigger
Advanced

Key capabilities

  • Query pathway information
  • Perform overrepresentation analysis
  • Analyze gene expression data
  • Map genes to pathways
  • Visualize analysis results

How it works

The skill interfaces with the Reactome REST API to perform computational analysis on biological data and retrieve pathway hierarchies.

Inputs & outputs

You give it
Gene list or expression data
You get back
Pathway analysis results

When to use reactome-database

  • Perform pathway enrichment analysis
  • Map genes to biological pathways
  • Analyze gene expression data
  • Identify disease-related pathways

About this skill

Reactome Database

Overview

Reactome is a free, open-source, curated pathway database with 2,825+ human pathways. Query biological pathways, perform overrepresentation and expression analysis, map genes to pathways, explore molecular interactions via REST API and Python client for systems biology research.

When to Use This Skill

This skill should be used when:

  • Performing pathway enrichment analysis on gene or protein lists
  • Analyzing gene expression data to identify relevant biological pathways
  • Querying specific pathway information, reactions, or molecular interactions
  • Mapping genes or proteins to biological pathways and processes
  • Exploring disease-related pathways and mechanisms
  • Visualizing analysis results in the Reactome Pathway Browser
  • Conducting comparative pathway analysis across species

Core Capabilities

Reactome provides two main API services and a Python client library:

1. Content Service - Data Retrieval

Query and retrieve biological pathway data, molecular interactions, and entity information.

Common operations:

  • Retrieve pathway information and hierarchies
  • Query specific entities (proteins, reactions, complexes)
  • Get participating molecules in pathways
  • Access database version and metadata
  • Explore pathway compartments and locations

API Base URL: https://reactome.org/ContentService

2. Analysis Service - Pathway Analysis

Perform computational analysis on gene lists and expression data.

Analysis types:

  • Overrepresentation Analysis: Identify statistically significant pathways from gene/protein lists
  • Expression Data Analysis: Analyze gene expression datasets to find relevant pathways
  • Species Comparison: Compare pathway data across different organisms

API Base URL: https://reactome.org/AnalysisService

3. reactome2py Python Package

Python client library that wraps Reactome API calls for easier programmatic access.

Installation:

uv pip install reactome2py

Note: The reactome2py package (version 3.0.0, released January 2021) is functional but not actively maintained. For the most up-to-date functionality, consider using direct REST API calls.

Querying Pathway Data

Using Content Service REST API

The Content Service uses REST protocol and returns data in JSON or plain text formats.

Get database version:

import requests

response = requests.get("https://reactome.org/ContentService/data/database/version")
version = response.text
print(f"Reactome version: {version}")

Query a specific entity:

import requests

entity_id = "R-HSA-69278"  # Example pathway ID
response = requests.get(f"https://reactome.org/ContentService/data/query/{entity_id}")
data = response.json()

Get participating molecules in a pathway:

import requests

event_id = "R-HSA-69278"
response = requests.get(
    f"https://reactome.org/ContentService/data/event/{event_id}/participatingPhysicalEntities"
)
molecules = response.json()

Using reactome2py Package

import reactome2py
from reactome2py import content

# Query pathway information
pathway_info = content.query_by_id("R-HSA-69278")

# Get database version
version = content.get_database_version()

For detailed API endpoints and parameters, refer to references/api_reference.md in this skill.

Performing Pathway Analysis

Overrepresentation Analysis

Submit a list of gene/protein identifiers to find enriched pathways.

Using REST API:

import requests

# Prepare identifier list
identifiers = ["TP53", "BRCA1", "EGFR", "MYC"]
data = "\n".join(identifiers)

# Submit analysis
response = requests.post(
    "https://reactome.org/AnalysisService/identifiers/",
    headers={"Content-Type": "text/plain"},
    data=data
)

result = response.json()
token = result["summary"]["token"]  # Save token to retrieve results later

# Access pathways
for pathway in result["pathways"]:
    print(f"{pathway['stId']}: {pathway['name']} (p-value: {pathway['entities']['pValue']})")

Retrieve analysis by token:

# Token is valid for 7 days
response = requests.get(f"https://reactome.org/AnalysisService/token/{token}")
results = response.json()

Expression Data Analysis

Analyze gene expression datasets with quantitative values.

Input format (TSV with header starting with #):

#Gene	Sample1	Sample2	Sample3
TP53	2.5	3.1	2.8
BRCA1	1.2	1.5	1.3
EGFR	4.5	4.2	4.8

Submit expression data:

import requests

# Read TSV file
with open("expression_data.tsv", "r") as f:
    data = f.read()

response = requests.post(
    "https://reactome.org/AnalysisService/identifiers/",
    headers={"Content-Type": "text/plain"},
    data=data
)

result = response.json()

Species Projection

Map identifiers to human pathways exclusively using the /projection/ endpoint:

response = requests.post(
    "https://reactome.org/AnalysisService/identifiers/projection/",
    headers={"Content-Type": "text/plain"},
    data=data
)

Visualizing Results

Analysis results can be visualized in the Reactome Pathway Browser by constructing URLs with the analysis token:

token = result["summary"]["token"]
pathway_id = "R-HSA-69278"
url = f"https://reactome.org/PathwayBrowser/#{pathway_id}&DTAB=AN&ANALYSIS={token}"
print(f"View results: {url}")

Working with Analysis Tokens

  • Analysis tokens are valid for 7 days
  • Tokens allow retrieval of previously computed results without re-submission
  • Store tokens to access results across sessions
  • Use GET /token/{TOKEN} endpoint to retrieve results

Data Formats and Identifiers

Supported Identifier Types

Reactome accepts various identifier formats:

  • UniProt accessions (e.g., P04637)
  • Gene symbols (e.g., TP53)
  • Ensembl IDs (e.g., ENSG00000141510)
  • EntrezGene IDs (e.g., 7157)
  • ChEBI IDs for small molecules

The system automatically detects identifier types.

Input Format Requirements

For overrepresentation analysis:

  • Plain text list of identifiers (one per line)
  • OR single column in TSV format

For expression analysis:

  • TSV format with mandatory header row starting with "#"
  • Column 1: identifiers
  • Columns 2+: numeric expression values
  • Use period (.) as decimal separator

Output Format

All API responses return JSON containing:

  • pathways: Array of enriched pathways with statistical metrics
  • summary: Analysis metadata and token
  • entities: Matched and unmapped identifiers
  • Statistical values: pValue, FDR (false discovery rate)

Helper Scripts

This skill includes scripts/reactome_query.py, a helper script for common Reactome operations:

# Query pathway information
python scripts/reactome_query.py query R-HSA-69278

# Perform overrepresentation analysis
python scripts/reactome_query.py analyze gene_list.txt

# Get database version
python scripts/reactome_query.py version

Additional Resources

For comprehensive API endpoint documentation, see references/api_reference.md in this skill.

Current Database Statistics (Version 94, September 2025)

  • 2,825 human pathways
  • 16,002 reactions
  • 11,630 proteins
  • 2,176 small molecules
  • 1,070 drugs
  • 41,373 literature references

When not to use it

  • When the data is not related to biological pathways
  • When real-time interaction with the database is not possible

Prerequisites

reactome2py

Limitations

  • Requires internet access
  • Analysis tokens expire after 7 days

How it compares

It provides programmatic access to curated biological data, replacing manual database searches.

Compared to similar skills

reactome-database side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
reactome-database (this skill)17moReviewAdvanced
kegg-database17moReviewIntermediate
ena-database27moReviewAdvanced
metabolomics-workbench-database17moReviewIntermediate

Try saying

Example prompts that trigger this skill in your AI assistant.

software-architecture

davila7

Guide for quality focused software architecture. This skill should be used when users want to write code, design architecture, analyze code, in any case that relates to software development.

333868

planning-with-files

davila7

Implements Manus-style file-based planning for complex tasks. Creates task_plan.md, findings.md, and progress.md. Use when starting complex multi-step tasks, research projects, or any task requiring >5 tool calls.

233106

telegram-bot-builder

davila7

Expert in building Telegram bots that solve real problems - from simple automation to complex AI-powered bots. Covers bot architecture, the Telegram Bot API, user experience, monetization strategies, and scaling bots to thousands of users. Use when: telegram bot, bot api, telegram automation, chat bot telegram, tg bot.

106130

scroll-experience

davila7

Expert in building immersive scroll-driven experiences - parallax storytelling, scroll animations, interactive narratives, and cinematic web experiences. Like NY Times interactives, Apple product pages, and award-winning web experiences. Makes websites feel like experiences, not just pages. Use when: scroll animation, parallax, scroll storytelling, interactive story, cinematic website.

101142

humanizer

davila7

Remove signs of AI-generated writing from text. Use when editing or reviewing text to make it sound more natural and human-written. Based on Wikipedia's comprehensive "Signs of AI writing" guide. Detects and fixes patterns including: inflated symbolism, promotional language, superficial -ing analyses, vague attributions, em dash overuse, rule of three, AI vocabulary words, negative parallelisms, and excessive conjunctive phrases. Credits: Original skill by @blader - https://github.com/blader/humanizer

90175

game-development

davila7

Game development orchestrator. Routes to platform-specific skills based on project needs.

70195

You might also like

kegg-database

davila7

Direct REST API access to KEGG (academic use only). Pathway analysis, gene-pathway mapping, metabolic pathways, drug interactions, ID conversion. For Python workflows with multiple databases, prefer bioservices. Use this for direct HTTP/REST work or KEGG-specific control.

11

ena-database

davila7

Access European Nucleotide Archive via API/FTP. Retrieve DNA/RNA sequences, raw reads (FASTQ), genome assemblies by accession, for genomics and bioinformatics pipelines. Supports multiple formats.

27

metabolomics-workbench-database

davila7

Access NIH Metabolomics Workbench via REST API (4,200+ studies). Query metabolites, RefMet nomenclature, MS/NMR data, m/z searches, study metadata, for metabolomics and biomarker discovery.

13

biopython

davila7

Primary Python toolkit for molecular biology. Preferred for Python-based PubMed/NCBI queries (Bio.Entrez), sequence manipulation, file parsing (FASTA, GenBank, FASTQ, PDB), advanced BLAST workflows, structures, phylogenetics. For quick BLAST, use gget. For direct REST API, use pubmed-database.

10

uniprot-database

davila7

Direct REST API access to UniProt. Protein searches, FASTA retrieval, ID mapping, Swiss-Prot/TrEMBL. For Python workflows with multiple databases, prefer bioservices (unified interface to 40+ services). Use this for direct HTTP/REST work or UniProt-specific control.

10

exploratory-data-analysis

K-Dense-AI

Perform comprehensive exploratory data analysis on scientific data files across 200+ file formats. This skill should be used when analyzing any scientific data file to understand its structure, content, quality, and characteristics. Automatically detects file type and generates detailed markdown reports with format-specific analysis, quality metrics, and downstream analysis recommendations. Covers chemistry, bioinformatics, microscopy, spectroscopy, proteomics, metabolomics, and general scientific data formats.

15114

Search skills

Search the agent skills registry