mims-harvard/tooluniverse

tooluniverse-spatial-omics-analysis

Spatial multi-omics interpretation pipeline.

First seen Feb 19, 2026

Installation

$ npx skills add mims-harvard/tooluniverse --skill tooluniverse-spatial-omics-analysis

Summary

  • Spatial multi-omics interpretation pipeline.
  • Transforms spatially variable genes (SVGs), domain annotations, and tissue context into biological insights via domain-by-domain characterization, cell-type composition, spatial gene expression patterns, RNA+protein+metabolite integration.
  • Use for Visium, MERFISH, seqFISH, Slide-seq, spatial proteomics, and spatial multi-omics interpretation.
  • Goes beyond statistics to disease mechanisms and therapeutic opportunities.

Also in this package

Other skills from mims-harvard/tooluniverse · top by installs.

npx skills add mims-harvard/tooluniverse

Browse all from mims-harvard/tooluniverse

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Repository health

Stars 1.7K
License LICENSE
Default branch main
Open issues 9
Status Active

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 12,644 B
  • docs SUMMARY.md 505 B

History

  1. First seen on skills.sh
  2. First recorded snapshot · 345 installs

SKILL.md

Spatial Multi-Omics Analysis Pipeline

Comprehensive biological interpretation of spatial omics data. Transforms spatially variable genes (SVGs), domain annotations, and tissue context into actionable biological insights.

KEY PRINCIPLES:

  1. Report-first approach - Create report file FIRST, then populate progressively
  2. Domain-by-domain analysis - Characterize each spatial region independently before comparison
  3. Gene-list-centric - Analyze user-provided SVGs and marker genes with ToolUniverse databases
  4. Biological interpretation - Go beyond statistics to explain biological meaning of spatial patterns
  5. Disease focus - Emphasize disease mechanisms and therapeutic opportunities when disease context is provided
  6. Evidence grading - Grade all evidence as T1 (human/clinical) to T4 (computational)
  7. Multi-modal thinking - Integrate RNA, protein, and metabolite information when available
  8. Validation guidance - Suggest experimental validation approaches for key findings
  9. Source references - Every statement must cite tool/database source
  10. English-first queries - Always use English terms in tool calls

LOOK UP, DON'T GUESS

When uncertain about any scientific fact, SEARCH databases first (PubMed, UniProt, ChEMBL, ClinVar, etc.) rather than reasoning from memory. A database-verified answer is always more reliable than a guess.


COMPUTE, DON'T DESCRIBE

When analysis requires computation (statistics, data processing, scoring, enrichment), write and run Python code via Bash. Don't describe what you would do — execute it and report actual results. Use ToolUniverse tools to retrieve data, then Python (pandas, scipy, statsmodels, matplotlib) to analyze it.

When to Use This Skill

Apply when users:

  • Provide spatially variable genes from spatial transcriptomics experiments
  • Ask about biological interpretation of spatial domains/clusters
  • Need pathway enrichment of spatial gene expression data
  • Want to understand cell-cell interactions from spatial data
  • Ask about tumor microenvironment heterogeneity from spatial omics
  • Need druggable targets in specific spatial regions
  • Ask about tissue zonation patterns (liver, brain, kidney)
  • Want to integrate spatial transcriptomics + proteomics data

NOT for: Single gene interpretation (use target-research), variant interpretation, drug safety, bulk RNA-seq, GWAS analysis.


Input Parameters

Parameter Required Description Example
svgs Yes Spatially variable genes ['EGFR', 'CDH1', 'VIM', 'MYC', 'CD3E']
tissue_type Yes Tissue/organ type brain, liver, lung, breast
technology No Spatial omics platform 10x Visium, MERFISH, DBiTplus
disease_context No Disease if applicable breast cancer, Alzheimer disease
spatial_domains No Domain -> marker genes dict {'Tumor core': ['MYC','EGFR']}
cell_types No Cell types from deconvolution ['Epithelial', 'T cell']
proteins No Proteins detected (multi-modal) ['CD3', 'PD-L1', 'Ki67']
metabolites No Metabolites (SpatialMETA) ['glutamine', 'lactate']

Spatial Omics Integration Score (0-100)

Data Completeness (0-30): SVGs (5), Disease context (5), Spatial domains (5), Cell types (5), Multi-modal (5), Literature (5)

Biological Insight (0-40): Pathway enrichment FDR<0.05 (10), Cell-cell interactions (10), Disease mechanism (10), Druggable targets (10)

Evidence Quality (0-30): Cross-database validation 3+ DBs (10), Clinical validation (10), Literature support (10)

Score Tier Interpretation
80-100 Excellent Comprehensive characterization, strong insights, druggable targets
60-79 Good Good pathway/interaction analysis, some therapeutic context
40-59 Moderate Basic enrichment, limited domain comparison
0-39 Limited Minimal data, gene-level annotation only

Evidence Grading

Tier Criteria Examples
[T1] Direct human/clinical evidence FDA-approved drug, validated biomarker
[T2] Experimental evidence Validated spatial pattern, known L-R pair
[T3] Computational/database evidence PPI prediction, pathway enrichment
[T4] Annotation/prediction only GO annotation, text-mined association

Analysis Phases Overview

Phase 0: Input Processing & Disambiguation (ALWAYS FIRST)

Resolve tissue/disease identifiers, establish analysis context. Get MONDO/EFO IDs for disease queries.

  • Tools: OpenTargetsgetdiseaseiddescriptionbyname, OpenTargetsgetdiseasedescriptionbyefoId, HPAsearchgenesby_query

Phase 1: Gene Characterization

Resolve gene IDs, annotate functions, tissue specificity, subcellular localization.

  • Tools: MyGenequerygenes, UniProtgetfunctionbyaccession, HPAgetsubcellularlocation, HPAgetrnaexpressionbysource, HPAgetcomprehensivegenedetailsbyensemblid, HPAgetcancerprognosticsbygene, UniProtIDMapgeneto_uniprot

Phase 2: Pathway & Functional Enrichment

Identify enriched pathways globally and per-domain. Filter FDR < 0.05.

  • Tools: STRINGfunctionalenrichment (PRIMARY), ReactomeAnalysispathwayenrichment, GOgetannotationsforgene, keggsearchpathway, WikiPathways_search

Phase 3: Spatial Domain Characterization

Characterize each domain biologically, assign cell types from markers, compare domains.

  • Tools: Phase 2 tools + HPAgetbiologicalprocessesbygene, HPAgetproteininteractionsbygene

Phase 4: Cell-Cell Interaction Inference

Predict communication from spatial patterns. Check ligand-receptor pairs across domains.

  • Tools: STRINGgetinteractionpartners, STRINGgetproteininteractions, intactsearchinteractions, Reactomegetinteractor, DGIdbgetdruggeneinteractions

Phase 5: Disease & Therapeutic Context

Connect to disease mechanisms, identify druggable targets, find clinical trials.

  • Tools: OpenTargetsgetassociatedtargetsbydiseaseefoId, OpenTargetsgettargettractabilitybyensemblID, OpenTargetsgetassociateddrugsbytargetensemblID, searchclinicaltrials, DGIdbgetgenedruggability, civicsearchgenes

Phase 6: Multi-Modal Integration

Integrate protein/RNA/metabolite data. Compare spatial RNA with protein detection.

  • Tools: HPAgetsubcellularlocation, HPAgetrnaexpressioninspecifictissues, Reactomemapuniprottopathways, kegggetpathwayinfo

Phase 7: Immune Microenvironment (Cancer/Inflammation only)

Classify immune cells, check checkpoint expression, assess Hot vs Cold vs Excluded patterns.

  • Tools: STRINGfunctionalenrichment, OpenTargetsgettargettractabilitybyensemblID, iedbsearch_epitopes

Phase 8: Literature & Validation Context

Search published evidence, suggest validation experiments (smFISH, IHC, PLA).

  • Tools: PubMedsearcharticles, openalexliteraturesearch

Data Discovery: HuBMAP Spatial Atlas Tools

Use HuBMAP tools to find published spatial biology reference datasets for comparison, validation, or cross-study analysis.

Tool Purpose Key Parameters
HuBMAPsearchdatasets Search published spatial datasets by organ/assay/keyword organ (code: "LK"=Kidney, "BR"=Brain, "LU"=Lung, etc.), dataset_type ("RNAseq", "CODEX", "MALDI"), query, limit
HuBMAPlistorgans List all available organs with codes and UBERON IDs (no required params)
HuBMAPgetdataset Get detailed metadata for a specific HuBMAP dataset hubmap_id (e.g. "HBM626.FHJD.938")

When to use: Phase 0 (find reference datasets for the tissue), Phase 8 (cross-reference findings with published HuBMAP atlas data).

See phase-procedures.md for detailed workflows, decision logic, and tool parameter specifications per phase.


Report Structure

Create file: {tissue}{disease}spatialomicsreport.md

# Spatial Multi-Omics Analysis Report: {Tissue Type}
**Report Generated**: {date} | **Technology**: {platform}
**Tissue**: {tissue_type} | **Disease**: {disease or "Normal tissue"}
**Total SVGs**: {count} | **Spatial Domains**: {count}
**Spatial Omics Integration Score**: (calculated after analysis)

## Executive Summary
## 1. Tissue & Disease Context
## 2. Spatially Variable Gene Characterization
  - 2.1 Gene ID Resolution
  - 2.2 Tissue Expression Patterns
  - 2.3 Subcellular Localization
  - 2.4 Disease Associations
## 3. Pathway Enrichment Analysis
  - 3.1 STRING, 3.2 Reactome, 3.3-3.5 GO (BP, MF, CC)
## 4. Spatial Domain Characterization (per-domain + comparison)
## 5. Cell-Cell Interaction Inference
  - 5.1 PPI, 5.2 Ligand-Receptor, 5.3 Signaling Pathways
## 6. Disease & Therapeutic Context
  - 6.1 Disease Gene Overlap, 6.2 Druggable Targets, 6.3 Drug Mechanisms, 6.4 Trials
## 7. Multi-Modal Integration (if data available)
## 8. Immune Microenvironment (if relevant)
## 9. Literature & Validation Context
## Spatial Omics Integration Score (breakdown table)
## Completeness Checklist
## References (tools used, database versions)

See report-template.md for full template with table structures.


Completeness Checklist

  • Gene ID resolution complete
  • Tissue expression patterns analyzed (HPA)
  • Subcellular localization checked (HPA)
  • Pathway enrichment complete (STRING + Reactome)
  • GO enrichment complete (BP + MF + CC)
  • Spatial domains characterized individually
  • Domain comparison performed
  • PPI analyzed (STRING)
  • Ligand-receptor pairs identified
  • Disease associations checked (OpenTargets)
  • Druggable targets identified
  • Multi-modal integration performed (if data available)
  • Immune microenvironment characterized (if relevant)
  • Literature search completed
  • Validation recommendations provided
  • Integration Score calculated
  • Executive summary written
  • All sections have source citations

Common Use Cases

  1. Cancer Spatial Heterogeneity: Visium with tumor/stroma/immune domains -> pathways, immune infiltration, druggable targets, checkpoints
  2. Brain Tissue Zonation: MERFISH with neuronal subtypes -> synaptic signaling, receptors, hippocampal zonation
  3. Liver Metabolic Zonation: Periportal vs pericentral -> CYP450, Wnt gradient, drug metabolism enzymes
  4. Tumor-Immune Interface: DBiTplus RNA+protein -> checkpoint L-R pairs, immune exclusion, multi-modal concordance
  5. Developmental Patterns: Morphogen gradients (Wnt, BMP, FGF, SHH), TF patterns, cell fate genes
  6. Disease Progression: Disease gradient -> inflammatory response, neuronal loss, therapeutic windows

Reference Files

  • phase-procedures.md - Detailed phase workflows, decision logic, tool usage per phase
  • tool-reference.md - Tool parameter names, response formats, fallback strategies, limitations
  • reference-data.md - Cell type markers, ligand-receptor pairs, immune checkpoint reference
  • report-template.md - Full report template with all table structures
  • testspatialomics.py - Test suite

Summary

Spatial Multi-Omics Analysis provides:

  1. Gene characterization (ID resolution, function, localization, tissue expression)
  2. Pathway & functional enrichment (STRING, Reactome, GO, KEGG)
  3. Spatial domain characterization (per-domain and cross-domain)
  4. Cell-cell interaction inference (PPI, ligand-receptor, signaling)
  5. Disease & therapeutic context (disease genes, druggable targets, trials)
  6. Multi-modal integration (RNA-protein concordance, metabolic pathways)
  7. Immune microenvironment (cell types, checkpoints, immunotherapy)
  8. Literature context & validation recommendations

Outputs: Markdown report with Spatial Omics Integration Score (0-100) Uses: 70+ ToolUniverse tools across 9 analysis phases Time: ~10-20 minutes depending on gene list size