smithery/crossxwill

ydata-eda-profiling

Generate and compare ydata-profiling EDA reports with sampling, consistent random seeds, and HTML outputs; often follows duckdb-parquet-lab-workflow when data is queried from Parquet.

Installation

$ npx skills add smithery/crossxwill --skill ydata-eda-profiling

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from smithery/crossxwill.

npx skills add smithery/crossxwill

Browse all from smithery/crossxwill

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 977 B
  • docs SUMMARY.md 210 B

History

  1. First recorded snapshot · 0 installs

SKILL.md

YData Profiling EDA

Purpose

Create consistent EDA reports for train and test or accepted and rejected datasets using ProfileReport, including comparisons and saved HTML outputs.

Usage

  • "generate ydata-profiling report"
  • "compare train and test EDA"
  • "create HTML EDA report"

Instructions

  1. Set a sampling fraction for large datasets and a fixed random seed.
  2. Create ProfileReport objects with progress_bar=False, and disable duplicates and interactions when speed matters.
  3. Compare reports using .compare() and save to HTML with .to_file().
  4. Use ./scripts/generateedareport.py for consistent report creation.
  5. Use ./templates/edacompareblock.md to document inputs, sample fractions, and output paths.