smithery.ai

duckdb-parquet-lab-workflow

Use DuckDB to query Parquet files, inspect metadata, join tables, and convert results to pandas for analysis; commonly precedes ydata-eda-profiling for EDA on extracted tables.

First seen Apr 6, 2026

Installation

$ npx skills add https://smithery.ai

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from smithery.ai · top by installs.

npx skills add https://smithery.ai

Browse all from smithery.ai

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 947 B
  • docs SUMMARY.md 211 B

History

  1. First seen on skills.sh
  2. First recorded snapshot · 1 installs

SKILL.md

DuckDB Parquet Lab Workflow

Purpose

Standardize the pattern of loading Parquet files into DuckDB, inspecting schema, running SQL joins, and converting results to pandas DataFrames.

Usage

  • "load Parquet with DuckDB and join tables"
  • "describe DuckDB table schema"
  • "convert DuckDB query to pandas"

Instructions

  1. Read Parquet data with duckdb.query or duckdb.sql using SQL strings.
  2. Inspect schema using DESCRIBE SELECT * FROM <table> and display with .show().
  3. Use explicit joins with clear LEFT or RIGHT semantics to preserve row counts.
  4. Convert results to pandas with .to_df() for downstream modeling.
  5. Use ./templates/duckdb_snippets.md for the standard SQL patterns.