smithery.ai

gemini-image-analysis

Analyze one or multiple images with Gemini Flash (vision) via the generateContent REST API. Uses Bun + TypeScript to base64-encode images and print model output. Use for captioning, OCR-like extraction, UI/screenshot analysis, and multi-image comparison.

First seen Apr 6, 2026

Installation

$ npx skills add https://smithery.ai

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from smithery.ai · top by installs.

npx skills add https://smithery.ai

Browse all from smithery.ai

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Declared
Cursor Not declared
Codex Not declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Declared
Cline Not declared
OpenCode Not declared

Skill metadata

Parsed from SKILL.md frontmatter.

Declared agents claude-code gemini

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 1,732 B
  • docs SUMMARY.md 283 B

History

  1. First seen on skills.sh
  2. First recorded snapshot · 1 installs

SKILL.md

Gemini Image Analysis (Bun + TypeScript)

When to use

Use this Skill when you need to analyze one or multiple images (screenshots, photos, diagrams) with Gemini Flash via the REST API.

Quick start

GEMINI_API_KEY="YOUR_KEY" \
bun .claude/skills/gemini-image-analysis/scripts/gemini-image-analyze.ts \
  --prompt "Caption this image." \
  /path/to/image.jpg

macOS Keychain (optional)

If you store GEMINIAPIKEY in the macOS Keychain:

GEMINI_API_KEY="$(security find-generic-password -a "$(whoami)" -s "GEMINI_API_KEY" -w)" \
bun .claude/skills/gemini-image-analysis/scripts/gemini-image-analyze.ts \
  --prompt "Caption this image." \
  /path/to/image.jpg

Multiple images

Send multiple images in one request (useful for comparison or “before/after”):

GEMINI_API_KEY="YOUR_KEY" \
bun .claude/skills/gemini-image-analysis/scripts/gemini-image-analyze.ts \
  --prompt "Compare these images and describe the differences." \
  /path/to/image1.jpg \
  /path/to/image2.png

Options

  • --prompt <text>: Text instruction for the model (default: Caption this image.)
  • --model <model>: Model name (default: gemini-3-flash-preview)
  • --json: Print the full JSON response instead of extracted text

Notes

  • The script reads GEMINIAPIKEY from the environment and never stores it.
  • Supported image types: .jpg, .jpeg, .png, .webp, .gif.