Source

calesthio/generative-media-skills

151 skills · 8.5K combined installs

Skills from this source

#
Skill
Source
8W Activity
Installs
1
elevenlabs-scribe Use ElevenLabs Scribe for speech-to-text production workflows: transcribing audio or video files, diarization and spe…
calesthio/generative-media-skills
52
2
gemini-live-audio Build and evaluate low-latency spoken, multimodal, and translation experiences with Google Gemini Live API and Gemini…
calesthio/generative-media-skills
52
3
google-gemini-omni-video Generate, reference, and conversationally edit short videos with Google's Gemini Omni Flash through the Gemini Develo…
calesthio/generative-media-skills
52
4
heygen-avatar-video Produce HeyGen avatar videos with Direct Video, Video Agent, Digital Twin, Avatar Realtime, and related avatar/voice/…
calesthio/generative-media-skills
52
5
ideogram-image Generate, remix, edit, inpaint, reframe, background-process, describe, layerize, and upscale images with the Ideogram…
calesthio/generative-media-skills
52
6
luma-ray-video Direct and operate Luma Ray video generation through the current Luma Agents API and distinguish it from the consumer…
calesthio/generative-media-skills
52
7
move-ai-motion-capture Use this skill when planning, capturing, processing, validating, troubleshooting, or exporting markerless human motio…
calesthio/generative-media-skills
52
8
nvidia-cosmos-video Select, run, and govern NVIDIA Cosmos world-video generation across Cosmos 3 Generator, Predict2.5, Transfer2.5, down…
calesthio/generative-media-skills
52
9
qwen3-tts Produce text-to-speech with Alibaba/Qwen Qwen3-TTS through DashScope/Model Studio or open-weight Qwen3-TTS checkpoint…
calesthio/generative-media-skills
52
10
recraft-image-design Design and produce raster images, native SVG artwork, brand-controlled visuals, and documented image edits with Recra…
calesthio/generative-media-skills
52
11
remotion-video-composition Provider-independent production workflow for assembling generated or sourced media into React/Remotion videos. Use wh…
calesthio/generative-media-skills
52
12
twelvelabs-video-understanding Use when an agent must make TwelveLabs do real video-understanding work: indexing footage, semantic/visual search acr…
calesthio/generative-media-skills
52
13
world-labs-marble Generate persistent, explorable 3D worlds/environments with Marble by World Labs — from text, a single image, multipl…
calesthio/generative-media-skills
52
14
adobe-firefly-image Use Adobe Firefly Services to generate, edit, expand, fill, match, composite, and upscale still images through the cu…
calesthio/generative-media-skills
51
15
amazon-nova-canvas Produce and review production image-generation and image-editing workflows with Amazon Nova Canvas on Amazon Bedrock,…
calesthio/generative-media-skills
51
16
amazon-nova-reel Produce and operate Amazon Nova Reel video-generation jobs through Amazon Bedrock. Use for Nova Reel text-to-video, i…
calesthio/generative-media-skills
51
17
amazon-polly Use Amazon Polly for production text-to-speech work: selecting Standard, Neural, Long-form, or Generative engines and…
calesthio/generative-media-skills
51
18
amazon-transcribe Use Amazon Transcribe for AWS-based speech-to-text production: batch S3 transcription, real-time streaming, captions/…
calesthio/generative-media-skills
51
19
black-forest-labs-flux Plan, prompt, execute, troubleshoot, and quality-control Black Forest Labs FLUX image generation and editing across t…
calesthio/generative-media-skills
51
20
bria-fibo-image Build and operate rights-aware Bria FIBO and FIBO Lite image-generation workflows with structured prompts, reference …
calesthio/generative-media-skills
51
21
d-id-avatar-video Use D-ID to plan, generate, stream, localize, and QA avatar/talking-head videos, including V2 Photo Avatar Talks, V3 …
calesthio/generative-media-skills
51
22
deepgram-speech Use for Deepgram speech and voice production workflows: speech-to-text transcription, live captions, diarization, aud…
calesthio/generative-media-skills
51
23
elevenlabs-agents Build, configure, and ship production voice agents on the ElevenLabs Agents platform (branded "ElevenAgents," formerl…
calesthio/generative-media-skills
51
24
google-cloud-speech Use this skill when a media-production agent needs Google Cloud speech and voice services for transcription, captions…
calesthio/generative-media-skills
51
25
hume-evi Build production realtime voice agents with Hume's Empathic Voice Interface (EVI): speech-to-speech sessions, empathi…
calesthio/generative-media-skills
51
26
hume-octave Use Hume Octave for emotionally expressive speech and voice production: text-to-speech, voice design, voice cloning, …
calesthio/generative-media-skills
51
27
immersive-spatial-video-production Use this skill to plan, direct, finish, and QA provider-independent mono or stereo 180-degree, 360-degree, spherical,…
calesthio/generative-media-skills
51
28
kling-kolors-image Plan, generate, edit, reference, and quality-control still images with Kling AI's hosted IMAGE surfaces or Kuaishou's…
calesthio/generative-media-skills
51
29
localization-dubbing-production Provider-independent localization and dubbing production direction for AI agents producing translated videos, dubbed …
calesthio/generative-media-skills
51
30
luma-photon Use Luma AI's first-party Photon and Photon Flash image API for text-to-image, image/style/character references, imag…
calesthio/generative-media-skills
51
31
moonvalley-marey Produce video with Moonvalley's Marey model family (Marey Realism v1.5) — a filmmaker-oriented, 1080p/24fps generativ…
calesthio/generative-media-skills
51
32
nvidia-maxine-audio-effects Use NVIDIA Maxine / NVIDIA AI for Media audio effects for speech cleanup and enhancement in live or offline media pip…
calesthio/generative-media-skills
51
33
nvidia-speech-nim Use NVIDIA Speech NIM microservices for speech production and voice workflows, including self-hosted ASR/STT, TTS, te…
calesthio/generative-media-skills
51
34
odyssey-interactive-video Plan and scope production work with Odyssey's real-time interactive video world models (odyssey.ml — Odyssey-1/Odysse…
calesthio/generative-media-skills
51
35
podcast-production Produce audio-first podcast episodes with generative tools — design the show and episode format (interview, narrative…
calesthio/generative-media-skills
51
36
procedural-canvas-animation Provider-independent production guidance for deterministic Canvas 2D and p5.js animation. Use for particles, fields, …
calesthio/generative-media-skills
51
37
sync-labs-lipsync Generate AI lip sync and visual dubbing with Sync Labs (sync.so) — choose the right model (sync-3, lipsync-2, lipsync…
calesthio/generative-media-skills
51
38
synthesia-avatar-video Produce presenter-led AI avatar videos with Synthesia Studio and Synthesia API. Use for Synthesia-specific avatar vid…
calesthio/generative-media-skills
51
39
tavus-replica-video Use for producing Tavus AI-human videos and real-time avatar conversations with Tavus Faces/Replicas, PALs/Personas, …
calesthio/generative-media-skills
51
40
tencent-hunyuanvideo Generate and operate Tencent Hunyuan video through the managed TokenHub HY-Video-1.5 API or official local HunyuanVid…
calesthio/generative-media-skills
51
41
virtual-production-icvfx Plan, troubleshoot, and hand off provider-independent in-camera VFX and LED volume virtual-production work, with Unre…
calesthio/generative-media-skills
51
42
gsap-animation-composition Production guidance for authoring, integrating, and reviewing GSAP animation in browser-rendered media. Use for deter…
calesthio/generative-media-skills
50
43
audio-reactive-video-composition Provider-independent production guidance for translating measured audio features into deterministic video timing and …
calesthio/generative-media-skills
49
44
d3-animated-data-visualization Production guidance for converting sourced data and approved claims into truthful, accessible, deterministic animated…
calesthio/generative-media-skills
49
45
precise-video-description Provider-independent production guidance for converting observed video into precise, objective, temporally ordered la…
calesthio/generative-media-skills
48
46
threejs-scene-composition Production guidance for planning, building, animating, capturing, and reviewing complete Three.js scenes for rendered…
calesthio/generative-media-skills
48
47
volcengine-doubao-speech-tts Production guidance for mainland-China Volcengine Doubao Speech text-to-speech. Use for TTS 1.0/2.0 selection, V3 bid…
calesthio/generative-media-skills
48
48
byteplus-seed-speech-tts Production guidance for international BytePlus Seed Speech text-to-speech. Use for selecting TTS 1.0 versus 2.0, bidi…
calesthio/generative-media-skills
47
49
kling-advanced-lip-sync Production guidance for Kling AI Open Platform Advanced Lip-Sync. Use for identifying/selecting one face in an existi…
calesthio/generative-media-skills
47
50
lottie-animation-delivery Production guidance for assessing, exporting, packaging, integrating, capturing, validating, and handing off Lottie v…
calesthio/generative-media-skills
47
Page 3 · 151 total Previous Next