Source

calesthio/generative-media-skills

151 skills · 8.5K combined installs

Skills from this source

#
Skill
Source
8W Activity
Installs
1
fashion-campaign-production Provider-independent production workflow for AI-assisted fashion campaigns, lookbooks, editorial fashion films, ecomm…
calesthio/generative-media-skills
56
2
higgsfield-video Produce video (and its supporting stills) through Higgsfield (higgsfield.ai), a multi-model creative platform that fr…
calesthio/generative-media-skills
56
3
hyperframes-video-composition Provider-independent production workflow for AI agents assembling generated or source media into HyperFrames HTML/CSS…
calesthio/generative-media-skills
56
4
minimax-speech Use this skill when producing speech, narration, dubbing, localization, advertising voice, voice-clone previews, or i…
calesthio/generative-media-skills
56
5
video-to-audio-foley Use for video-conditioned audio and automated Foley production: adding synchronized sound effects, ambience, impacts,…
calesthio/generative-media-skills
56
6
alibaba-image-models Plan, prompt, call, edit, iterate, and productionize Alibaba Cloud Model Studio image generation with current Wan 2.7…
calesthio/generative-media-skills
55
7
deepmotion-animate-3d Use DeepMotion Animate 3D for markerless video-to-3D human motion capture through the cloud portal, sales-gated API, …
calesthio/generative-media-skills
55
8
kling-video Plan, prompt, generate, reference, edit, motion-control, and quality-review videos with Kuaishou's direct Kling AI vi…
calesthio/generative-media-skills
55
9
manim-explainer-animation Provider-independent production workflow for creating Manim-based explainer animations. Use when an agent must plan, …
calesthio/generative-media-skills
55
10
openai-realtime-voice Build production OpenAI Realtime voice agents and low-latency spoken interactions with live audio sessions, WebRTC or…
calesthio/generative-media-skills
55
11
pika-video Produce short-form video with Pika (pika.art) — its effect-driven tools (Pikaffects, Pikadditions, Pikaswaps, Pikafra…
calesthio/generative-media-skills
55
12
resemble-chatterbox Use Resemble AI Chatterbox for text-to-speech and voice-cloning workflows, including local open-weight Chatterbox, Ch…
calesthio/generative-media-skills
55
13
seedance-2-0 Direct ByteDance Dreamina Seedance 2.0 Standard, Fast, and Mini video production across BytePlus ModelArk and verifie…
calesthio/generative-media-skills
55
14
suno-music Generate music with Suno (v5 / v5.5 era, 2026) and advise a production team on what they may legally do with the outp…
calesthio/generative-media-skills
55
15
xai-grok-imagine-video Produce, edit, extend, and govern short videos with xAI's direct Grok Imagine Video API. Use for text-to-video, image…
calesthio/generative-media-skills
55
16
hedra-character-video Produce Hedra character, talking-avatar, lip-sync, motion-avatar, and live avatar work. Use when an AI agent needs to…
calesthio/generative-media-skills
54
17
livestream-event-production Use this skill to plan, direct, troubleshoot, and hand off provider-independent live or hybrid livestream event produ…
calesthio/generative-media-skills
54
18
ltx-2-video Generate and edit synchronized audio-video with LTX-2.3 using the official hosted LTX API or official local/open-weig…
calesthio/generative-media-skills
54
19
openai-audio Produce and understand audio with OpenAI request-based audio APIs and audio-capable chat models, including text-to-sp…
calesthio/generative-media-skills
54
20
stable-audio Use for Stability AI Stable Audio production work: selecting Stable Audio hosted API or open-weight models, generatin…
calesthio/generative-media-skills
54
21
tencent-hunyuan3d Generate 3D assets with Tencent's Hunyuan3D family — open-weight self-hosted models (Hunyuan3D-2.0/2.1 and HunyuanWor…
calesthio/generative-media-skills
54
22
xai-grok-imagine-image Generate and edit production images with xAI's first-party Grok Imagine API, including model selection, multiple refe…
calesthio/generative-media-skills
54
23
bytedance-seedream Build and operate production image generation and natural-language image editing with ByteDance Seedream through firs…
calesthio/generative-media-skills
53
24
educational-animation-production Provider-independent production workflow for AI agents creating educational animated lessons, classroom explainers, S…
calesthio/generative-media-skills
53
25
elevenlabs-music Generate and iterate music with ElevenLabs Eleven Music for production deliverables. Use when planning, prompting, AP…
calesthio/generative-media-skills
53
26
food-beverage-content-production Provider-independent production guidance for AI agents creating or retouching appetizing food and beverage images, re…
calesthio/generative-media-skills
53
27
google-gemini-image Plan, generate, edit, and quality-check still images with Google's Gemini native image models (Nano Banana), includin…
calesthio/generative-media-skills
53
28
google-lyria Use Google Lyria music generation models through Google Cloud / Vertex AI / Gemini Enterprise Agent Platform for prod…
calesthio/generative-media-skills
53
29
image-generation-gateways Select, integrate, and operate multi-model image-generation gateways with model-specific schema discovery, version po…
calesthio/generative-media-skills
53
30
leonardo-image Create, edit, guide, upscale, and quality-control still images with Leonardo.Ai's official Production API, including …
calesthio/generative-media-skills
53
31
meshy-3d Produce 3D assets with Meshy (meshy.ai) through its REST API and web platform — text-to-3D, image-to-3D, multi-image-…
calesthio/generative-media-skills
53
32
midjourney-video Plan and direct Midjourney Video V1 image-to-video work through the official website or Discord, with human operator …
calesthio/generative-media-skills
53
33
minimax-hailuo-video Use MiniMax's first-party Hailuo video API safely and reproducibly across global and mainland-China platforms. Covers…
calesthio/generative-media-skills
53
34
minimax-music Produce music with MiniMax Music 2.6 and MiniMax cover/lyrics APIs for songs, instrumentals, AI-generated lyrics, ref…
calesthio/generative-media-skills
53
35
music-video-production Provider-independent production workflow for AI agents creating music videos, lyric videos, visualizers, performance/…
calesthio/generative-media-skills
53
36
openai-gpt-image Generate, edit, composite, stream, and production-review still images with OpenAI GPT Image models. Use when an agent…
calesthio/generative-media-skills
53
37
runway-image Generate, edit, and iterate still images with Runway's official API, especially native Gen-4 Image and Gen-4 Image Tu…
calesthio/generative-media-skills
53
38
runway-video Build and operate production-safe Runway API video generation, video editing, and character-performance workflows wit…
calesthio/generative-media-skills
53
39
stability-ai-image Operate Stability AI image generation, image-to-image, edit, control, background, and upscale APIs safely and reprodu…
calesthio/generative-media-skills
53
40
topaz-video-enhancement Use when enhancing, upscaling, restoring, denoising, sharpening, deinterlacing, stabilizing, motion-deblurring, color…
calesthio/generative-media-skills
53
41
vidu-video Plan and integrate ShengShu Vidu Open Platform video generation with current model/mode selection, reference consiste…
calesthio/generative-media-skills
53
42
ace-step Use ACE-Step and ACE-Step 1.5 for local or hosted AI music generation, including text-to-music, lyrics-to-song, instr…
calesthio/generative-media-skills
52
43
alibaba-wan-video Plan, implement, and review Alibaba Wan video generation using Alibaba Cloud Model Studio/DashScope hosted APIs or of…
calesthio/generative-media-skills
52
44
assemblyai-transcription Use this skill when an agent needs AssemblyAI for speech-to-text or speech-understanding work in media production, in…
calesthio/generative-media-skills
52
45
audiobook-production Produce full-length audiobooks and long-form narration with generative voice tools. Use when the task is to turn a ma…
calesthio/generative-media-skills
52
46
avatar-spokesperson-production Provider-independent production workflow for AI avatar spokesperson videos, including presenter briefs, consent and l…
calesthio/generative-media-skills
52
47
azure-speech Use Microsoft Azure Speech in Foundry Tools for media-production speech workflows: speech-to-text, fast and batch tra…
calesthio/generative-media-skills
52
48
cartesia-sonic Use Cartesia Sonic and related Cartesia voice APIs for production speech: text-to-speech, realtime WebSocket TTS, voi…
calesthio/generative-media-skills
52
49
comfyui-media-workflows Provider-independent production workflow for agents assembling, auditing, executing, and handing off ComfyUI node-gra…
calesthio/generative-media-skills
52
50
elevenlabs-dubbing-voice-conversion Use for ElevenLabs provider-specific dubbing, localization, voice changer, speech-to-speech, and voice isolation/enha…
calesthio/generative-media-skills
52
Page 2 · 151 total Previous Next