seedance-podcast-visual
Generate podcast clip visualization video prompts for Seedance 2.0 on Higgsfield. Use for podcast clip videos, audio-to-visual content, audiogram alternatives, podcast highlight reels, interview clip visuals, or any video that transforms audio content into engaging visual format. Triggers on podcast, audio clip, audiogram, interview clip, sound bite, audio visual, podcast video, episode highlight, podcast clip.
git clone --depth 1 https://github.com/rediumvex/ai-video-generator-claude /tmp/seedance-podcast-visual && cp -r /tmp/seedance-podcast-visual/skills/10-podcast-visual ~/.claude/skills/seedance-podcast-visualSKILL.md
# Podcast Visual — Audio-to-Video Transformation Prompts Transform podcast audio into cinematic visual content using Seedance 2.0 on Higgsfield. This skill produces video prompts that replace static audiograms with storytelling-driven visual experiences built entirely from constructed imagery. --- ## Input Specifications **Primary inputs:** - Up to 3 audio files (podcast clips, interview excerpts, sound bites, episode highlights) - Transcript or key quote text from the audio - Speaker name(s) and brief context (topic, show name, tone) - Desired visual style (abstract, cinematic, interview reconstruction, kinetic) - Target platform (Instagram Reels, YouTube Shorts, LinkedIn, TikTok) - Aspect ratio: 9:16 (vertical/mobile-first), 16:9 (widescreen), or 1:1 (square) **Audio file handling:** - File 1: Primary clip — the main sound bite or key quote being visualized - File 2 (optional): Intro or context clip — sets up the narrative before the hook - File 3 (optional): Reaction or follow-up clip — speaker response, co-host moment, audience reaction - Duration guidance: each clip should be 15–90 seconds; total sequence up to 3 minutes **What you extract from audio before writing prompts:** - The single most quotable sentence (becomes the visual anchor) - The emotional register: contemplative, fired-up, vulnerable, instructive, funny - Pacing: fast and punchy vs. slow and deliberate delivery - Natural pauses: where silence lives (these become visual breath moments) - Speaker energy level: seated calm, animated gesturing, emotional peak --- ## Philosophy | Old model (audiogram) | New model (podcast visual) | |---|---| | Show the waveform | Show what the words feel like | | Static background image | Constructed cinematic environment | | Speaker photo as thumbnail | Speaker reconstructed in scene | | Generic brand colors | Lighting and atmosphere matched to tone | | Passive viewing | Active emotional engagement | | Optimized for "audio on" | Compelling even on mute | --- ## 2-Second Hook Patterns The hook is the opening frame that stops the scroll. It must communicate emotion, intrigue, or tension before a single word is heard. Four proven structures: ### The Quote Impact Display the most provocative line from the clip as large kinetic text before audio begins. The text arrives with weight — not a gentle fade, but a hard cut or a push-in. The visual behind it is blurred or dark, forcing the text into full focus. **When to use:** clips with a single devastating sentence, contrarian takes, counterintuitive statistics, direct challenges to conventional wisdom. **Visual execution in prompt:** specify "bold white sans-serif typography slams onto dark background, camera holds for 1.5 seconds, then cuts to speaker close-up, shallow depth of field, background softly bokeh'd." ### The Reaction Shot Open on the speaker's face at the moment of peak emotional expression — surprise, laughter, conviction, vulnerability — before any context is given. This creates a curiosity gap: the viewer needs to hear what caused that expression. **When to use:** interview moments where a genuine reaction occurs, storytelling clips where the speaker relives something visceral, moments of realization or revelation. **Visual execution in prompt:** specify "extreme close-up on speaker's face, caught mid-expression, eyes slightly wide, ambient room sound implied by environment, camera slowly eases back over 3 seconds to reveal setting." ### The Visual Metaphor Instead of showing the speaker at all, open with an environmental or abstract image that represents the core concept of the clip. A podcast about burnout opens on dying embers. A clip about compounding returns opens on a single drop rippling outward. The metaphor does expository work so the audio can focus on depth. **When to use:** concept-heavy clips, philosophical discussions, any clip where the idea is more powerful than the person delivering it. **Visual execution in prompt:** specify the metaphor object explicitly, its lighting, its motion quality, and a precise camera behavior (slow push, orbital, static hold with foreground element drifting through). ### The Sound Wave Art Not a functional audiogram waveform — instead, an artistic rendering of sound as visual sculpture. Particles forming and dissolving in rhythm with imagined speech cadence. Light bending through air as if vibrated by voice. Sound made beautiful, not informational. **When to use:** music-adjacent podcasts, high-production brand content, moments where you want to foreground the craft of the medium itself. **Visual execution in prompt:** specify particle behavior, color palette tied to the emotional register of the clip, and whether motion is rhythmic/predictable or fluid/organic. Avoid the word "waveform" — describe it as "acoustic particle field" or "resonant light diffusion." --- ## Visual Formats ### Abstract Visualization The audio inspires a visual world that does not contain the speaker at all. Instead, abstract imagery — light, texture, particle systems, color gradients, fluid dynamics — evolves in response to the imagined emotional arc of the audio. **Core parameters:** - Color temperature must match emotional tone (cool/blue for analytical, warm/amber for intimate, high-contrast for confrontational) - Motion should breathe with speech rhythm — slowing during pauses, accelerating during emphasis - Avoid literal representation; the visual is interpretive, not illustrative - Works best at 9:16 for mobile, full-bleed composition **Prompt elements to always include:** dominant color palette, motion behavior (fluid, particle, crystalline, liquid, smoke), camera behavior (static, slow push, orbital), and whether the environment is finite (a room implied by light edges) or infinite (void space) ### Cinematic B-Roll Narrative Construct a series of visuals that would, in a traditional documentary, accompany the audio as b-roll. Except here every frame is generated — n
Generate scroll-stopping viral hook video prompts for Seedance 2.0 on Higgsfield. Use whenever the user wants viral content, TikTok hooks, Instagram Reels openers, YouTube Shorts, attention-grabbing video, scroll-stopper, pattern interrupt, or any short-form video designed to maximize retention and views. Triggers on any mention of viral, hook, scroll-stop, retention, views, engagement, short-form, TikTok, Reels, Shorts.
Generate SaaS product launch and software demo video prompts for Seedance 2.0 on Higgsfield. Use whenever the user wants a product launch video, app demo, software walkthrough, feature showcase, startup promo, tech product reveal, or any SaaS/software marketing video. Triggers on SaaS, app, software, product launch, demo, feature, startup, tech, UI, dashboard, landing page video.
Generate personal brand and founder story video prompts for Seedance 2.0 on Higgsfield. Use for authority content, day-in-the-life, founder story, personal brand building, thought leadership, creator content, lifestyle videos, behind-the-scenes, or any video meant to build a personal brand presence. Triggers on personal brand, founder, creator, authority, lifestyle, behind the scenes, day in the life, thought leader.
Generate online course and coaching program promotional video prompts for Seedance 2.0 on Higgsfield. Use for course trailers, coaching ads, educational content promos, masterclass teasers, webinar promotions, or any educational product video. Triggers on course, coaching, masterclass, webinar, tutorial, education, teaching, training, class, program, academy.
Generate faceless content video prompts for Seedance 2.0 on Higgsfield. Use for faceless YouTube channels, TikTok content without showing face, anonymous creator content, narration-driven videos, stock-footage-style AI content, or any video where the creator doesn't appear on camera. Triggers on faceless, no face, anonymous, narration, stock footage, b-roll, background video, voiceover video.
Generate luxury and premium aesthetic video prompts for Seedance 2.0 on Higgsfield. Use for luxury brand content, premium product showcases, high-end lifestyle, minimalist aesthetic, elegant brand videos, or any content requiring sophisticated visual treatment. Triggers on luxury, premium, high-end, elegant, minimalist, sophisticated, exclusive, refined, designer, couture, bespoke.
Generate before-and-after transformation video prompts for Seedance 2.0 on Higgsfield. Use for transformation reveals, glow-ups, makeovers, renovation reveals, fitness transformations, design before-after, business growth visuals, or any content showing dramatic change. Triggers on before after, transformation, reveal, glow up, makeover, renovation, redesign, progress, results.
Generate customer testimonial and social proof video prompts for Seedance 2.0 on Higgsfield. Use for customer stories, case study videos, review showcases, social proof content, success story videos, or any content featuring customer results and experiences. Triggers on testimonial, review, case study, social proof, customer story, success story, results, feedback, client.