Assemble a cosmic-mythology-voiceover reel from a config — a warm spoken voiceover carries the whole narrative while N curated cosmic stills are weighted beat-synced across the delivered VO duration (cut_dur = VO_dur times weight over the weight sum, so emotional beats hold longer), Ken-Burns-zoomed per still (scale 2x, center crop, zoompan, fade-in first and fade-out last), ffmpeg-concatenated, the VO composited under the picture (libx264 crf18 plus aac), the ONE on-screen hook line faded on over the open with a drawtext alpha window, and Whisper/VEED captions burned along the bottom — never in-world text on a still. This is the FREE deterministic assembly stage (weighted sequence plus Ken-Burns plus concat plus VO composite plus hook overlay plus caption burn); the VO and the stills come from create-vo-elevenlabs and create-image-fal. Use for the cosmic-mythology-voiceover format.
npx skills add https://github.com/gooseworks-ai/goose-skills --skill render-cosmic-mythology-voiceover
Assemble a cosmic-mythology-voiceover reel from a config: a faceless, cinematic storytelling
video where a warm, contemplative spoken voiceover carries the whole narrative (a "myth as
teacher" reframe) over a slow, weighted Ken-Burns zoom across curated cosmic / mythology stills
in ONE ethereal deep-indigo + gold look, with ONE on-screen hook line and burned captions. This
capability is the FREE, deterministic assembly — the weighted beat-sync sequencing, the
Ken-Burns render, the concat, the VO composite, the hook overlay, and the caption burn.
scripts/config.example.json is the worked example (WishAstro "Saturn isn't your villain", ~31s
1080×1920 9:16, 12 weighted Ken-Burns cuts); scripts/PIPELINE.md maps every config block to its
source step and scripts/README.md documents the free assembly.
There is a single runnable script — scripts/render.py (config-driven, ffmpeg + Pillow only, NO
API keys and NO drawtext/libass required):
python3 scripts/render.py --config config.json --vo working/vo2/vo_atempo.mp3 \
--stills-dir working/stills --out working/final.mp4 \
[--words working/vo2/words.json] [--endcard working/endcard.png]
This is the FREE, deterministic assembly stage — it spends nothing. The paid inputs are
separate capabilities — the spoken VO (create-vo-elevenlabs, ElevenLabs eleven_v3 from a
tone-tagged script, atempo time-stretched so the delivered duration sets the timeline) and the 4–6
hero stills in one look pack (create-image-fal, Flux Pro 1.1, reused as repeats to reach the
~10–12 cuts). Given the VO + the stills + the per-cut weight array + the hook line, render.py
distributes the cuts across the VO duration by the weighted formula, Ken-Burns-renders each still,
concats, composites the VO, fades the hook line on over the open, burns the captions, and (if
--endcard is passed) appends a brand end card → the master + a poster. Re-cuts reuse the existing
VO / stills and cost $0. See scripts/README.md §0 for the full arg contract.
VO IS the audio bed (no music bed by default); do not add a presenter or a second bed.
the factor ≤ ~1.25 so the voice never chipmunks); its delivered length sets the timeline — never
trim the VO to a pre-planned grid.
cut_dur = VO_dur × weight / Σweights —heavier weights hold longer on the emotional beats (the open, the reframe, the close); the setup
cuts run shorter. Every cut stays proportional to the whole VO.
zoompan to theconfigured zoom_end (~1.10); apply zoom_out on the flagged cuts; fade_in on the FIRST cut
and fade_out on the LAST. Stills are reusable — the sequence repeats a few across the cuts.
volumetric-light look; the reel's only text is the hook + the captions, added in post (the "no
text, no words" descriptor keeps words off the stills).
(fade in ~0.5s, hold, fade out ~0.6s) — never a persistent caption, never in-world. render.py
does this with a PIL PNG + ffmpeg fade=…:alpha=1 (no drawtext dependency, since stock ffmpeg
often lacks it); an ffmpeg drawtext alpha window is an equivalent alternative where available.
bottom third, white #FFFFFF. If the host ffmpeg lacks libass (no subtitles/ass filter),
render the cues as timed PIL PNG overlays (ffmpeg overlay=…:enable='between(t,st,en)') at the
same bottom placement — a free local Whisper + ffmpeg burn is the fallback to the VEED tier.
under the picture (libx264 crf 18 + aac 192k), burn the hook alpha-fade + the caption track →
a 1080×1920 h264+aac master (~31s). No paid calls, no keys (the local caption fallback is free).
Create beautiful visual art in .png and .pdf documents using design philosophy. You should use this skill when the user asks to create a poster, piece of art, design, or other static piece. Create original visual designs, never copying existing artists' work to avoid copyright violations.
Creating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, generative art, algorithmic art, flow fields, or particle systems. Create original algorithmic art rather than copying existing artists' work to avoid copyright violations.
Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, documentation, or social media posts.
Downloads videos from YouTube and other platforms for offline viewing, editing, or archival. Handles various formats and quality options.
Lightweight WSI tile extraction and preprocessing. Use for basic slide processing tissue detection, tile extraction, stain normalization for H&E images. Best for simple pipelines, dataset preparation, quick tile-based analysis. For advanced spatial proteomics, multiplexed imaging, or deep learning pipelines use pathml.
Microscopy data management platform. Access images via Python, retrieve datasets, analyze pixels, manage ROIs/annotations, batch processing, for high-content screening and microscopy workflows.
Python library for working with DICOM (Digital Imaging and Communications in Medicine) files. Use this skill when reading, writing, or modifying medical imaging data in DICOM format, extracting pixel data from medical images (CT, MRI, X-ray, ultrasound), anonymizing DICOM files, working with DICOM metadata and tags, converting DICOM images to other formats, handling compressed DICOM data, or processing medical imaging datasets. Applies to tasks involving medical image analysis, PACS systems, radiology workflows, and healthcare imaging applications.
This skill should be used when working with pre-trained transformer models for natural language processing, computer vision, audio, or multimodal tasks. Use for text generation, classification, question answering, translation, summarization, image classification, object detection, speech recognition, and fine-tuning models on custom datasets.
Take gooseworks-ai/render-cosmic-mythology-voiceover from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.