seedance
Generate videos with ByteDance Seedance 2.0 via inference.sh CLI. Unified model for text-to-video, image-to-video, and reference-to-video with synchronized audio, up to 1080p, 4-15s duration. Pro and Fast variants. Studio variants with private asset library for portrait consistency. Use for: social media videos, music videos, product demos, animated content, AI video with sound. Triggers: seedance, seedance 2, bytedance video, seedance t2v, seedance i2v, seedance r2v, video with audio, seedance
Security Assessment
About seedance
seedance is a skill for generating videos with ByteDance's Seedance 2.0 model family through the inference.sh `belt` command-line tool. Seedance 2.0 is a unified model that handles text-to-video, image-to-video, first-and-last-frame control, and multimodal reference-to-video with synchronized audio at up to 1080p and 4–15 second durations. The skill packages this so an agent can produce sound-synced AI video from a single shell command without local model infrastructure.
Four variants are documented: Seedance 2.0 (best quality, up to 1080p), Seedance 2.0 Fast (quicker, up to 720p), and two Studio variants that automatically upload reference images to the BytePlus private virtual portrait library for stronger face and character consistency. Generation mode is inferred from the inputs — prompt only for text-to-video, prompt plus image for image-to-video, plus an end image for first/last-frame control, or reference image/video/audio arrays for multimodal guidance. The skill includes a detailed prompt guide (referencing assets by type and index such as Image 1 or Video 1), formulas for video editing, extension, and stitching, and worked examples for product ads, video element replacement, and reference-driven character scenes.
It targets creators and marketers producing social media videos, music videos, product demos, and animated content with sound. Because it relies entirely on the hosted inference.sh/BytePlus service, it is a documented front-end to those platforms. The Studio variants notably upload user-supplied portrait references to a third-party private library, which is a normal feature of the service but worth noting for anyone handling images of real people.
FAQ
What do I need to run it?
The inference.sh CLI (`belt`), authenticated with `belt login`. Commands take the form `belt app run bytedance/seedance-2-0 --input '{...}'`.
What generation modes does Seedance 2.0 support?
Text-to-video, image-to-video, first-and-last-frame control, and multimodal reference-to-video, all with optional synchronized audio via `generate_audio`.
What is the difference between the Studio variants and the standard ones?
Studio variants automatically upload reference images to the BytePlus private virtual portrait library for enhanced face and character consistency, particularly for branded characters or specific people.
How do I reference multiple input assets in a prompt?
Reference by type and index — Image 1, Image 2, Video 1, Audio 1 — matching the position in the arrays you provide. The skill says not to use asset IDs in prompts.
Are the modes combinable?
No. First-frame/last-frame inputs and reference inputs are mutually exclusive; the model picks one mode based on the inputs supplied.
Install seedance
Quick Setup:
- Copy the skill folder to
.claude/skills/ - Claude will automatically detect and use the skill
Repository
skills-101/superpowers