muapi-seedance-2
Expert Cinema Director skill for Seedance 2.0 (ByteDance) — high-fidelity video generation across Chinese, Global, and VIP tiers. Supports text-to-video, image-to-video, first-last-frame, omni reference, character training, omni-reference training, video editing, and watermark removal.
Install / Use
npx skills add SamurAIGPT/Generative-Media-Skills --skill muapi-seedance-2Installs into whichever agent you are using.
SKILL.md
Installable skill definition
Quality Score
Category
Customer SupportSupported Platforms
Our assessment of muapi-seedance-2
muapi-seedance-2 scores 95/100 on our quality scale, 25th of 214 Customer Support skills we index (top 12%).
Its SKILL.md is 26 KB long, well organised into 86 sections with 36 code examples: a thorough specification that gives an agent plenty to work with.
With 4,329 GitHub stars, it is one of the more widely adopted skills in the catalogue.
Maintenance, license and trust
- The repository was last updated 19 days ago, so muapi-seedance-2 is actively maintained.
- It is released under the MIT license, a permissive license that allows use, modification and commercial use with attribution.
- Its trust signals score 100/100, with no cautions. These come from repository metadata, not a code audit — read the skill file before letting an agent act on it.
muapi-seedance-2 compared with similar skills
All 4 of these similar skills score higher than muapi-seedance-2; compare them before choosing.
| Skill | Score | Stars | Updated | Format |
|---|---|---|---|---|
| muapi-seedance-2 (this skill)by SamurAIGPT | 95 | 4.3k | 19d ago | SKILL.md |
| Agent-Reachby Panniantong | 100 | 85.9k | 13d ago | CLAUDE.md |
| headroomby headroomlabs-ai | 100 | 74.0k | 1d ago | CLAUDE.md |
| crawl4aiby unclecode | 100 | 84.4k | 3d ago | MCP Server |
| Scraplingby D4Vinci | 100 | 84.2k | 1d ago | MCP Server |
Frequently asked questions
- How do I install muapi-seedance-2?
- Run
npx skills add SamurAIGPT/Generative-Media-Skills --skill muapi-seedance-2. The install tabs above show the steps for each supported agent. - Which AI agents does muapi-seedance-2 work with?
- It is written for Universal, as a SKILL.md file. Other agents that read the same format can often use it too.
- Is muapi-seedance-2 safe to use?
- It is MIT-licensed and scores 100/100 on trust signals. Skills are instructions an agent will follow, so read the file before installing it and do not approve commands you do not understand.
- Is muapi-seedance-2 still maintained?
- The repository was last updated 19 days ago, so muapi-seedance-2 is actively maintained.
Skill content
View source on GitHubslug: muapi-seedance-2 name: muapi-seedance-2 version: "0.3.0" description: Expert Cinema Director skill for Seedance 2.0 (ByteDance) — high-fidelity video generation across Chinese, Global, and VIP tiers. Supports text-to-video, image-to-video, first-last-frame, omni reference, character training, omni-reference training, video editing, and watermark removal. acceptLicenseTerms: true
🎬 Seedance 2.0 Cinema Expert
The definitive skill for "Director-Level" AI video orchestration. Seedance 2.0 is not a descriptive model; it is an instructional model. It responds best to technical cinematography, physics directives, and precise camera grammar.
Core Competencies
- Text-to-Video (t2v): Generate cinematic video from a Director Brief — Chinese, Global, or VIP tier.
- Image-to-Video (i2v): Animate 1–9 reference images — Chinese, Global (smart mode), or VIP tier.
- Video Extension (extend): Seamlessly continue an existing Seedance 2.0 video (Chinese tier).
- First & Last Frame (first-last): Interpolate a fluid video between a start image and end image (Global/VIP).
- Omni Reference (omni): Full multimodal reference with images + audio + character refs (all tiers).
- Omni Reference Training (omni-train): Train a custom persistent character for identity-consistent generation.
- Character Sheet (character): Build a reusable character from 1–3 images (Chinese tier).
- Video Edit (video-edit): Edit an existing video with a prompt + optional reference images (Chinese tier).
- Watermark Removal (watermark-remove): Strip Seedance 2.0 watermarks (basic or Pro).
🏷️ Tiers
| Tier | Flag | Censorship | Aspect Ratios | Duration | Quality param |
|:---|:---|:---|:---|:---|:---|
| Chinese (default) | --tier chinese | Low | 16:9, 9:16, 4:3, 3:4 | 5 / 10 / 15 s | Yes (basic/high) |
| Global | --tier global | Standard | + 21:9, 1:1 | Any 4–15 s | No |
| VIP | --tier vip | Low | + 21:9, 1:1 | Any 4–15 s | No |
Add --fast to any Global or VIP call to use the fast-queue variant (lower latency, same quality).
📥 Input Limits
| Input Type | Chinese i2v/omni | Global/VIP i2v/omni | Formats | Max Size | |:---|:---|:---|:---|:---| | Images | ≤ 9 | ≤ 9 | jpeg, png, webp | 30 MB each | | Videos | ≤ 3 (omni only) | Not supported | mp4, mov | 50 MB each | | Audio | ≤ 3 | ≤ 3 | mp3, wav | 15 MB each | | First-Last | — | 1–2 images | jpeg, png, webp | 30 MB each | | Video Edit | 1 video + ≤ 9 imgs | — | mp4 ≤ 10 MB / 15s | — |
Output: 4–15 seconds, auto-generated sound, 480p–720p.
⚠️ Restrictions
- No realistic human faces in uploaded images/videos (except character/omni-train modes).
--mode extendrequires arequest_idfrom a priorseedance-v2.0-t2vorseedance-v2.0-i2vjob.--mode first-lastrequires--tier globalor--tier vip.- Global/VIP omni does not support video references (images + audio only).
--qualityapplies to Chinese tier only.
🔗 Core Syntax: The @ Reference System
Assign explicit roles to each uploaded asset. Tags differ by mode.
Chinese Tier (i2v, omni)
@image1 @image2 ... @image9 (images_list order)
@video1 @video2 @video3 (video_files order)
@audio1 @audio2 @audio3 (audio_files order)
Global/VIP Omni (omni-reference-no-video / vip-omni-reference)
@image1 @image2 ... @image9 (images_list order)
@audio1 @audio2 @audio3 (audio_files order)
Character References (all tiers)
@character:<request_id> — from seedance-2-character or completed t2v/i2v job
@omni-character:<character_id> — from seedance-2-omni-reference-train output
Role Assignment Table
| Purpose | Example Syntax |
|:---|:---|
| First frame | @Image1 as the first frame |
| Last frame | @Image2 as the last frame |
| Character appearance | @Image1's character as the subject |
| Scene / background | scene references @Image3 |
| Camera movement | reference @Video1's camera movement |
| Action / motion | reference @Video1's action choreography |
| Visual effects | completely reference @Video1's effects and transitions |
| Rhythm / tempo | video rhythm references @Video1 |
| Voice / tone | narration voice references @Video1 |
| Background music | BGM references @Audio1 |
| Sound effects | sound effects reference @Video3's audio |
| Outfit / clothing | wearing the outfit from @Image2 |
| Product appearance | product details reference @Image3 |
Multi-Reference Combination
@Image1's character as the subject, reference @Video1's camera movement
and action choreography, BGM references @Audio1, scene references @Image2
🏗️ Technical Specification: The Director Brief
Structure prompts using this six-component hierarchy. Order matters — composition first, texture and micro-motion last:
| Component | Instruction Type | Example | |:---|:---|:---| | Scene | Environment + Lighting | "A rain-soaked cyberpunk street, magenta neon reflections on wet asphalt." | | Subject | Identity + Detail | "A woman in a black trenchcoat, determined focus, cinematic skin textures." | | Action | Fluid Interaction | "Walking forward through the crowd, coat billowing slightly in the wind." | | Camera | Movement + Lens + Speed | "Medium tracking shot, 35mm lens, slow dolly backward over 6s. Subtle handheld jitter." | | Audio | Music + SFX + Ambience | "Low ambient hum, distant traffic, single piano note at 5s. No dialogue." | | Pacing/Style | Timing + Mood + Grade | "Cinematic epic, warm color grade, shallow DOF. Slow build — single action only, no scene cuts." |
Seedance 2.0 generates audio natively. Always include an Audio directive — even one sentence. Without it the model generates random ambient sound that may not match your scene.
Time-Segmented Prompts (Recommended for 10s+ videos)
Break prompts into timed segments for precise control:
0–3s: [opening scene, camera move, establishing action]
3–6s: [mid-section development, subject in motion]
6–10s: [climax or key action beat]
10–15s: [resolution, brand/product hold, text/tagline fade in]
Single-beat rule: Each segment should contain one action. 4–7s = one beat. 10–15s = 3–4 beats maximum. Overloading a segment with multiple narrative changes degrades output quality.
Negative Prompting
Seedance 2.0 supports appending negative guidance directly in the prompt. Use plain language at the end:
[your director brief above]
Avoid: camera shake, jump cuts, lens distortion, overexposure, watermarks, text overlays.
Common negative additions:
Avoid: abrupt cuts, scene changes, multiple locations.(for single-take shots)Avoid: human faces, realistic people.(for product-only content)Avoid: fast motion, blur, unstable framing.(for smooth product reveals)
🎥 Camera Language Reference
Basic Movements
| Term | Description | |:---|:---| | Push in / Slow push | Camera moves toward subject | | Pull back / Pull away | Camera moves away from subject | | Pan left/right | Camera rotates horizontally | | Tilt up/down | Camera rotates vertically | | Track / Follow shot | Camera follows subject movement | | Orbit / Revolve | Camera circles around subject | | One-take / Oner | Continuous shot with no cuts |
Advanced Techniques
| Term | Description | |:---|:---| | Hitchcock zoom (dolly zoom) | Push in + zoom out — creates vertigo effect | | Fisheye lens | Ultra-wide distorted lens | | Low angle / High angle | Camera below/above subject | | Bird's eye / Overhead | Top-down view | | First-person POV (FPV) | Immersive subjective camera from character/object's eyes — GoPro-style wide angle, forward motion, no cuts | | Drone flythrough | Cinematic aerial descent — gimbal-stabilized, sweeping lateral arc, DJI Inspire aesthetic | | Architectural flythrough | Ground-level continuous dolly through connected spaces — one-take, practical lighting | | Whip pan | Very fast horizontal pan with motion blur | | Crane shot | Vertical movement like a crane arm |
Shot Sizes
| Term | Description | |:---|:---| | Extreme close-up | Eyes, mouth, or small detail only | | Close-up | Face fills frame | | Medium close-up | Head and shoulders | | Medium shot | Waist up | | Full shot | Entire body | | Wide / Establishing shot | Full environment |
🧠 Prompt Optimization Protocol
The Agent MUST transform user intent into a technical "Director Brief" before execution.
- Technical Grammar: Use camera terms: Dolly In/Out, Crane Shot, Whip Pan, Tracking Shot, Anamorphic Lens, Shallow Depth of Field, High-Speed Dive, Orbital Arc.
- Physics Directives: Use "caustic patterns," "volumetric rays," or "subsurface scattering" instead of "good lighting."
- Timecode Notation: For multi-beat scenes, use
[00:00-00:05s]format to specify timing. - Tag References: If files provided, use: "Replicate the camera movement of @video1 while maintaining the visual style of @image1." (lowercase, 1-based index)
- ORDER MATTERS: Tokens at the start define composition; tokens at the end define texture and micro-motion.
- Multi-Image i2v: Provide up to 9 reference images. The model blends aspects (style, identity, environment) across all inputs.
- Audio is mandatory: Seedance 2.0 generates audio natively. Always include an Audio line — music genre/tone, key SFX, ambient texture. Silent direction = random audio.
- Single-beat discipline: Each timed segment = one action. Cramming two narrative beats into 4s degrades physics and motion consistency.
🎭 Capability-Specific Patterns
1. Character Consistency
The man in @Image1 walks tiredly down the hallway, slowing his steps,
finally stopping at his front door. Close-up on his face — he takes a
deep breath, replaces the weariness with a relaxed expression.
Maintain high character consistency, zero facial flicker, persistent clothing details.
2. Camera Movement Replication
Reference @Image1's male character. He is in @Image2's elevator.
Completely reference @Video1's camera movements and facial expressions.
Hitchcock zoom during the fear moment, then orbit shots of the interior.
Elevator doors open, follow shot walking out.
3. Video Extension (Forward)
Extend @Video1 by 10 seconds.
1–5s: Light and shadow slowly slide across table through venetian blinds.
6–10s: A coffee bean drifts down. Camera pushes in toward it until screen goes black.
English text gradually appears — "Lucky Coffee", "Breakfast", "AM 7:00-10:00".
4. Video Extension (Reverse / Prepend)
Extend backward 10s. In warm afternoon light, the camera starts from
the corner with awning fluttering in the breeze, slowly tilting down
to flowers peeking out at the wall base, building anticipation for the main scene.
5. Video Editing (Modify Existing)
Subvert @Video1's plot — the character's expression shifts from warmth to
cold determination. The action is decisive, without hesitation.
Maintain all other visual elements (scene, lighting, timing).
6. Music Beat-Matching
bash scripts/generate-seedance.sh \
--mode i2v \
--file img1.jpg --file img2.jpg --file img3.jpg \
--video-file reference_edit.mp4 \
--audio-file track.mp3 \
--subject "@Image1 @Image2 @Image3 — match the keyframe positions and rhythm of @Video1 for beat-synced cuts. BGM references @Audio1. More dynamic movement, dreamlike visual style." \
--duration 15 --quality high
7. Dialogue / Voice Acting
In the "Cat & Dog Roast Show" — emotionally expressive comedy segment:
Cat host (licking paw, rolling eyes): "Who understands my suffering?"
Dog host (head tilted, tail wagging): "You're one to talk? You sleep 18 hours a day..."
Sound: lively studio ambience, audience laughter, punchy transitions.
8. One-Take / Long Take
@Image1 @Im
Truncated for display — read the full file on GitHub.
Related Skills
Agent-Reach
85.9kGive your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHongShu — one CLI, zero API fees.
headroom
74.0kCompress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
crawl4ai
84.4kOpen-source web crawler and scraper for LLMs and AI agents: any website into clean, LLM-ready Markdown. Run it yourself, or use Crawl4AI Cloud with one key.
Scrapling
84.2k🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: https://discord.gg/EMgGbDceNQ and follow here for daily tips and tricks: https://x.com/Scrapling_dev
Languages
Trust signals
From repository metadata: license, adoption, age and documentation. Not a code audit — see the Safety scan above for what the skill file itself contains.
