Try Google Gemini Omni Flash Video Generator from Google →
Models
Agents
WorkflowsStudioPricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
Explore models/
Video GenerationActive

ByteDance / Seedance 2.0 Fast

Seedance 2.0 Fast

bybytedance

Seedance 2.0 Fast by ByteDance generates short MP4 videos faster and cheaper than Seedance 2.0, with the same multimodal capabilities, limited to 480p / 720p output.

Text to VideoImage to Video
Model ID
Seedance 2.0 Fast
Provider
bytedance
Updated
1781696229
10
Comments
Average rating : 5 (11 users)
Providerbytedance
ModelSeedance 2.0 Fast
Text to VideoImage to Video
wiro playground—bytedance/Seedance 2.0 Fast
Reset to defaults

Optional. The prompt to generate the video.

0 / 1
Maximum 1 image allowed
Drop image to upload

OR

Click to browse your device

Supports: JPG, JPEG, PNG, GIF, WEBP, HEIC

Optional. The provider may sometimes reject human faces. First frame of the video (image-to-video).

Sample outputs
Click To See More Samples
Updated 1781696229

Seedance 2.0 Fast is the faster, lower-cost variant of Seedance 2.0. It shares the same multimodal capabilities (text-to-video, image-to-video, first-and-last frame, and reference-to-video) but is limited to 480p and 720p output (1080p is not supported).

## Overview Seedance 2.0 Fast is ByteDance’s multimodal video generation model. It generates video and audio together, so sound effects and dialogue timing match the visuals. You describe a scene and the model produces a short, edited-feeling clip with strong motion stability and controllable camera language. This helps teams prototype ads, story beats, and visual concepts without a shoot. ## What you can build - Text-to-video scenes with camera moves like tracking shots, push-ins, and POV switches - Image-to-video animations that start from a provided first frame - First-to-last frame transitions for outfit changes, product reveals, and before/after shots - Reference-driven clips that keep a character or style consistent across a short sequence - Music- or rhythm-guided visuals using reference audio as a timing and mood anchor - Social video loops for Reels, Shorts, and TikTok in vertical or square formats ## Inputs - A scene description written as plain text. Include actions, setting, lighting, and camera direction. - Optional spoken lines written inside double quotes if you want dialogue with lip-synced audio. - Optional first-frame image to anchor the opening shot for image-to-video. Use common web image formats such as JPG, PNG, WebP, GIF, or AVIF. - Optional last-frame image to guide how the clip should end. This works only when you also provide a first frame. - Optional reference image set (up to 9 images) to lock in character appearance, wardrobe, props, or visual style. In your description, call them out as Image 1, Image 2, and so on. - Optional reference audio clips (up to 3 files) in WAV or MP3 to guide mood, pacing, and sound design. Some workflows require at least one visual reference when you include audio. - A target output resolution choice. Higher settings can increase detail, but availability depends on the deployment. - An output aspect ratio choice, or an adaptive option that lets the model pick based on your prompt and references. - A clip duration choice of 5, 10, or 15 seconds. - A toggle to include or omit generated audio. - An optional watermark toggle. - An optional numeric seed to reproduce results. Set it to 0 to randomize. ## Outputs The model returns a rendered MP4 video. If audio is enabled, the MP4 includes a synchronized audio track that can contain ambience, sound effects, music, and dialogue. When you provide reference images, the output typically preserves key identity cues like clothing and overall look within the clip’s length. ## Recommended settings - First tests and prompt iteration: generate 5 seconds at 720p with adaptive aspect ratio. - Dialogue scenes: keep spoken lines short and wrap them in double quotes. - Image-to-video with a specific ending: provide both a first frame and a last frame, then describe the transition. - Character consistency: use 3–9 reference images and explicitly refer to them in the description. - Repeatable results for reviews: reuse the same seed and keep references identical. ## Limitations - Single generations are limited to short clips, with an upper bound of 15 seconds. - Resolution support varies by provider and endpoint. Fast is limited to 480p and 720p (1080p not supported). - Face and likeness controls can be strict. Uploads that include real human faces may be blocked or require authorization. - Requests that reference copyrighted characters, brand IP, or public figures may be filtered or refused. - Reference quality matters. Blurry images, heavy compression, screenshots with UI overlays, or inconsistent reference sets can cause identity drift and visual artifacts. - Audio quality can vary by prompt. Very dense sound instructions can introduce distortion or muddiness. ## Safety & compliance - Only upload images, audio, and prompts you have rights to use. - Don’t generate content that impersonates real people without consent and legal permission. - Avoid prompts that ask for copyrighted characters, trademarked brands, or scenes copied from films. - Use the model for lawful, non-deceptive content. Don’t create misleading deepfakes or fake news footage.

Example prompts

Great starting points for Seedance 2.0 Fast.

A lone lighthouse stands on a jagged cliff as a violent storm hurls waves forty feet into the air, its beam cutting through sheets of rain in rotating golden arcs. The camera begins tight on the rain-streaked glass at the top of the tower, then pulls back and spirals down and around the structure revealing the full fury of the ocean. Lightning fractures the sky behind, briefly silhouetting the lighthouse. Moody storm palette of slate blue and warm amber beam, cinematic anamorphic lens, filmic grain, roaring wind, thunder cracks and crashing surf audio.Video Generation
A bustling night market in Bangkok glows with strings of warm yellow bulbs and red paper lanterns, steam rising from dozens of street-food woks. The camera drifts forward in a smooth handheld push between stalls, a wok erupts in a sudden burst of orange flame just as the lens passes, then it whip-pans to a smiling vendor tossing noodles. Motorbike headlights streak through the background. Rich saturated palette of ember orange and deep green, cinematic 35mm feel, sizzling oil, chatter and distant tuk-tuk horns audio.Video Generation
A flamenco dancer in a blood-red ruffled dress stamps her heel on a wooden stage in a candlelit Spanish tablao, a burst of dust rising in slow motion around her ankles. The camera begins low at floor level catching the stomp, then cranes upward and orbits her as the skirt fans out in a dramatic spiral. A single guitarist sits in shadow behind her. Theatrical chiaroscuro lighting, deep crimson and warm gold palette, filmic shallow depth of field, sharp heel strikes, flamenco guitar and rhythmic hand claps audio.Video Generation
A caravan of camels traverses endless orange sand dunes at the last light of sunset, their silhouettes stretched long across rippled sand. The camera tracks parallel at low height watching the hooves sink softly, then cranes high into a sweeping aerial revealing the full procession winding like a serpent toward distant purple mountains. A crescent moon rises. Lawrence-of-Arabia cinematography, warm burnt-orange to violet gradient palette, anamorphic lens, soft hoofbeats, bells and gentle desert wind audio.Video Generation
A figure skater carves elegant spirals across a perfectly smooth frozen lake at dawn, each blade stroke spraying fine ice crystals that catch the pink morning light. The camera starts in a tight low-angle tracking shot following the blades, then lifts into a sweeping overhead spiral matching her spin as she launches into a triple axel. Snow-dusted pines ring the lake. Ethereal cold palette of pastel pink and icy blue, cinematic slow motion, soft classical strings and crisp blade-on-ice audio.Video Generation

API quick start

Run Seedance 2.0 Fast with a single API call.

POST https://api.wiro.ai/v1/Run/ByteDance/Seedance 2.0 Fast
{
  "prompt": "A lone lighthouse stands on a jagged clif…",
  "resolution": "720p",
  "ratio": "adaptive",
  "duration": 4
}
View full API docs

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion