Try Google Gemini Omni Flash Video Generator from Google →
Models
Agents
WorkflowsStudioPricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
Explore models/
Video GenerationActive

alibaba / Wan 2.7 Reference

Wan 2.7 Reference

byalibaba

Generate 2-10s videos in 720p or 1080p from reference images or a short clip. Wan 2.7 R2V keeps subject identity while following your prompt.

Image to VideoVideo to Video
Model ID
Wan 2.7 Reference
Provider
alibaba
Updated
1777299058
14
Comments
Average rating : 5 (15 users)
Provideralibaba
ModelWan 2.7 Reference
Image to VideoVideo to Video
wiro playground—alibaba/Wan 2.7 Reference
Reset to defaults

Required. 'videoedit' edits an input video. 'r2v' generates a video from reference videos and/or reference images.

Delete All
0 / 5
Maximum 5 video allowed
Drop video to upload

OR

Click to browse your device

Supports: MP4, WEBM, MOV

In videoedit mode: required, single input video to edit. In r2v mode: optional, up to 5 reference videos. Max 100 MB each.

Delete All
0 / 5
Maximum 5 image allowed
Drop image to upload

OR

Click to browse your device

Supports: JPG, JPEG, PNG, GIF, WEBP, HEIC

Optional in videoedit (up to 4). In r2v: required if no input video is provided; total references + reference videos cannot exceed 5.

Delete All
0 / 5
Maximum 5 audio allowed
Drop audio to upload

OR

Click to browse your device

Supports: MP3, WAV, M4A, WEBM, OPUS

Optional, r2v only. Voice samples mapped 1:1 (by order) to reference videos then reference images. Ignored in videoedit mode.

Required. Describes the elements and visual characteristics for the generated video. Up to 5,000 characters.

Sample outputs
Updated 1777299058
## Overview Wan 2.7 R2V by Alibaba generates short videos from visual references plus a written prompt. It uses your reference images or clips as identity anchors, then creates new motion and camera moves that match your description. You can also edit an existing clip by describing changes in plain words, with an optional reference image for guided edits. This helps you keep the same character or product look across scenes without manual keyframing. ## What you can build - Character-consistent clips from 1 to 3 reference images - Product shots that keep packaging and color details stable across scenes - Brand mascot animation that stays recognizable across multiple generations - Instruction-based edits of a short clip, like style changes or scene rewrites - Element replacement using an extra reference image, like swapping clothes or props - Social-friendly vertical videos by selecting a 9:16 aspect ratio ## Inputs - A required text description of the scene, action, and camera behavior, up to 5,000 characters - An optional negative description of what must not appear, up to 500 characters - A source video to edit when you want video editing mode (MP4 or MOV, 2 to 10 seconds, up to 100 MB) - Reference images when you want reference-to-video mode (JPEG/JPG/PNG/WEBP). You can provide up to 3 images. - A target output duration between 2 and 10 seconds - A resolution tier selection: 720p or 1080p - An aspect ratio choice (16:9, 9:16, 1:1, 4:3, 3:4), or auto selection based on your first reference image. When you edit a video, the aspect ratio follows the input clip. - An audio choice when a source video is provided: keep the original audio or generate a new audio track - An optional prompt expansion toggle that rewrites short prompts into longer scene directions - An optional watermark toggle that adds an "AI Generated" mark in the lower-right corner - An optional seed number for repeatable results (0 to 2,147,483,647) ## Outputs - A generated MP4 video file - Video metadata that describes the clip, such as width, height, fps, duration, and frame count - The seed used for the run, so you can recreate or iterate on the same motion - When prompt expansion is enabled, the final rewritten prompt may be returned alongside the video ## Recommended settings - For the most stable identity match, use 1 clear reference image and set duration to 2 to 5 seconds - For maximum detail, pick 1080p when your scene has faces, hands, or product close-ups - Use multiple references only when they show the same subject with consistent styling ## Limitations - Identity can drift in busy scenes, especially with multiple moving subjects - Fine details like small text, logos, and tiny accessories may change across frames - Long or highly choreographed actions can reduce resemblance to the references - Conflicting references can confuse the model and cause mixed features - Low-quality inputs reduce results. Blurry images, heavy compression, fast cuts, and noisy frames can cause flicker and artifacts. - Editing mode only accepts short clips within the supported duration and file size limits ## Safety & compliance - This model can apply content moderation to both inputs and outputs, which may block some requests - Only upload videos and images you have rights to use - Don’t generate or edit content that impersonates real people without consent, especially for deceptive or harmful uses - Follow Wiro’s content rules for prohibited or restricted content

Example prompts

Great starting points for Wan 2.7 Reference.

A woman seen from behind with long dark hair blowing in the wind, wearing a long white dress and wrapped tightly in a thick olive-green blanket, standing on the edge of a dark, charred cliff. Below her, a vast, bubbling sea of glowing red and orange lava stretches to the horizon. A hellish atmosphere with a dark, ominous sky choked with thick black smoke and falling grey ash. The fiery glow of the magma illuminates her from below, intense heat distortion in the air.Video Generation
An aerial drone shot flies forward, following a winding asphalt road with double yellow lines as it recedes deep into a canyon. The previously rocky, sparse mountainsides are now a dense, lush, vibrant green forest. Towering slopes are completely covered in varied green trees—conifers, oaks, and dense undergrowth—creating a rich, verdant tapestry. The canyon is now entirely filled with thriving green vegetation. The winding road cuts a clear path through the sea of green. Clear, bright daylight illuminates the vibrant colors of the forest. The perspective is maintained as the camera moves over the road, into the heart of the lush green mountain pass.Video Generation
Insert the astronaut from image 1 and the Earth from image 2 into the provided starfield video. Place the astronaut in the lower-left foreground and the Earth in the upper-right background. Animate the astronaut with a very subtle, slow-motion floating effect while the background stars continue their timelapse movement. Ensure the lighting on the astronaut and Earth is consistent with a distant sun.Video Generation
Three lions from the reference image, including a majestic male with a dark mane and a white male lion, standing in the shallow waters of a fast-flowing, wide river. The river water rushes around their legs. The background is a dense, lush green tropical forest with palm trees under a bright, sunny sky. Realistic sunlight reflecting off the water and the lions\' fur.Video Generation

API quick start

Run Wan 2.7 Reference with a single API call.

POST https://api.wiro.ai/v1/Run/alibaba/Wan 2.7 Reference
{
  "prompt": "A woman seen from behind with long dark h…",
  "mode": "videoedit",
  "negativePrompt": "...",
  "duration": 5
}
View full API docs

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion