Try Google Gemini Omni Flash Video Generator from Google →
Models
Agents
WorkflowsStudioPricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
DISCOVER AI MODELS

Search the full model catalog.

Discover and compare production-ready models. Filter by modality, provider, pricing, and production fit.

696 models0 active filters

Wiro AI

Video Generation

Image Generation

Audio & Speech

Realtime Stream

Music Generation

3D Generation

LLM & Chat

Workflow

Partners

Need help choosing?

Compare models by capability, cost, and provider fit.

Showing 20 of 696 models

ugc creator v2

wiro

Wiro’s UGC Creator V2 turns a product photo and short script into a 5–10s social ad clip using preset creator scenes and multiple aspect ratios.

Image to Video
Run
gemini-3.5-flash-lite - AI model cover image

gemini-3.5-flash-lite

google

Gemini 3.5 Flash-Lite is Google’s fast multimodal model for high-volume extraction, classification, and sub-agent tasks, with up to 1,048,576 input tokens.

Partner LLMgoogle
Run
gemini-3.6-flash - AI model cover image

gemini-3.6-flash

google

Gemini 3.6 Flash by Google is a fast, natively multimodal reasoning model for agentic coding and long-context analysis. It takes text plus media and returns text.

Partner LLMgoogle
Run
upscaler - AI model cover image

upscaler

google

Google Upscaler enlarges images by 2×, 3×, or 4× while preserving detail. Export as PNG or JPEG, with final resolution capped at 17 megapixels.

Image to Image
Run

gen 4.5

Runway

Runway Gen-4.5 turns text or a first-frame image into 2 to 10s 720p video clips. It follows sequenced actions and camera moves with strong fidelity.

Text to Video
Run
seedream-v5-pro - AI model cover image

seedream-v5-pro

ByteDance

ByteDance Seedream V5 Pro generates images from text and edits a base image using up to 10 references. It’s tuned for infographics, layouts, and clear typography.

Text to ImageByteDance
Run
gpt-5-6-luna - AI model cover image

gpt-5-6-luna

openai

OpenAI’s GPT 5.6 Luna is a fast GPT‑5.6 tier for vision Q&A and chat. Add one or more images plus a prompt to get grounded text answers.

LLM
Run
gpt-5-6-terra - AI model cover image

gpt-5-6-terra

openai

OpenAI GPT 5.6 Terra is a balanced multimodal model for everyday work. Add images and a prompt to get clear, grounded text analysis and writing.

LLM
Run
gpt-5-6-sol - AI model cover image

gpt-5-6-sol

openai

OpenAI’s GPT 5.6 Sol is a flagship reasoning model that can analyze images and produce detailed text answers for coding, research, and visual Q&A.

LLM
Run
grok-4-5 - AI model cover image

grok-4-5

xai

Grok 4.5 is xAI’s flagship reasoning model for coding and long-context work. It supports tool calling and image understanding, returning detailed text answers.

LLM
Run

gemini-omni-flash-r2v

google

Generate 3–10s videos from reference images, or edit an existing clip with one instruction. Built on Google Gemini Omni Flash for coherent motion and style.

Image to Videogoogle
Run

gemini-omni-flash

google

Generate 5s,10s 720p MP4 videos from text, or animate a still image as the opening frame. Google Gemini Omni Flash supports multi-turn, chat-based video edits.

Text to Videogoogle
Run
nano-banana-2-lite - AI model cover image

nano-banana-2-lite

google

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built for rapid iteration and draft assets.

Text to Imagegoogle
Run
remove-background - AI model cover image

remove-background

ideogram

Remove image backgrounds and get a clean transparent PNG cutout. Ideogram’s generative matting keeps hair, glass, and fine text edges crisp.

Image to Imagevideo generatorByteDance
Run

seedance 2.0 mini v2v

ByteDance

ByteDance’s Seedance 2.0 Mini V2V edits short clips from 1–3 reference videos and a prompt. It outputs 480p or 720p MP4 video, with optional synced audio.

Image to Videovideo generatorByteDance
Run

seedance 2.0 mini

ByteDance

Seedance 2.0 Mini by ByteDance generates short MP4 videos from a prompt, with optional start and end frames or reference assets. Add synced audio for social and ads.

Image to Videovideo generatorByteDance
Run

kling-v3-turbo

klingai

Kling V3 Turbo turns a prompt, or a still image plus motion instructions, into a short MP4 video with native audio and lip-sync in 720p or 1080p.

Text to Videovideo generatorklingai
Run

Hero Vox Effects

wiro

Create a short talking-portrait video from a human photo and a script. Choose an effect preset, aspect ratio, and 5, 10, or 15s duration.

Image to Video
Run

happyhorse 1.1 reference

alibaba

Create 3–15s videos from up to 9 reference images plus a scene description. Designed for character and product consistency at 720p or 1080p.

Image to Videovideo generatoralibaba
Run

happyhorse 1.1

alibaba

HappyHorse 1.1 by Alibaba generates 720p or 1080p videos with synced audio from text, or animates a first-frame image into motion.

Text to Videovideo generatoralibaba
Run
Showing 20 of 696 models

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion