Try MiniMax H3 (Text-to-Video) (Image-to-Video) from MiniMax →
Models
Agents
WorkflowsStudioPricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
DISCOVER AI MODELS

Search the full model catalog.

Discover and compare production-ready models. Filter by modality, provider, pricing, and production fit.

27 models1 active filters

Active filters

Wiro AI

Video Generation

Image Generation

Audio & Speech

Realtime Stream

Music Generation

3D Generation

LLM & Chat

Workflow

Partners

1
Need help choosing?

Compare models by capability, cost, and provider fit.

Showing 20 of 27 models
gemini-3.5-flash-lite - AI model cover image

gemini-3.5-flash-lite

google

Gemini 3.5 Flash-Lite is Google’s fast multimodal model for high-volume extraction, classification, and sub-agent tasks, with up to 1,048,576 input tokens.

Partner LLMgoogle
Run
gemini-3.6-flash - AI model cover image

gemini-3.6-flash

google

Gemini 3.6 Flash by Google is a fast, natively multimodal reasoning model for agentic coding and long-context analysis. It takes text plus media and returns text.

Partner LLMgoogle
Run

gemini-omni-flash-r2v

google

Generate 3–10s videos from reference images, or edit an existing clip with one instruction. Built on Google Gemini Omni Flash for coherent motion and style.

Image to Videogoogle
Run

gemini-omni-flash

google

Generate 5s,10s 720p MP4 videos from text, or animate a still image as the opening frame. Google Gemini Omni Flash supports multi-turn, chat-based video edits.

Text to Videogoogle
Run
nano-banana-2-lite - AI model cover image

nano-banana-2-lite

google

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built for rapid iteration and draft assets.

Text to Imagegoogle
Run
gemini-3.5-flash - AI model cover image

gemini-3.5-flash

google

Google’s Gemini 3.5 Flash is a fast multimodal reasoning model built for agents, coding, and long-context analysis. It takes text plus media and returns text.

Partner LLMgoogle
Run
lyria 3 pro - AI model cover image

lyria 3 pro

google

Google’s Lyria 3 Pro generates up to 3-minute songs from detailed prompts and an optional image. It returns 48 kHz stereo MP3 audio with structure.

Google
Run
lyria 3 - AI model cover image

lyria 3

google

Google’s Lyria 3 generates a 30-second, 48 kHz stereo music clip from a detailed description, with an optional image to guide mood and style.

Text to Song
Run
nano-banana-2 - AI model cover image

nano-banana-2

google

An image editing tool designed for quick transformations using reference images and prompts. Supports multi-image mixing and aspect ratio adjustments.

Text to Imagegoogle
Run
gemini-3-pro - AI model cover image

gemini-3-pro

google

Gemini 3 Pro is Google's advanced AI model designed for complex reasoning and natural language understanding tasks.

Partner LLMgoogle
Run
translate-gemma-4b-it-image - AI model cover image

translate-gemma-4b-it-image

google

TranslateGemma is a family of lightweight, state-of-the-art open translation models from Google, based on the Gemma 3 family of models. TranslateGemma models are designed to handle translation tasks across 55 languages. Their relatively small size makes it possible to deploy them in environments with limited resources such as laptops, desktops or your own cloud infrastructure, democratizing access to state of the art translation models and helping foster innovation for everyone.

Fast Inference
Run
translate-gemma-12b-it-image - AI model cover image

translate-gemma-12b-it-image

google

TranslateGemma is a family of lightweight, state-of-the-art open translation models from Google, based on the Gemma 3 family of models. TranslateGemma models are designed to handle translation tasks across 55 languages. Their relatively small size makes it possible to deploy them in environments with limited resources such as laptops, desktops or your own cloud infrastructure, democratizing access to state of the art translation models and helping foster innovation for everyone.

Fast Inference
Run
translate-gemma-27b-it-image - AI model cover image

translate-gemma-27b-it-image

google

TranslateGemma is a family of lightweight, state-of-the-art open translation models from Google, based on the Gemma 3 family of models. TranslateGemma models are designed to handle translation tasks across 55 languages. Their relatively small size makes it possible to deploy them in environments with limited resources such as laptops, desktops or your own cloud infrastructure, democratizing access to state of the art translation models and helping foster innovation for everyone.

Fast Inference
Run
translate-gemma-27b-it - AI model cover image

translate-gemma-27b-it

google

TranslateGemma is a family of lightweight, state-of-the-art open translation models from Google, based on the Gemma 3 family of models. TranslateGemma models are designed to handle translation tasks across 55 languages. Their relatively small size makes it possible to deploy them in environments with limited resources such as laptops, desktops or your own cloud infrastructure, democratizing access to state of the art translation models and helping foster innovation for everyone.

Fast Inference
Run
translate-gemma-12b-it - AI model cover image

translate-gemma-12b-it

google

TranslateGemma is a family of lightweight, state-of-the-art open translation models from Google, based on the Gemma 3 family of models. TranslateGemma models are designed to handle translation tasks across 55 languages. Their relatively small size makes it possible to deploy them in environments with limited resources such as laptops, desktops or your own cloud infrastructure, democratizing access to state of the art translation models and helping foster innovation for everyone.

Fast Inference
Run
translate-gemma-4b-it - AI model cover image

translate-gemma-4b-it

google

TranslateGemma is a family of lightweight, state-of-the-art open translation models from Google, based on the Gemma 3 family of models. TranslateGemma models are designed to handle translation tasks across 55 languages. Their relatively small size makes it possible to deploy them in environments with limited resources such as laptops, desktops or your own cloud infrastructure, democratizing access to state of the art translation models and helping foster innovation for everyone.

Fast Inference
Run
gemini-3-flash - AI model cover image

gemini-3-flash

google

gemini-3-flash

Partner LLMgoogle
Run
gemini-2-5-flash - AI model cover image

gemini-2-5-flash

google

gemini-2-5-flash

Partner LLMgoogle
Run
nano-banana-pro - AI model cover image

nano-banana-pro

google

Google's Gemini 3 Pro Image Preview, also known as Nano Banana, model for text-to-image and image-to-image generation.

Text to Imagegoogle
Run

veo3.1-fast

google

Create high-fidelity 720p/1080p videos with audio using the Veo 3.1 Fast API. Fast integration, flexible pricing, image-to-video, scene control, and seamless creative tools.

Text to Videogoogle
Run
Showing 20 of 27 models

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion