Try MiniMax H3 (Text-to-Video) (Image-to-Video) from MiniMax →
Models
Agents
WorkflowsStudioPricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
AI MODELS

Production-ready models,organized for real work.

Browse and compare image, video, audio, and language models from leading providers, all available through the Wiro API.

More filters

Recently Added

711 models in this collection.

View all
h3-r2v
minimax

MiniMax H3 R2V generates 4–15s videos with synced stereo audio from reference images, clips, and optional audio. Use it…

Image to VideoVideo to Video
h3
minimax

MiniMax H3 I2V generates 4–15s videos from a prompt and optional first or last frame. It outputs MP4 video at 24 FPS wi…

Text to VideoImage to Video
flux-3-v2v
blackforestlabs

Extend a source video past its final frame. FLUX 3 V2V generates 5–20 seconds of 720p or 1080p footage, with optional s…

Video to Video
flux-3
blackforestlabs

FLUX 3 I2V turns 1–10 images and a prompt into a 5–20s MP4 clip at 720p or 1080p, with optional native audio and keyfra…

Text to VideoImage to Video
wiro/audio-fade model cover
audio-fade
wiro

Add smooth fade-in and fade-out to any audio file, then export as MP3, WAV, FLAC, AAC, OGG, or Opus with optional bitra…

Utility
wiro/audio-trim model cover
audio-trim
wiro

Cut any audio file into a precise clip using a start time plus end time or duration. Export MP3, WAV, M4A, FLAC, and mo…

Utility
wiro/audio-merge model cover
audio-merge
wiro

Combine up to 10 audio files into one track by concatenating or mixing them, add optional crossfades, and export to com…

Utility
wiro/audio-to-audio-conv model cover
audio-to-audio-conv
wiro

Convert up to 10 audio files per run to MP3, WAV, FLAC, AAC, OGG, or Opus. Choose lossless exports or set a target bitr…

Utility
wiro/video-fade model cover
video-fade
wiro

Add fade-in and fade-out to a video clip with a chosen color. Control fade durations and optionally fade audio for clea…

Utility
wiro/video-loop model cover
video-loop
wiro

Repeat a video clip back-to-back and export one longer file. Video Loop trims inputs over 180s, supports 1–10 extra rep…

Utility
wiro/video-boomerang model cover
video-boomerang
wiro

Turn a short clip into a boomerang that plays forward then backward once. Optionally keep the original audio and choose…

Utility
wiro/video-thumbnail model cover
video-thumbnail
wiro

Extract 1–20 evenly spaced still frames from a video and save them as JPG, PNG, or WebP. Capture a timestamp and resize…

Utility
wiro/video-watermark model cover
video-watermark
wiro

Burn a logo or text watermark into any video with placement presets, opacity, margin, and proportional sizing controls.…

Utility
wiro/video-crop model cover
video-crop
wiro

Crop videos to a fixed pixel box or a preset aspect ratio without scaling. Video Crop by Wiro keeps the largest fitting…

Utility
wiro/video-resize model cover
video-resize
wiro

Resize and reframe videos to new pixel sizes or aspect ratios using padding, crop, stretch, or blurred background fills…

Utility
wiro/video-trim model cover
video-trim
wiro

Wiro Video Trim cuts a continuous segment from a source video using precise timestamps. Set a start time plus an end ti…

Utility
wiro/video-audio-merge model cover
video-audio-merge
wiro

Merge a video with a separate audio track. Replace original sound or mix both tracks, adjust added-audio volume, and lo…

Utility
wiro/video-audio-extract model cover
video-audio-extract
wiro

Extract the audio track from a video and export it as MP3, WAV, FLAC, AAC, OGG, Opus, or M4A. Choose bitrate for smalle…

Utility
wiro/video-audio-remover model cover
video-audio-remover
wiro

Strip all audio tracks from a video and export a silent copy. It preserves the original video stream and keeps the same…

Utility
pruna/p-image-ideogram-custom model cover
p-image-ideogram-custom
pruna

Pruna’s P Image Ideogram Custom generates poster-ready images from a short prompt. It’s tuned for readable text, typogr…

Text to Image

Popular Models

711 models in this collection.

View all
bytedance/seedream-v4-5-uncensored model cover
seedream-v4-5-uncensored
bytedance

Generate high-resolution images using Seedream v4.5 Uncensored. Supports text-to-image and image-to-image transformatio…

Text to ImageImage to Image
kling-v2.6-motion-control
klingai

Generates videos from images and reference videos with motion control. Supports custom prompts and character orientatio…

Image to Video
kling-v3
klingai

Generate high-quality videos from text prompts using Kling V3. Supports custom frames, duration, and aspect ratios.

Text to VideoImage to Video
google/nano-banana-2 model cover
nano-banana-2
google

An image editing tool designed for quick transformations using reference images and prompts. Supports multi-image mixin…

Text to ImageImage to Image
seedance-pro-v1.5-uncensored
bytedance

Seedance Pro v1.5 Uncensored by ByteDance generates short videos from text with optional native audio, strong prompt fo…

Text to VideoImage to Video
gemini-omni-flash-r2v
google

Generate 3–10s videos from reference images, or edit an existing clip with one instruction. Built on Google Gemini Omni…

Image to VideoVideo to Video
gemini-omni-flash
google

Generate 5s,10s 720p MP4 videos from text, or animate a still image as the opening frame. Google Gemini Omni Flash supp…

Text to VideoImage to Video
google/nano-banana-2-lite model cover
nano-banana-2-lite
google

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built f…

Text to ImageImage to Image
google/nano-banana-pro model cover
nano-banana-pro
google

Google's Gemini 3 Pro Image Preview, also known as Nano Banana, model for text-to-image and image-to-image generation.

Text to ImageImage to Image
wiro/Video Converter model cover
Video Converter
wiro

Convert a video to MP4, MOV, WebM, MKV, AVI, MPEG, or M4V with adjustable compression. Built by Wiro for clean exports…

Video to VideoUtility
wiro/Image Converter model cover
Image Converter
wiro

Convert an image to JPEG, PNG, WebP, TIFF, or AVIF. Set output quality from 0 to 100 to balance file size and visual fi…

Image to ImageUtility
panini-card
wiro

Turn a selfie into a Panini-style player card video. Enter name, height, weight, and birth date, then choose a country…

Image to VideoSocial Media & Viral
openai/gpt-image-2-custom model cover
gpt-image-2-custom
openai

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optiona…

Text to ImageImage to Image
wiro/smart resize model cover
smart resize
wiro

Smart Resize by Wiro converts one image into multiple exact sizes, using AI recomposition to keep key subjects in frame…

Image to Image
openai/gpt-image-2 model cover
gpt-image-2
openai

Generate or edit images with GPT Image 2 from OpenAI. It delivers strong instruction following, sharp text rendering, a…

Text to ImageImage to Image
google/lyria 3 pro model cover
lyria 3 pro
google

Google’s Lyria 3 Pro generates up to 3-minute songs from detailed prompts and an optional image. It returns 48 kHz ster…

Text To AudioImage To Audio
google/lyria 3 model cover
lyria 3
google

Google’s Lyria 3 generates a 30-second, 48 kHz stereo music clip from a detailed description, with an optional image to…

Text to SongImage to Song
h3-r2v
minimax

MiniMax H3 R2V generates 4–15s videos with synced stereo audio from reference images, clips, and optional audio. Use it…

Image to VideoVideo to Video
Seedance 2.0
bytedance

Seedance 2.0 by ByteDance generates short MP4 videos from text, with optional image and audio references. It supports s…

Text to VideoImage to Video
h3
minimax

MiniMax H3 I2V generates 4–15s videos from a prompt and optional first or last frame. It outputs MP4 video at 24 FPS wi…

Text to VideoImage to Video

Generate Videos

168 models in this collection.

View all
h3-r2v
minimax

MiniMax H3 R2V generates 4–15s videos with synced stereo audio from reference images, clips, and optional audio. Use it…

Image to VideoVideo to Video
h3
minimax

MiniMax H3 I2V generates 4–15s videos from a prompt and optional first or last frame. It outputs MP4 video at 24 FPS wi…

Text to VideoImage to Video
flux-3-v2v
blackforestlabs

Extend a source video past its final frame. FLUX 3 V2V generates 5–20 seconds of 720p or 1080p footage, with optional s…

Video to Video
flux-3
blackforestlabs

FLUX 3 I2V turns 1–10 images and a prompt into a 5–20s MP4 clip at 720p or 1080p, with optional native audio and keyfra…

Text to VideoImage to Video
ugc creator v2
wiro

Wiro’s UGC Creator V2 turns a product photo and short script into a 5–10s social ad clip using preset creator scenes an…

Image to VideoSocial Media & Viral
gen 4.5
runway

Runway Gen-4.5 turns text or a first-frame image into 2 to 10s 720p video clips. It follows sequenced actions and camer…

Text to VideoImage to Video
gemini-omni-flash-r2v
google

Generate 3–10s videos from reference images, or edit an existing clip with one instruction. Built on Google Gemini Omni…

Image to VideoVideo to Video
gemini-omni-flash
google

Generate 5s,10s 720p MP4 videos from text, or animate a still image as the opening frame. Google Gemini Omni Flash supp…

Text to VideoImage to Video
seedance 2.0 mini v2v
bytedance

ByteDance’s Seedance 2.0 Mini V2V edits short clips from 1–3 reference videos and a prompt. It outputs 480p or 720p MP4…

Image to VideoVideo to Video
seedance 2.0 mini
bytedance

Seedance 2.0 Mini by ByteDance generates short MP4 videos from a prompt, with optional start and end frames or referenc…

Image to VideoVideo to Video
kling-v3-turbo
klingai

Kling V3 Turbo turns a prompt, or a still image plus motion instructions, into a short MP4 video with native audio and…

Text to VideoImage to Video
Hero Vox Effects
wiro

Create a short talking-portrait video from a human photo and a script. Choose an effect preset, aspect ratio, and 5, 10…

Image to VideoSocial Media & Viral
happyhorse 1.1 reference
alibaba

Create 3–15s videos from up to 9 reference images plus a scene description. Designed for character and product consiste…

Image to Video
happyhorse 1.1
alibaba

HappyHorse 1.1 by Alibaba generates 720p or 1080p videos with synced audio from text, or animates a first-frame image i…

Text to VideoImage to Video
Birthday Mode Effects
wiro

Turn your photo into a fun birthday celebration while keeping your face natural and unchanged for instantly share-ready…

Image to VideoImage to Image
Birthday Mode Effects with Caption
wiro

Turn your photo into a fun birthday celebration with your age featured on the scene, while keeping your face natural an…

Image to VideoImage to Image
Hair Style Effects with Caption
wiro

Transform your photo with trending men's and women's hairstyles while keeping your face, expression, and background unc…

Image to VideoSocial Media & Viral
Euphoria Effects
wiro

Transform images into trending Euphoria effects with customizable styles and durations. Apply various creative filters…

Image to VideoSocial Media & Viral
reference-to-video-v6
pixverse

Generate short cinematic videos from up to 3 reference images plus a prompt. PixVerse V6 keeps subjects consistent and…

Image to Video
video-extend-v6
pixverse

Extend an existing clip with PixVerse Video Extend V6. Upload a short MP4 or MOV, describe the continuation, and get a…

Video to Video

Generate Images

239 models in this collection.

View all
pruna/p-image-ideogram-custom model cover
p-image-ideogram-custom
pruna

Pruna’s P Image Ideogram Custom generates poster-ready images from a short prompt. It’s tuned for readable text, typogr…

Text to Image
pruna/p-image-ideogram model cover
p-image-ideogram
pruna

P-Image Ideogram by Pruna generates high-quality text-to-image outputs with strong typography, offering 1K or 2K resolu…

Text to Image
bytedance/seedream-v5-pro model cover
seedream-v5-pro
bytedance

ByteDance Seedream V5 Pro generates images from text and edits a base image using up to 10 references. It’s tuned for i…

Text to ImageImage to Image
google/nano-banana-2-lite model cover
nano-banana-2-lite
google

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built f…

Text to ImageImage to Image
microsoft/lens model cover
lens
microsoft

A foundational text-to-image model designed for efficient, high-resolution image generation with strong prompt followin…

Text to ImageFast Inference
openai/gpt-image-2-custom model cover
gpt-image-2-custom
openai

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optiona…

Text to ImageImage to Image
sensenova/U1-8B-Text-to-Image model cover
U1-8B-Text-to-Image
sensenova

SenseNova U1-8B turns text into high-detail images with strong layout control and clearer in-image text. Use it for pos…

Text to ImageFast Inference
openai/gpt-image-2 model cover
gpt-image-2
openai

Generate or edit images with GPT Image 2 from OpenAI. It delivers strong instruction following, sharp text rendering, a…

Text to ImageImage to Image
xai/grok-imagine-image model cover
grok-imagine-image
xai

xAI’s Grok Imagine Image generates images from prompts and can edit a source image using instructions. Pick an aspect r…

Text to ImageImage to Image
openai/gpt-image-1-5 model cover
gpt-image-1-5
openai

Generate or edit images using text prompts or image edits with GPT Image 1.5. Supports multiple sizes, formats, and qua…

Text to ImageImage to Image
bytedance/seedream-v4-5-uncensored model cover
seedream-v4-5-uncensored
bytedance

Generate high-resolution images using Seedream v4.5 Uncensored. Supports text-to-image and image-to-image transformatio…

Text to ImageImage to Image
fireredteam/FireRed-Image-Edit-1.1 model cover
FireRed-Image-Edit-1.1
fireredteam

FireRed-Image-Edit-1.1 significantly enhances identity consistency, multi-image conditioning, and domain-specialized ed…

Text to ImageImage to Image
bytedance/seedream-v5-lite-uncensored model cover
seedream-v5-lite-uncensored
bytedance

Generate high-quality images from text prompts or image inputs using the Seedream v5 Lite Uncensored model. Supports mu…

Text to ImageImage to Image
google/nano-banana-2 model cover
nano-banana-2
google

An image editing tool designed for quick transformations using reference images and prompts. Supports multi-image mixin…

Text to ImageImage to Image
bytedance/seedream-v5-lite model cover
seedream-v5-lite
bytedance

Generate high-quality images using Seedream V5 Lite, supporting both image-to-image and text-to-image transformations w…

Text to ImageImage to Image
fireredteam/FireRed-Image-Edit model cover
FireRed-Image-Edit
fireredteam

FireRed-Image-Edit is a general-purpose image editing model that delivers high-fidelity and consistent editing across a…

Text to ImageImage to Image
zai-org/GLM-IMAGE model cover
GLM-IMAGE
zai-org

GLM-Image is an image generation model adopts a hybrid autoregressive + diffusion decoder architecture. In general imag…

Text to ImageImage to Image
black-forest-labs/FLUX.2-klein-base-9B model cover
FLUX.2-klein-base-9B
black-forest-labs

FLUX.2 [klein] 9B Base is a 9 billion parameter rectified flow transformer capable of generating images from text descr…

Text to ImageImage to Image
black-forest-labs/FLUX.2-klein-base-4B model cover
FLUX.2-klein-base-4B
black-forest-labs

FLUX.2 [klein] 4B Base is a 4 billion parameter rectified flow transformer capable of generating images from text descr…

Text to ImageImage to Image
black-forest-labs/FLUX.2-klein-4B model cover
FLUX.2-klein-4B
black-forest-labs

FLUX.2 [klein] 4B is a 4 billion parameter rectified flow transformer capable of generating images from text descriptio…

Text to ImageImage to Image

Edit Images

131 models in this collection.

View all
bytedance/seedream-v5-pro model cover
seedream-v5-pro
bytedance

ByteDance Seedream V5 Pro generates images from text and edits a base image using up to 10 references. It’s tuned for i…

Text to ImageImage to Image
briaai/RMBG-2.0 model cover
RMBG-2.0
briaai

RMBG-2.0 removes image backgrounds with an 8-bit alpha matte for smooth edges. Built by BRIA AI for production cutouts…

Image to ImageEcommerce
google/nano-banana-2-lite model cover
nano-banana-2-lite
google

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built f…

Text to ImageImage to Image
ideogram/remove-background model cover
remove-background
ideogram

Remove image backgrounds and get a clean transparent PNG cutout. Ideogram’s generative matting keeps hair, glass, and f…

Image to ImageIdeogram
Birthday Mode Effects
wiro

Turn your photo into a fun birthday celebration while keeping your face natural and unchanged for instantly share-ready…

Image to VideoImage to Image
Birthday Mode Effects with Caption
wiro

Turn your photo into a fun birthday celebration with your age featured on the scene, while keeping your face natural an…

Image to VideoImage to Image
pruna/p-image-try-on model cover
p-image-try-on
pruna

Pruna’s p-image-try-on creates virtual try-on images by applying one or more garment photos to a person photo. Generate…

Image to ImageFast Inference
nvidia/LocateAnything-3B model cover
LocateAnything-3B
nvidia

LocateAnything 3B by Nvidia localizes objects, UI elements, and text in images. It returns normalized boxes or points f…

Image to ImageFast Inference
wiro/Image Converter model cover
Image Converter
wiro

Convert an image to JPEG, PNG, WebP, TIFF, or AVIF. Set output quality from 0 to 100 to balance file size and visual fi…

Image to ImageUtility
wiro/smart resize model cover
smart resize
wiro

Smart Resize by Wiro converts one image into multiple exact sizes, using AI recomposition to keep key subjects in frame…

Image to Image
sensenova/U1-8B-Interleave model cover
U1-8B-Interleave
sensenova

u1-8b-interleave by SenseNova generates step-by-step text with matching images, optionally conditioned on up to 5 refer…

Image to ImageImage to Text
openai/gpt-image-2-custom model cover
gpt-image-2-custom
openai

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optiona…

Text to ImageImage to Image
openai/gpt-image-2 model cover
gpt-image-2
openai

Generate or edit images with GPT Image 2 from OpenAI. It delivers strong instruction following, sharp text rendering, a…

Text to ImageImage to Image
xai/grok-imagine-image model cover
grok-imagine-image
xai

xAI’s Grok Imagine Image generates images from prompts and can edit a source image using instructions. Pick an aspect r…

Text to ImageImage to Image
openai/gpt-image-1-5 model cover
gpt-image-1-5
openai

Generate or edit images using text prompts or image edits with GPT Image 1.5. Supports multiple sizes, formats, and qua…

Text to ImageImage to Image
wiro/Instagram Pose Multi model cover
Instagram Pose Multi
wiro

Generate stylish Instagram-style pose images with trendy angles, natural expressions, and a modern aesthetic. Built by…

Image to ImageSocial Media & Viral
bytedance/seedream-v4-5-uncensored model cover
seedream-v4-5-uncensored
bytedance

Generate high-resolution images using Seedream v4.5 Uncensored. Supports text-to-image and image-to-image transformatio…

Text to ImageImage to Image
fireredteam/FireRed-Image-Edit-1.1 model cover
FireRed-Image-Edit-1.1
fireredteam

FireRed-Image-Edit-1.1 significantly enhances identity consistency, multi-image conditioning, and domain-specialized ed…

Text to ImageImage to Image
bytedance/seedream-v5-lite-uncensored model cover
seedream-v5-lite-uncensored
bytedance

Generate high-quality images from text prompts or image inputs using the Seedream v5 Lite Uncensored model. Supports mu…

Text to ImageImage to Image
google/nano-banana-2 model cover
nano-banana-2
google

An image editing tool designed for quick transformations using reference images and prompts. Supports multi-image mixin…

Text to ImageImage to Image

Generate Text

27 models in this collection.

View all
sensenova/U1-8B-Interleave model cover
U1-8B-Interleave
sensenova

u1-8b-interleave by SenseNova generates step-by-step text with matching images, optionally conditioned on up to 5 refer…

Image to ImageImage to Text
sensenova/U1-8B-Visual-Understanding model cover
U1-8B-Visual-Understanding
sensenova

SenseNova U1 8B Visual Understanding creates infographic-style images and prompt-based edits from text and an optional…

Image to TextFast Inference
nvidia/parakeet-tdt-0.6b-v3 model cover
parakeet-tdt-0.6b-v3
nvidia

Multilingual speech-to-text for 25 European languages with auto language detection, punctuation, capitalization, and op…

Speech to TextFast Inference
coherelabs/cohere-transcribe-03-2026 model cover
cohere-transcribe-03-2026
coherelabs

CohereLabs cohere-transcribe-03-2026 is a 2B Conformer speech-to-text model for 14 languages. It creates accurate trans…

Speech to TextFast Inference
qwen/Qwen3-ASR-1.7B model cover
Qwen3-ASR-1.7B
qwen

A lightweight speech-to-text model optimized for fast inference. Converts audio input into text with support for multip…

Speech to TextFast Inference
mistralai/Voxtral-Mini-4B-Realtime-2602 model cover
Voxtral-Mini-4B-Realtime-2602
mistralai

Voxtral Mini 4B Realtime 2602 is a multilingual, realtime speech-transcription model and among the first open-source so…

Speech to TextRealtime STT
nvidia/nemotron model cover
nemotron
nvidia

Nemotron-Speech-Streaming-En-0.6b is the first unified model in the Nemotron Speech family, engineered to deliver high-…

Speech to TextAudio
elevenlabs/speech-to-text model cover
speech-to-text
elevenlabs

Speech to text model from ElevenLabs

Speech to Text
moondream3-preview/detect model cover
detect
moondream3-preview

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Image to TextBf16
moondream3-preview/point model cover
point
moondream3-preview

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Image to TextBf16
moondream3-preview/caption model cover
caption
moondream3-preview

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Image to TextBf16
moondream3-preview/query model cover
query
moondream3-preview

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Image to TextBf16
openai/whisper-large-v3-turbo-turkish model cover
whisper-large-v3-turbo-turkish
openai

Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of…

Speech to TextWhisper
wiro/video-nsfw-detection model cover
video-nsfw-detection
wiro

NSFW video detection automatically analyzes video content to identify inappropriate or explicit material, ensuring comp…

Video to Text
VideoLLaMA3-2B
damo-nlp-sg

VideoLLaMA3-2B is a model designed for video understanding.

Video to Text
VideoLLaMA3-2B-Image
damo-nlp-sg

VideoLLaMA3-2B-Image is a model designed for image understanding.

Image to Text
wiro/VideoLLaMA3-7B-Image model cover
VideoLLaMA3-7B-Image
wiro

VideoLLaMA3-7B-Image is a model designed for image understanding.

Image to Text
wiro/VideoLLaMA3-7B model cover
VideoLLaMA3-7B
wiro

VideoLLaMA3-7B is a model designed for video understanding.

Video to Text
salesforce/blip2-flan-t5-xl model cover
blip2-flan-t5-xl
salesforce

BLIP-2 creates captions or detailed descriptions for images. This is BLIP-2 model, leveraging Flan T5-xl.

Image to Text
salesforce/blip-image-captioning-large model cover
blip-image-captioning-large
salesforce

BLIP is a model that is able to perform various multi-modal tasks including visual question answering and image caption…

Image to Text

Generate 3D

4 models in this collection.

View all
tencentarc/Pixal3D model cover
Pixal3D
tencentarc
3D Generation
HY-World-2.0-World-Reconstruction
tencent
3D Generation
microsoft/Trellis-2 model cover
Trellis-2
microsoft
3D Generation
tencent/Hunyuan3D-2.1 model cover
Hunyuan3D-2.1
tencent
3D Generation

Generate Audio

26 models in this collection.

View all
openmoss/MOSS-TTS-v1.5 model cover
MOSS-TTS-v1.5
openmoss
Text to Speech
openbmb/VoxCPM2 model cover
VoxCPM2
openbmb
Text to Speech
k2-fsa/OmniVoice model cover
OmniVoice
k2-fsa
Text to Speech
humeai/tada-3b-ml model cover
tada-3b-ml
humeai
Text to Speech
fishaudio/s2-pro model cover
s2-pro
fishaudio
Text to Speech
nineninesix/kani-tts-2-en model cover
kani-tts-2-en
nineninesix
Text to Speech
resemble-ai/chatterbox-turbo model cover
chatterbox-turbo
resemble-ai
Text to Speech
resemble-ai/chatterbox-multilingual  model cover
chatterbox-multilingual
resemble-ai
Text to Speech
openmoss/MOSS-TTSD model cover
MOSS-TTSD
openmoss
Text to Speech
openmoss/MOSS-TTS-Realtime model cover
MOSS-TTS-Realtime
openmoss
Text to Speech
elevenlabs/Realtime Conversational AI model cover
Realtime Conversational AI
elevenlabs
Speech to Speech
openai/gpt-realtime-mini model cover
gpt-realtime-mini
openai
Speech to Speech
openai/gpt-realtime model cover
gpt-realtime
openai
Speech to Speech
qwen/Qwen3-TTS-12Hz-1.7B model cover
Qwen3-TTS-12Hz-1.7B
qwen
Text to Speech
nvidia/PersonaPlex-Realtime model cover
PersonaPlex-Realtime
nvidia
Speech to Speech
microsoft/VibeVoice-Realtime model cover
VibeVoice-Realtime
microsoft
Text to Speech
openbmb/VoxCPM model cover
VoxCPM
openbmb
Text to Speech
google/gemini-2.5-tts model cover
gemini-2.5-tts
google
Text to Speech
elevenlabs/text-to-speech model cover
text-to-speech
elevenlabs
Text to Speech
Faceless-Video-Generator
wiro
Text to Video

Generate Music

17 models in this collection.

View all
stabilityai/stable-audio-3-small-sfx model cover
stable-audio-3-small-sfx
stabilityai
Text to Music
stabilityai/stable-audio-3-small-music model cover
stable-audio-3-small-music
stabilityai
Text to Music
stabilityai/stable-audio-3-medium model cover
stable-audio-3-medium
stabilityai
Text to Music
google/lyria 3 model cover
lyria 3
google
Text to Song
tencent-ailab/SongGeneration 2 model cover
SongGeneration 2
tencent-ailab
Music Generation
wiro/video-background-music-v2 model cover
video-background-music-v2
wiro
Video to Video
ace-step/text-to-song-ACE-Step1.5 model cover
text-to-song-ACE-Step1.5
ace-step
Music Generation
wiro/Song Frame model cover
Song Frame
wiro
Image to Video
Faceless-Video-Generator
wiro
Text to Video
video-background-music-gen
wiro
Video to Video
ace-step/image-to-song-ACE-Step-v1-3.5B model cover
image-to-song-ACE-Step-v1-3.5B
ace-step
Music Generation
ace-step/text-to-song-ACE-Step-v1-3.5B model cover
text-to-song-ACE-Step-v1-3.5B
ace-step
Music Generation
wiro/image-to-song-with-reference-YuE model cover
image-to-song-with-reference-YuE
wiro
Music Generation
wiro/image-to-song-YuE model cover
image-to-song-YuE
wiro
Music Generation
wiro/text-to-song-with-reference-YuE model cover
text-to-song-with-reference-YuE
wiro
Music Generation
wiro/text-to-song-YuE model cover
text-to-song-YuE
wiro
Music Generation
wiro/music_gen model cover
music_gen
wiro
Music Generation

Realtime Stream

7 models in this collection.

View all
openbmb/VoxCPM2 model cover
VoxCPM2
openbmb
Text to Speech
mistralai/Voxtral-Mini-4B-Realtime-2602 model cover
Voxtral-Mini-4B-Realtime-2602
mistralai
Speech to Text
openmoss/MOSS-TTS-Realtime model cover
MOSS-TTS-Realtime
openmoss
Text to Speech
elevenlabs/Realtime Conversational AI model cover
Realtime Conversational AI
elevenlabs
Speech to Speech
openai/gpt-realtime-mini model cover
gpt-realtime-mini
openai
Speech to Speech
openai/gpt-realtime model cover
gpt-realtime
openai
Speech to Speech
nvidia/PersonaPlex-Realtime model cover
PersonaPlex-Realtime
nvidia
Speech to Speech

LLM & Chat

82 models in this collection.

View all
google/gemini-3.5-flash-lite model cover
gemini-3.5-flash-lite
google
Partner LLM
google/gemini-3.6-flash model cover
gemini-3.6-flash
google
Partner LLM
openai/gpt-5-6-luna model cover
gpt-5-6-luna
openai
LLM
openai/gpt-5-6-terra model cover
gpt-5-6-terra
openai
LLM
openai/gpt-5-6-sol model cover
gpt-5-6-sol
openai
LLM
xai/grok-4-5 model cover
grok-4-5
xai
LLM
openai/gpt-5.4-nano model cover
gpt-5.4-nano
openai
Partner LLM
openai/gpt-5.4-mini model cover
gpt-5.4-mini
openai
Partner LLM
openai/gpt-5.4 model cover
gpt-5.4
openai
Partner LLM
openai/gpt-5.4-pro model cover
gpt-5.4-pro
openai
Partner LLM
openai/gpt-5.5 model cover
gpt-5.5
openai
Partner LLM
openai/gpt-5.5-pro model cover
gpt-5.5-pro
openai
Partner LLM
google/gemini-3.5-flash model cover
gemini-3.5-flash
google
Partner LLM
qwen/Qwen3.6-27B model cover
Qwen3.6-27B
qwen
Chat
xai/grok-4-1-fast model cover
grok-4-1-fast
xai
LLM
xai/grok-4-20 model cover
grok-4-20
xai
LLM
qwen/Qwen3.5-4B-heretic model cover
Qwen3.5-4B-heretic
qwen
Chat
qwen/Qwen3.5-9B-heretic model cover
Qwen3.5-9B-heretic
qwen
Chat
qwen/Qwen3.5-4B model cover
Qwen3.5-4B
qwen
Chat
qwen/Qwen3.5-9B model cover
Qwen3.5-9B
qwen
Chat

AI Models for E-commerce

17 models in this collection.

View all
ugc creator v2
wiro
Image to Video
briaai/RMBG-2.0 model cover
RMBG-2.0
briaai
Image to Image
ugc creator
wiro
Image to Video
wiro/Shopify Template model cover
Shopify Template
wiro
Image to Image
Product Studio
wiro
Image to Video
Product with Model
wiro
Image to Video
wiro/Virtual Try-On-V2 model cover
Virtual Try-On-V2
wiro
Image to Video
Animated Logo
wiro
Image to Video
3D Text Animations
wiro
Text to Video
Product Ads with Caption
wiro
Image to Video
Product Ads with Logo
wiro
Image to Video
Product Ads
wiro
Image to Video
wiro/camera-angle-editor model cover
camera-angle-editor
wiro
Image to Image
wiro/Product Photoshoot model cover
Product Photoshoot
wiro
Image to Image
wiro/Virtual Try-On model cover
Virtual Try-On
wiro
Image to Image
wiro/text-removal model cover
text-removal
wiro
Image to Image
wiro/remove-background model cover
remove-background
wiro
Image to Image

AI Models for Social Media Creators

64 models in this collection.

View all
ugc creator v2
wiro
Image to Video
Hero Vox Effects
wiro
Image to Video
Birthday Mode Effects
wiro
Image to Video
Birthday Mode Effects with Caption
wiro
Image to Video
Hair Style Effects with Caption
wiro
Image to Video
Euphoria Effects
wiro
Image to Video
World Cup 2026 Effects
wiro
Image to Video
World Cup 2026 Effects with Caption
wiro
Image to Video
Scream Effects
wiro
Image to Video
panini-card
wiro
Image to Video
Sport Trend Effects
wiro
Image to Video
Insta Hot Girl Effects
wiro
Image to Video
Queer Editorial Effects
wiro
Image to Video
Tiktok Trend Effects
wiro
Image to Video
Wildlife Documentary Effect
wiro
Image to Video
Transformation Effect
wiro
Image to Video
Supernatural Presence Effect
wiro
Image to Video
Superhero Powers Effect
wiro
Image to Video
Sports Extreme Effect
wiro
Image to Video
Scale Shift Effect
wiro
Image to Video
Built for Production

Everything you need
to ship with confidence.

Wiro is the production layer between your product and every model.

One unified API

Same auth, same format, every model.

Transparent pricing

Know your costs before you run.

Run history & files

Inputs, outputs, and cost logged.

Webhooks & workflows

Automate completions and build at scale.

Explore the full catalog

Hundreds of models.
One place to find them.

Filter by modality, provider, pricing, context, and more.

Advanced filters

Pinpoint the right model.

Compare side by side

Evaluate quality, cost, and performance.

Save and organize

Bookmark favorites and build collections.

View all models

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion