Production-ready models,organized for real work.
Browse and compare image, video, audio, and language models from leading providers, all available through the Wiro API.



Recently Added
523 models in this collection.
FastH3 V2 is an eight-forward distilled MiniMax H3 model that generates synchronized video and stereo audio from a text…
Create 2–30s reference-guided videos from a prompt plus up to 10 images, 5 clips, and 5 audio tracks. Export 1080p, 2K,…
Generate 5–30 second videos from a written shot prompt, with optional first and last frame keyframes, audio on/off, and…
Generate 2–30s 480p–1080p MP4 videos from a prompt plus up to 10 reference images, with optional reference video and au…
Create 5–30s MP4 videos from a prompt, optionally anchored by first and last frame images with optional audio. Powered…
Generate uncensored reference-to-video clips with Wan 3.0 Pro R2V. Add up to 10 images plus optional video and audio re…
Create 5–30s cinematic video from one prompt, with optional first and last frame anchors, 1080p–4K output tiers, and an…
Generate 2–30s reference-guided videos from a prompt plus optional images, clips, and audio. This R2V build targets str…
Generate 5–30s MP4 videos from a prompt, with optional first and last frame images. Pick 480p–1080p resolution, aspect…
Animate a still image into a 5–15s HD video from your motion prompt, with optional negative prompts, audio control, and…
Wan 3.0 Prime R2V by Alibaba generates 2–30s MP4 videos from prompts plus reference images, clips, and audio for consis…
Wan 3.0 Prime generates 5–30 second MP4 videos from a written prompt, with optional first and last frames. Pick 480p–10…
Generate 2–30s MP4 videos from a prompt plus reference images, video, and audio. Wan 3.0 R2V helps keep subjects, motio…
Wan 3.0 by Alibaba generates cinematic videos up to 30 seconds at 1080p. Start from text alone or lock the first and la…
Transform a single portrait into a high-stakes cinematic adventure. Face frozen cliffs, desert canyons, volcanic edges,…

Qwen Image Max uncensored generates or edits images from instructions, with optional reference photos for stable identi…
Generate 5–15s 720p/1080p clips from a prompt, with optional first-frame image and audio. WAN 2.6 uncensored is tuned f…
Generate 5–15 second videos from a prompt or first and last frame images. Choose Speed or Quality, and export 480p or 7…
Transform a single car photo into a cinematic vehicle adventure. Turn your car into a powerful robot, a stylized cartoo…
Transform a single portrait into a high-stakes cinematic rescue video. Face frozen wilderness, desert emergencies, coll…
Popular Models
523 models in this collection.

Generate high-resolution images using Seedream v4.5 Uncensored. Supports text-to-image and image-to-image transformatio…
Generates videos from images and reference videos with motion control. Supports custom prompts and character orientatio…
Generate high-quality videos from text prompts using Kling V3. Supports custom frames, duration, and aspect ratios.

An image editing tool designed for quick transformations using reference images and prompts. Supports multi-image mixin…
Seedance Pro v1.5 Uncensored by ByteDance generates short videos from text with optional native audio, strong prompt fo…
Generate 3–10s videos from reference images, or edit an existing clip with one instruction. Built on Google Gemini Omni…
Generate 5s,10s 720p MP4 videos from text, or animate a still image as the opening frame. Google Gemini Omni Flash supp…

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built f…

Google's Gemini 3 Pro Image Preview, also known as Nano Banana, model for text-to-image and image-to-image generation.
Create 4–30s videos from a description, or steer motion using first and last frame images. Choose 480p or 720p output w…

Chat model based on Qwen 3.8 27B with refusal behavior reduced through direction removal. Its 262k context fits long do…

Refusal-reduced variant of Qwen 3.8 27B for long-context chat and coding. It can emit or hide thinking traces and suppo…

Run Unsloth’s GGUF quantizations of Qwen 3.8 Flash Next for long-context chat and reasoning. Includes optional thinking…

Convert a video to MP4, MOV, WebM, MKV, AVI, MPEG, or M4V with adjustable compression. Built by Wiro for clean exports…

Convert an image to JPEG, PNG, WebP, TIFF, or AVIF. Set output quality from 0 to 100 to balance file size and visual fi…
Turn a selfie into a Panini-style player card video. Enter name, height, weight, and birth date, then choose a country…

Generate 1K or 2K images with strong typography and reference-guided edits. Built on ByteDance Seedream 5.0 Pro for pos…

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optiona…

Smart Resize by Wiro converts one image into multiple exact sizes, using AI recomposition to keep key subjects in frame…

A custom Sunburst build of OpenAI GPT Image 2.5 for high-fidelity image generation and precise edits with masks, transp…
Generate Videos
180 models in this collection.
FastH3 V2 is an eight-forward distilled MiniMax H3 model that generates synchronized video and stereo audio from a text…
Create 2–30s reference-guided videos from a prompt plus up to 10 images, 5 clips, and 5 audio tracks. Export 1080p, 2K,…
Generate 5–30 second videos from a written shot prompt, with optional first and last frame keyframes, audio on/off, and…
Generate 2–30s 480p–1080p MP4 videos from a prompt plus up to 10 reference images, with optional reference video and au…
Create 5–30s MP4 videos from a prompt, optionally anchored by first and last frame images with optional audio. Powered…
Generate uncensored reference-to-video clips with Wan 3.0 Pro R2V. Add up to 10 images plus optional video and audio re…
Create 5–30s cinematic video from one prompt, with optional first and last frame anchors, 1080p–4K output tiers, and an…
Generate 2–30s reference-guided videos from a prompt plus optional images, clips, and audio. This R2V build targets str…
Generate 5–30s MP4 videos from a prompt, with optional first and last frame images. Pick 480p–1080p resolution, aspect…
Animate a still image into a 5–15s HD video from your motion prompt, with optional negative prompts, audio control, and…
Wan 3.0 Prime R2V by Alibaba generates 2–30s MP4 videos from prompts plus reference images, clips, and audio for consis…
Wan 3.0 Prime generates 5–30 second MP4 videos from a written prompt, with optional first and last frames. Pick 480p–10…
Generate 2–30s MP4 videos from a prompt plus reference images, video, and audio. Wan 3.0 R2V helps keep subjects, motio…
Wan 3.0 by Alibaba generates cinematic videos up to 30 seconds at 1080p. Start from text alone or lock the first and la…
Transform a single portrait into a high-stakes cinematic adventure. Face frozen cliffs, desert canyons, volcanic edges,…
Generate 5–15s 720p/1080p clips from a prompt, with optional first-frame image and audio. WAN 2.6 uncensored is tuned f…
Generate 5–15 second videos from a prompt or first and last frame images. Choose Speed or Quality, and export 480p or 7…
Transform a single car photo into a cinematic vehicle adventure. Turn your car into a powerful robot, a stylized cartoo…
Transform a single portrait into a high-stakes cinematic rescue video. Face frozen wilderness, desert emergencies, coll…
Transform a single portrait into a nostalgic 1980s-style video. Step into retro arcades, neon diners, summer drives, sc…
Generate Images
72 models in this collection.

Qwen Image Max uncensored generates or edits images from instructions, with optional reference photos for stable identi…

A custom Sunburst build of OpenAI GPT Image 2.5 for high-fidelity image generation and precise edits with masks, transp…

Generate new images or edit reference photos with GPT Image 2.5 Flare Custom by OpenAI. Export PNG/JPEG/WebP up to 3840…

OpenAI’s GPT Image 2.5 Sunburst generates and edits images with tight control using prompts, reference photos, and opti…

Generate and edit images with GPT Image 2.5 Flare from OpenAI. Use reference images, optional masks, size controls, and…

Create new images or edit a source photo with xAI’s Grok Imagine Image V2. Choose 1K or 2K output, quality level, and a…

Generate 1K or 2K images with strong typography and reference-guided edits. Built on ByteDance Seedream 5.0 Pro for pos…

Pruna’s P Image Ideogram Custom generates poster-ready images from a short prompt. It’s tuned for readable text, typogr…

P-Image Ideogram by Pruna generates high-quality text-to-image outputs with strong typography, offering 1K or 2K resolu…

ByteDance Seedream V5 Pro generates images from text and edits a base image using up to 10 references. It’s tuned for i…

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built f…

A foundational text-to-image model designed for efficient, high-resolution image generation with strong prompt followin…

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optiona…

SenseNova U1-8B turns text into high-detail images with strong layout control and clearer in-image text. Use it for pos…

Generate or edit images with GPT Image 2 from OpenAI. It delivers strong instruction following, sharp text rendering, a…

xAI’s Grok Imagine Image generates images from prompts and can edit a source image using instructions. Pick an aspect r…

Generate or edit images using text prompts or image edits with GPT Image 1.5. Supports multiple sizes, formats, and qua…

Generate high-resolution images using Seedream v4.5 Uncensored. Supports text-to-image and image-to-image transformatio…

FireRed-Image-Edit-1.1 significantly enhances identity consistency, multi-image conditioning, and domain-specialized ed…

Generate high-quality images from text prompts or image inputs using the Seedream v5 Lite Uncensored model. Supports mu…
Edit Images
86 models in this collection.

Qwen Image Max uncensored generates or edits images from instructions, with optional reference photos for stable identi…

A custom Sunburst build of OpenAI GPT Image 2.5 for high-fidelity image generation and precise edits with masks, transp…

Generate new images or edit reference photos with GPT Image 2.5 Flare Custom by OpenAI. Export PNG/JPEG/WebP up to 3840…

OpenAI’s GPT Image 2.5 Sunburst generates and edits images with tight control using prompts, reference photos, and opti…

Generate and edit images with GPT Image 2.5 Flare from OpenAI. Use reference images, optional masks, size controls, and…

Create new images or edit a source photo with xAI’s Grok Imagine Image V2. Choose 1K or 2K output, quality level, and a…

Generate 1K or 2K images with strong typography and reference-guided edits. Built on ByteDance Seedream 5.0 Pro for pos…

ByteDance Seedream V5 Pro generates images from text and edits a base image using up to 10 references. It’s tuned for i…

RMBG-2.0 removes image backgrounds with an 8-bit alpha matte for smooth edges. Built by BRIA AI for production cutouts…

Google’s Nano Banana 2 Lite generates and edits 1K images from text, with optional multi-image references. It’s built f…

Remove image backgrounds and get a clean transparent PNG cutout. Ideogram’s generative matting keeps hair, glass, and f…
Turn your photo into a fun birthday celebration while keeping your face natural and unchanged for instantly share-ready…
Turn your photo into a fun birthday celebration with your age featured on the scene, while keeping your face natural an…

Pruna’s p-image-try-on creates virtual try-on images by applying one or more garment photos to a person photo. Generate…

LocateAnything 3B by Nvidia localizes objects, UI elements, and text in images. It returns normalized boxes or points f…

Convert an image to JPEG, PNG, WebP, TIFF, or AVIF. Set output quality from 0 to 100 to balance file size and visual fi…

Smart Resize by Wiro converts one image into multiple exact sizes, using AI recomposition to keep key subjects in frame…

u1-8b-interleave by SenseNova generates step-by-step text with matching images, optionally conditioned on up to 5 refer…

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optiona…

Generate or edit images with GPT Image 2 from OpenAI. It delivers strong instruction following, sharp text rendering, a…
Generate Text
28 models in this collection.

u1-8b-interleave by SenseNova generates step-by-step text with matching images, optionally conditioned on up to 5 refer…

SenseNova U1 8B Visual Understanding creates infographic-style images and prompt-based edits from text and an optional…

Multilingual speech-to-text for 25 European languages with auto language detection, punctuation, capitalization, and op…

CohereLabs cohere-transcribe-03-2026 is a 2B Conformer speech-to-text model for 14 languages. It creates accurate trans…

A lightweight speech-to-text model optimized for fast inference. Converts audio input into text with support for multip…

Voxtral Mini 4B Realtime 2602 is a multilingual, realtime speech-transcription model and among the first open-source so…

Nemotron-Speech-Streaming-En-0.6b is the first unified model in the Nemotron Speech family, engineered to deliver high-…

Speech to text model from ElevenLabs

OpenAI Whisper Medium transcribes speech to text with timestamps and optional speaker diarization. Works across many la…

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detecti…

Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of…

NSFW video detection automatically analyzes video content to identify inappropriate or explicit material, ensuring comp…
VideoLLaMA3-2B is a model designed for video understanding.
VideoLLaMA3-2B-Image is a model designed for image understanding.

VideoLLaMA3-7B-Image is a model designed for image understanding.

VideoLLaMA3-7B is a model designed for video understanding.

BLIP-2 creates captions or detailed descriptions for images. This is BLIP-2 model, leveraging Flan T5-xl.
Generate 3D
4 models in this collection.
Generate Audio
31 models in this collection.




















Generate Music
18 models in this collection.
















Realtime Stream
10 models in this collection.










LLM & Chat
99 models in this collection.




















AI Models for E-commerce
17 models in this collection.








AI Models for Social Media Creators
70 models in this collection.
Everything you need
to ship with confidence.
Wiro is the production layer between your product and every model.
Same auth, same format, every model.
Know your costs before you run.
Inputs, outputs, and cost logged.
Automate completions and build at scale.
Hundreds of models.
One place to find them.
Filter by modality, provider, pricing, context, and more.
Pinpoint the right model.
Evaluate quality, cost, and performance.
Bookmark favorites and build collections.


