Try FastH3 (Text-to-Audio-Video) from Fastvideo →
Models
Agents
Workflows
Studio
PricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
Generative Media AgentCreate and edit media by chattingWorkflow AgentBuild visual workflows with Agent
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
DISCOVER AI MODELS

Search the full model catalog.

Discover and compare production-ready models. Filter by modality, provider, pricing, and production fit.

28 models1 active filters

Active filters

Wiro AI

Video Generation

Image Generation

Audio & Speech

Realtime Stream

Music Generation

3D Generation

LLM & Chat

Workflow

Partners

1
Need help choosing?

Compare models by capability, cost, and provider fit.

Showing 20 of 28 models
gpt-realtime-whisper - AI model cover image

gpt-realtime-whisper

openai

OpenAI’s GPT Realtime Whisper turns live audio into streaming transcript updates. Adjust delay levels to trade latency for higher transcription accuracy.

Run
gpt-realtime-translate - AI model cover image

gpt-realtime-translate

openai

Stream live speech and get translated speech plus live captions with low delay. Made by OpenAI for calls, meetings, broadcasts, and video chat translation.

Speech to Speech
Run
gpt-realtime-2.1-mini - AI model cover image

gpt-realtime-2.1-mini

openai

Build low-latency voice agents with OpenAI’s GPT‑Realtime 2.1 Mini. It takes live audio or text and replies with spoken audio plus a transcript. ([developers.openai.com](https://developers.openai.com/api/docs/models/gpt-realtime-2.1-mini))

Speech to Speech
Run
gpt-realtime-2.1 - AI model cover image

gpt-realtime-2.1

openai

OpenAI's GPT Realtime 2.1 runs speech-to-speech conversations with configurable reasoning and tool calls. It’s built for voice agents that handle noise and interruptions.

Speech to Speech
Run
gpt-live-transcribe - AI model cover image

gpt-live-transcribe

openai

Low-latency speech-to-text for live streams, calls, and meetings. Add context, expected terms, and language hints to improve accuracy.

Run
gpt-5-6-luna - AI model cover image

gpt-5-6-luna

openai

OpenAI’s GPT 5.6 Luna is a fast GPT‑5.6 tier for vision Q&A and chat. Add one or more images plus a prompt to get grounded text answers.

LLM
Run
gpt-5-6-terra - AI model cover image

gpt-5-6-terra

openai

OpenAI GPT 5.6 Terra is a balanced multimodal model for everyday work. Add images and a prompt to get clear, grounded text analysis and writing.

LLM
Run
gpt-5-6-sol - AI model cover image

gpt-5-6-sol

openai

OpenAI’s GPT 5.6 Sol is a flagship reasoning model that can analyze images and produce detailed text answers for coding, research, and visual Q&A.

LLM
Run
gpt-5.4-nano - AI model cover image

gpt-5.4-nano

openai

GPT-5.4 Nano is OpenAI’s smallest GPT-5.4 tier for fast classification, extraction, and image understanding. Built for high-volume workflows.

Partner LLM
Run
gpt-5.4-mini - AI model cover image

gpt-5.4-mini

openai

OpenAI GPT‑5.4 mini is a text model with image understanding and tool support. It’s built for high-volume coding assistants, subagents, and UI screenshot analysis.

Partner LLM
Run
gpt-5.4 - AI model cover image

gpt-5.4

openai

OpenAI GPT-5.4 takes text plus optional images and returns detailed answers, code, and extracted facts. It supports a 1,050,000-token context window.

Partner LLM
Run
gpt-5.4-pro - AI model cover image

gpt-5.4-pro

openai

GPT 5.4 Pro by OpenAI is a long-context reasoning model with image understanding. It produces precise text for complex analysis, coding, and document work.

Partner LLM
Run
gpt-5.5 - AI model cover image

gpt-5.5

openai

OpenAI GPT‑5.5 is a long-context reasoning model that accepts text and images, then generates reliable writing, code, and structured text outputs for professional work.

Partner LLM
Run
gpt-5.5-pro - AI model cover image

gpt-5.5-pro

openai

GPT-5.5 Pro is OpenAI’s highest-accuracy GPT‑5.5 variant for deep reasoning, coding, and analysis. It accepts text with optional images and returns text.

Partner LLM
Run
gpt-image-2-custom - AI model cover image

gpt-image-2-custom

openai

Create or edit images with OpenAI GPT Image 2 using custom pixel sizes, quality tiers, and format controls. Add optional masks to target specific edits.

Text to Image
Run
gpt-image-2 - AI model cover image

gpt-image-2

openai

Generate or edit images with GPT Image 2 from OpenAI. It delivers strong instruction following, sharp text rendering, and flexible sizing up to 4K.

Text to Image
Run
gpt-image-1-5 - AI model cover image

gpt-image-1-5

openai

Generate or edit images using text prompts or image edits with GPT Image 1.5. Supports multiple sizes, formats, and quality settings.

Text to Image
Run
gpt-realtime-mini - AI model cover image

gpt-realtime-mini

openai

GPT Mini Realtime enables low-latency, bidirectional streaming for voice and text. Build interactive, responsive AI experiences that feel natural and immediate.

Speech to Speech
Run
gpt-realtime - AI model cover image

gpt-realtime

openai

GPT Realtime enables low-latency, bidirectional streaming for voice and text. Build interactive, responsive AI experiences that feel natural and immediate.

Speech to Speech
Run
gpt-5-nano - AI model cover image

gpt-5-nano

openai

A compact AI model optimized for efficient processing of complex prompts and multi-modal inputs with support for images and structured data.

Partner LLM
Run
Showing 20 of 28 models

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion