Try Seedance 2.5 Uncensored Video from ByteDance →
Models
Agents
Workflows
Studio
PricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
Generative Media AgentCreate and edit media by chattingWorkflow AgentBuild visual workflows with Agent
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
Explore models/
Image GenerationActive

moondream3-preview / caption

caption

bymoondream3-preview

Moondream3 is a cutting-edge vision-language model that delivers advanced visual reasoning with built-in object detection, pointing, and OCR capabilities—bringing fast, cost-effective, and scalable inference to real-world applications.

Image to TextBf16
Model ID
caption
Provider
moondream3-preview
Updated
1770899032
caption
0
Comments
Average rating : 0 (0 users)
Providermoondream3-preview
Modelcaption
Image to TextBf16
wiro playground—moondream3-preview/caption
Reset to defaults
0 / 1
Maximum 1 image allowed
Drop image to upload

OR

Click to browse your device

Supports: JPG, JPEG, PNG, GIF, WEBP, HEIC

Provide an image file or an image URL.

Sample outputs
moondream3-preview-caption-sample-1.txt
moondream3-preview-caption-sample-2.txt
moondream3-preview-caption-sample-3.txt
Updated 1770899032

The Moondream 3 "Caption" model is a core capability within the Moondream 3 vision-language model (VLM) for generating natural language descriptions of images. It automatically analyzes the visual content and provides a coherent, detailed narrative, with options to produce varying lengths of captions, from short, concise summaries to long, elaborate descriptions that capture style, attributes, pose, and contextual interpretation.

API quick start

Run caption with a single API call.

POST https://api.wiro.ai/v1/Run/moondream3-preview/caption
{
  "inputImage": "https://your-cdn.com/input.png",
  "length": "normal",
  "temperature": 0.7,
  "top_p": 0.95
}
View full API docs

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion