Try Seedance 2.5 Uncensored Video from ByteDance →
Models
Agents
Workflows
Studio
PricingBlogDocs
ExploreDiscover models by categoryBrowse All ModelsBrowse the complete catalogSee FavoritesSign in to view saved models
Generative Media AgentCreate and edit media by chattingWorkflow AgentBuild visual workflows with Agent
OverviewThe platform at a glanceLearnSkills, knowledge, guardrailsAnatomyWhat makes agents reasonBuild Your AgentPick skills, set tier, deploy
Pre-built AgentsBrowse the catalog
Agent Usecases
Ad Campaign ManagerApp Event ManagerApp Review RepliesBarber BookingCustomer Win-BackEcommerce ListingsRestaurant Reviews
Sign InStart Building

Task History

Click to see output list

No tasks yet

Go to Models
Explore models/
LLM & ChatActive

Qwen / Qwen3.5-9B

Qwen3.5-9B

byqwen

A dense 9-billion parameter language model optimized for chat and reasoning tasks. Designed for efficient deployment and high-quality responses.

ChatLLMReasoningBf16
Model ID
Qwen3.5-9B
Provider
qwen
Updated
1776064287
Qwen3.5-9B
4
Comments
Average rating : 5 (3 users)
Providerqwen
ModelQwen3.5-9B
ChatLLMReasoningBf16
wiro playground—qwen/Qwen3.5-9B
Reset to defaults

Prompt to send to the model.

Sample outputs

No samples yet

Run this model to create outputs and build up samples.

Updated 1776064287

Qwen3.5-9B Language Model

Overview

Qwen3.5-9B is a dense 9-billion parameter language model optimized for chat and reasoning tasks. It is designed for efficient deployment with support for various inference parameters to customize behavior.

What you can build

This model is ideal for building conversational agents, chatbots, and reasoning-based applications where balanced performance and accuracy are required.

Inputs

  • prompt: Text input to generate a response from the model
  • thinking: Enable or disable reasoning mode
  • user_id (optional): Unique identifier for maintaining chat history
  • session_id (optional): Session identifier for managing multiple chat sessions
  • system_prompt (optional): System-level instructions to guide model behavior
  • temperature, top_p, top_k, repetition_penalty, length_penalty: Parameters to control output randomness and structure
  • max_tokens, min_tokens: Limits on token generation
  • stop_sequences: Stop generation when encountering specified strings
  • seed: For reproducible outputs
  • quantization: Enable quantization for reduced memory usage
  • do_sample: Toggle between deterministic and probabilistic sampling

Outputs

  • Generated text responses based on the input prompt and configuration parameters

Recommended settings

For balanced performance and coherence, start with default values for temperature (0.7), top_p (0.95), and max_tokens (0). Adjust based on application requirements.

Limitations

  • Not optimized for extremely long contexts
  • Output quality depends on training data and parameter tuning
  • Quantization may slightly reduce output quality

Safety & compliance

This model adheres to safety guidelines by default. Ensure compliance with applicable regulations and consider additional safeguards for sensitive applications.

API quick start

Run Qwen3.5-9B with a single API call.

POST https://api.wiro.ai/v1/Run/Qwen/Qwen3.5-9B
{
  "prompt": "Explain the Second Law of Thermodynamics …",
  "user_id": "...",
  "session_id": "...",
  "enableThinking": "true"
}
View full API docs

Discover, test, and run AI models, build workflows and agents with one unified API.

All systems operational
WiroAboutBlogCareersContact
ProductModelsAgentsPricingPartnerChangelogStatusFAQ
Getting StartedIntroductionAuthenticationProjectsCode ExamplesWiro MCP ServerSelf-Hosted MCPn8n IntegrationLLMs.txt
API ReferenceModelsRun a ModelModel ParametersTasksLLM & Chat StreamingWebSocketRealtime VoiceFiles
© 2026 Wiro AI. All rights reserved.
PrivacyTermsData Deletion