alibaba/qwen3-vl-32b-instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual percepti...
Input⌘ / Ctrl + Enter
Try a prompt
OutputReady
Your response will appear here
Choose an example or write a prompt, then click Run.
API details and access
Model Details
ProviderAlibaba
TypeLlm
Model IDalibaba/qwen3-vl-32b-instruct
Capabilities
InputTextImage
OutputText
Context-
Max Output32,768
VisionSupported
Function CallingSupported
Access details
Chat CompletionsBase URLhttps://api.sandbase.ai
API Endpoint/v1/chat/completions
{
"model": "alibaba/qwen3-vl-32b-instruct",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"max_tokens": 512
}Pricing
Input$0.10 / 1M tokens
Output$0.42 / 1M tokens

