alibaba/qwen3-vl-8b-instruct
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interle...
Input⌘ / Ctrl + Enter
Try a prompt
OutputReady
Your response will appear here
Choose an example or write a prompt, then click Run.
API details and access
Model Details
ProviderAlibaba
TypeLlm
Model IDalibaba/qwen3-vl-8b-instruct
Capabilities
InputTextImage
OutputText
Context-
Max Output32,768
VisionSupported
Function CallingSupported
Access details
Chat CompletionsBase URLhttps://api.sandbase.ai
API Endpoint/v1/chat/completions
{
"model": "alibaba/qwen3-vl-8b-instruct",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"max_tokens": 512
}Pricing
Input$0.08 / 1M tokens
Output$0.50 / 1M tokens

