MODEL PROVIDER

Alibaba

Explore 181 models and APIs from Alibaba, available through one SandBase integration.

Browse all models
181models available
01

Large Language Models (LLMs)

21 on this page

LLM

alibaba/qwen3-32b

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode ...

$0.08 / $0.28 per 1M tokens131.1K context
LLM

alibaba/qwen3.7-max

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivi...

$1.25 / $3.75 per 1M tokens1M context
LLM

alibaba/qwen3.7-plus

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

$0.40 / $1.60 per 1M tokens1M context
LLM

alibaba/qwen3.6-plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to t...

$0.33 / $1.95 per 1M tokens1M context
LLM

alibaba/qwen3.8-2.4t-a95b

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen, with 95 billion active parameters out of 2.4 trillion total. It is the open-weight variant of Qwen3.8 Max.

$2.00 / $6.00 per 1M tokens262.1K context
LLM

alibaba/qwen-plus

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

$0.26 / $0.78 per 1M tokens1M context
LLM

alibaba/qwen3-vl-30b-a3b-instruct

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general mu...

$0.13 / $0.52 per 1M tokens262.1K context
LLM

alibaba/qwen3-235b-a22b

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex r...

$0.46 / $1.82 per 1M tokens131.1K context
LLM

alibaba/qwen-2.5-72b-instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

$0.36 / $0.40 per 1M tokens131.1K context
LLM

alibaba/qwen3.5-flash-02-23

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference effic...

$0.07 / $0.26 per 1M tokens1M context
LLM

alibaba/qwen3-coder

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-con...

$0.22 / $1.80 per 1M tokens1M context
LLM

alibaba/qwen3-vl-8b-instruct

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal...

$0.08 / $0.50 per 1M tokens256K context
LLM

alibaba/qwen3.5-9b

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified...

$0.10 / $0.15 per 1M tokens262.1K context
LLM

alibaba/qwen3-vl-8b-thinking

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences...

$0.12 / $1.36 per 1M tokens256K context
LLM

alibaba/qwen-vl-max

Qwen VL Max is a visual understanding model with 7500 tokens context length. It excels in delivering optimal performance for a broader spectrum of complex tasks.

$0.52 / $2.08 per 1M tokens131.1K context
LLM

alibaba/qwen3.6-max-preview

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic cod...

$1.04 / $6.24 per 1M tokens262.1K context
LLM

alibaba/qwen3.5-27b

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities a...

$0.20 / $1.56 per 1M tokens262.1K context
LLM

alibaba/qwen-turbo

Qwen-Turbo, based on Qwen2.5, is a 1M context model that provides fast speed and low cost, suitable for simple tasks.

$0.03 / $0.13 per 1M tokens131.1K context
LLM

alibaba/qwen3-235b-a22b-thinking-2507

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass...

$0.10 / $0.10 per 1M tokens262.1K context
LLM

alibaba/qwen3-coder-plus

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...

$0.65 / $3.25 per 1M tokens1M context
LLM

alibaba/qwen2.5-vl-72b-instruct

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

$0.25 / $0.75 per 1M tokens131.1K context
02

Image Models

13 on this page

IMAGE

alibaba/qwen-image-edit/lora

Qwen Image Edit Lora by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images ef...

$0.0350/call
IMAGE

alibaba/qwen-image-edit-2509-lora-gallery/remove-element

Qwen Image Edit 2509 Lora Gallery Remove Element by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change s...

$0.0350/call
IMAGE

alibaba/qwen-image-edit-plus/lora-gallery/lighting-restoration

Qwen Image Edit Plus Lora Gallery Lighting Restoration by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, ch...

$0.0350/call
IMAGE

alibaba/wan-vace

Wan Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.2000/call
IMAGE

alibaba/qwen-image-edit-plus/lora-gallery/next-scene

Qwen Image Edit Plus Lora Gallery Next Scene by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change style...

$0.0350/call
IMAGE

alibaba/qwen-image-edit/2509-lora-gallery/face-to-full-portrait

Qwen Image Edit 2509 Lora Gallery Face To Full Portrait is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement...

$0.0350/call
IMAGE

alibaba/z-image/turbo/inpaint/lora

Z Image Turbo Inpaint by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images e...

$0.0115/call
IMAGE

alibaba/qwen-image-edit/2509-lora-gallery/add-background

Qwen Image Edit 2509 Lora Gallery Add Background by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change s...

$0.0350/call
IMAGE

alibaba/qwen-image-edit/2509-lora-gallery/integrate-product

Qwen Image Edit 2509 Lora Gallery Integrate Product is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to ...

$0.0350/call
IMAGE

alibaba/face-swap

Face Swap replaces the face in a target image with the face from a source image, producing a realistic face-swapped result.

$0.0130/call
IMAGE

alibaba/head-swap

Head Swap replaces the head in a target image with the head from a source image, producing a realistic head-swapped result.

$0.0130/call
IMAGE

alibaba/qwen-image-3/edit

Alibaba Qwen Image 3 image editing model with support for up to three reference images.

$0.0750/call
IMAGE

alibaba/qwen-image-3

Qwen Image 3 by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

$0.0750/call
03

Video Models

15 on this page

VIDEO

alibaba/wan/2.1/image-to-video/lora

Wan 2.1 Lora is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

$0.7500/call
VIDEO

alibaba/wan/vision-enhancer

Wan Vision Enhancer is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.3000/call
VIDEO

alibaba/wan/2.1/vace/depth

Wan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.6400/call
VIDEO

alibaba/wan/2.2/fun-control

Wan 2.2 Fun Control is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.1000/call
VIDEO

alibaba/wan/alpha

Wan Alpha is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

$0.0400/call
VIDEO

alibaba/wan/2.2/vace-fun/depth

Wan 2.2 Vace Fun by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.1000/call
VIDEO

alibaba/wan/move

Wan Move by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

$0.2000/call
VIDEO

alibaba/wan/2.1/text-to-video/lora

Wan 2.1 Lora by Alibaba - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.

$0.7500/call
VIDEO

alibaba/wan/2.1/vace/pose

Wan 2.1 Vace is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.6400/call
VIDEO

alibaba/wan/2.1/vace

Wan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.6400/call
VIDEO

alibaba/wan/2.2/vace-fun/inpainting

Wan 2.2 Vace Fun is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.1000/call
VIDEO

alibaba/wan/ati

Wan Ati is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

$0.1500/call
VIDEO

alibaba/wan/v2.2-5b/text-to-video/distill

Wan V2.2 5b Distill is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

$0.0800/call
VIDEO

alibaba/wan-animate

Wan-Animate generates an animated video from an input image and a driving video, transferring motion onto the image subject.

$0.0530/call
VIDEO

alibaba/flashvsr

FlashVSR is a video super-resolution model that upscales videos to higher resolutions (720p / 1080p / 2K / 4K) with fast inference.

$0.0160/call
04

embedding Models

1 on this page

EXPLORE

Other model providers

View all providers ↗