Alibaba releases Qwen3.5-35B-A3B, a 35B multimodal model with Apache 2.0 license
Alibaba's Qwen team has released Qwen3.5-35B-A3B-Base, a 35-billion parameter multimodal model supporting image-text-to-text tasks. The model is available under the Apache 2.0 license and compatible with major inference endpoints including Azure deployment.
Qwen3.5-35B-A3B-Base — Quick Specs
Alibaba's Qwen division has released Qwen3.5-35B-A3B-Base, a 35-billion parameter multimodal language model designed for image-text-to-text tasks.
Model Details
The model was published on February 24, 2026 on Hugging Face and carries an Apache 2.0 license, allowing both commercial and research use without licensing restrictions. It is tagged as part of the Qwen3.5 MoE (mixture of experts) family, indicating the model uses conditional computation techniques to improve efficiency.
Qwen3.5-35B-A3B-Base supports multimodal inputs, processing both images and text to generate text outputs. The model is compatible with the Transformers library and uses SafeTensors format for weight storage, a security-focused serialization standard.
Availability and Deployment
The model has achieved 1,937 downloads and 62 likes on Hugging Face as of publication. It is compatible with inference endpoints through major cloud providers, including Azure deployment options, making it accessible for production use cases.
The base model variant indicates this is the foundational version without instruction-tuning or fine-tuning for specific tasks, leaving optimization to end users or downstream applications.
Context
This release continues Alibaba's Qwen series momentum in the open-weight model space. The Qwen3.5 line represents an iteration beyond Qwen3, with the A3B variant designation referring to a specific model configuration within the 35B parameter class.
The mixture-of-experts architecture employed in this model typically provides efficiency improvements during inference compared to dense models of equivalent parameter count, though exact computational requirements are not yet published.
What This Means
Alibaba is positioning Qwen3.5-35B-A3B as an open alternative for organizations needing multimodal capabilities at the 35B scale. The Apache 2.0 license removes commercial deployment barriers, and cloud provider integration lowers infrastructure barriers. The model joins a competitive field of open multimodal 30B+ parameter models from Meta, Mistral, and others, each with different architectural choices and trade-offs in performance, efficiency, and licensing.
Related Articles
Alibaba releases Qwen 3.8, a 2.4 trillion parameter open-weight model claiming second place behind Fable 5
Alibaba has released Qwen 3.8, a 2.4 trillion parameter open-weight model that the company claims trails only Fable 5. The multimodal model processes images, videos, and documents, with a preview available through Alibaba's platforms at 10 percent of standard pricing.
Alibaba previews Qwen3.8 with 2.4 trillion parameters, claims second place without benchmark data
Alibaba unveiled Qwen3.8 at the World Artificial Intelligence Conference in Shanghai, claiming the 2.4 trillion parameter model ranks second only to Anthropic's Fable 5. The company provided no benchmark scores, model card, or independent verification to support the claim.
Moonshot AI and Alibaba release 2.8T and 2.4T parameter models, claim performance near GPT-5.6 and Claude Fable 5
Within days, Moonshot AI and Alibaba unveiled what they claim are frontier-class models. Moonshot's Kimi K3, at 2.8 trillion parameters, and Alibaba's Qwen3.8, at 2.4 trillion parameters, will both be released as open-weight models with full weights available for download.
Moonshot AI Releases Kimi K3: 2.8T Parameter Open Model at $3/$15 Per Million Tokens
Moonshot AI has released Kimi K3, a 2.8 trillion parameter model with 1 million token context window and native multimodal input. The model ranks #1 in Frontend Code Arena and #9 in Text Arena, with pricing at $3 per million input tokens and $15 per million output tokens—comparable to Claude Sonnet 5 pricing while delivering performance the company claims is near Claude Opus 4.8 and GPT-5.5.
Comments
Loading...