Qwen 3.6 35B A3B — Multimodal LLM AI Model

Qwen 3.6 35B A3B is a mixture-of-experts vision-language model — 35B parameters total but only 3B active per token, so it answers fast for its size. 256K context, and strong on agentic coding and repo-scale reasoning.

Details

  • CategoryMultimodal LLM
  • Year2026
  • LicenseApache-2.0

Compliance & Provenance

  • ProviderAlibaba (open weights) · GPAI
  • EU AI Act RiskLimited Risk
  • Art. 50 TransparencyRequired — AI-generated outputs are marked

Inputs & Outputs

  • TextInput · string

    Text prompt or question

  • Image 1Input · image · optional

    Optional reference image for visual understanding. When provided, the model analyzes the image content together with the text prompt. Supports up to 4096x4096 resolution.

  • Image 2Input · image · optional

    Optional second reference image. When provided alongside Image 1, the model can compare, contrast, or jointly reason over both images with the text prompt. Supports up to 4096x4096 resolution.

  • TextOutput · string

    Generated text response

Tags

  • multimodal
  • visual-qa
  • image-understanding
  • reasoning
  • agentic-coding
  • moe
  • gpu-heavy

Alternatives in Multimodal Language Models

  • Gemma 4 31B

    Gemma 4 31B Dense flagship VLM, 256K context, thinking mode.

  • Gemma4-E2B

    Gemma 4 E2B compact VLM, 128K context, bf16.

Resources