MoE Architecture โ€ข 30B Total / 3B Active

๐Ÿ‘๏ธ Qwen3-VL-30B-A3B-Instruct Demo

Explore Alibaba's frontier Vision-Language model featuring deep visual perception, extended 256K context, visual coding, OCR across 32 languages, and spatial grounding.

Compute Engine

Choose between local ZeroGPU execution or cloud Inference Providers (DeepInfra, Novita).

0 1.5
0.1 1
128 4096
Ask a question or attach images...
Built with โค๏ธ using Qwen3-VL-30B-A3B-Instruct, Gradio, and Hugging Face ZeroGPU.