Text model · Alibaba
Qwen 3 VL 235B
Alibaba's Qwen 3 VL vision-language model
What is Qwen 3 VL 235B?
Qwen 3 VL 235B is Alibaba's large vision-language model, reasoning across text, images, and video.
Capabilities
- Reason over text
- Understand images and video alongside text
Qwen 3 VL 235B on Melius
Melius runs Qwen 3 VL 235B alongside every other major image, video, audio, and text model, so you can generate with it, compare it against the field, and let the Mel agent route to it automatically — all on one infinite canvas. You only pay for what you generate.
Frequently asked questions
What is Qwen 3 VL 235B?
Qwen 3 VL 235B is Alibaba's large vision-language model, reasoning across text, images, and video.
How do I use Qwen 3 VL 235B on Melius?
Melius is a node-based canvas: add a node, pick Qwen 3 VL 235B from the model picker, and connect it to whatever should feed it or follow it. You can also just describe what you want and let the Mel agent wire it up. You only pay for what you generate.