by Qwen
The Qwen3 series visual understanding model achieves an effective fusion of thinking and non-thinking modes. Its visual agent capabilities reach world-class levels on public test sets such as OS World. This version features comprehensive upgrades in visual coding, spatial perception, and multimodal reasoning; visual perception and recognition abilities are greatly enhanced, supporting ultra-long video understanding.
The Qwen3 series visual understanding model achieves an effective fusion of thinking and non-thinking modes. Its visual agent capabilities reach world-class levels on public test sets such as OS World. This version features comprehensive upgrades in visual coding, spatial perception, and multimodal reasoning; visual perception and recognition abilities are greatly enhanced, supporting ultra-long video understanding.
qwen3-vl-plus has a 256,000 token context window.
On AIHubMix, qwen3-vl-plus costs $0.14 per million input tokens and $1.37 per million output tokens. Cached input reads are billed at $0.03 per million tokens.
qwen3-vl-plus accepts text, image and video input.
qwen3-vl-plus supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
qwen3-vl-plus is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3-vl-plus — no other code changes needed.
qwen3-vl-plus is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in…
qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for…
qwen-audio-3.0-tts-plus is a high-performance speech synthesis large model designed for…
HappyHorse-1.1-I2V supports image-to-video generation, further enhancing visual texture…
HappyHorse-1.1-R2V supports reference-based video generation, further improving the…
HappyHorse-1.1-T2V supports text-to-video generation, further enhancing text semantic…
Use qwen3-vl-plus via the AIHubMix unified API — one interface for every major LLM.