by Qwen
The model provider is the Sophon platform. Qwen2.5-VL-72B-Instruct is the latest vision-language model released by the Qwen team. This model excels not only at recognizing common objects such as flowers, birds, fish, and insects, but also at efficiently analyzing text, charts, icons, graphics, and layouts within images. As a visual agent, it is capable of reasoning and dynamically guiding tool usage, supporting both computer and mobile operations. Moreover, it can understand videos longer than one hour and accurately locate relevant video segments.
The model provider is the Sophon platform. Qwen2.5-VL-72B-Instruct is the latest vision-language model released by the Qwen team. This model excels not only at recognizing common objects such as flowers, birds, fish, and insects, but also at efficiently analyzing text, charts, icons, graphics, and layouts within images. As a visual agent, it is capable of reasoning and dynamically guiding tool usage, supporting both computer and mobile operations. Moreover, it can understand videos longer than one hour and accurately locate relevant video segments.
On AIHubMix, Qwen2.5-VL-72B-Instruct costs $0.62 per million input tokens and $0.62 per million output tokens.
Qwen2.5-VL-72B-Instruct accepts text, image and video input.
Qwen2.5-VL-72B-Instruct is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to Qwen2.5-VL-72B-Instruct — no other code changes needed.
Qwen2.5-VL-72B-Instruct is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in…
qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for…
qwen-audio-3.0-tts-plus is a high-performance speech synthesis large model designed for…
HappyHorse-1.1-I2V supports image-to-video generation, further enhancing visual texture…
HappyHorse-1.1-R2V supports reference-based video generation, further improving the…
HappyHorse-1.1-T2V supports text-to-video generation, further enhancing text semantic…
Use Qwen2.5-VL-72B-Instruct via the AIHubMix unified API — one interface for every major LLM.