by Baidu
The new version of the Wenxin Yiyan large model significantly improves capabilities in image understanding, creation, translation, and coding. It supports a context length of up to 32K tokens for the first time, with a notable reduction in the latency of the first token.
The new version of the Wenxin Yiyan large model significantly improves capabilities in image understanding, creation, translation, and coding. It supports a context length of up to 32K tokens for the first time, with a notable reduction in the latency of the first token.
ernie-4.5-turbo-vl has a 139,000 token context window.
On AIHubMix, ernie-4.5-turbo-vl costs $0.4 per million input tokens and $1.2 per million output tokens.
ernie-4.5-turbo-vl accepts text and image input.
ernie-4.5-turbo-vl supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
ernie-4.5-turbo-vl is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to ernie-4.5-turbo-vl — no other code changes needed.
ernie-4.5-turbo-vl is developed by Baidu. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
ERNIE 5.1 is the latest model in the Wenxin series, with comprehensive upgrades to its…
ERNIE 5.0 is the next-generation natively multimodal foundation model in the ERNIE…
The Ernie-image-Turbo model is an 8-step distilled version of the Ernie-image model, also…
musesteamer-air-image is a text-to-image model developed by the Baidu Search team aimed…
Qianfan-OCR-Fast is a multimodal large model specialized for OCR, trained primarily on…
Qianfan-OCR-Fast is a multimodal large model specialized for OCR, trained primarily on…
Use ernie-4.5-turbo-vl via the AIHubMix unified API — one interface for every major LLM.