by KLing
VIDEO 3.0 Omni: All-in-One Multimodal Input, Voice-Driven Characters, Direct Audio-Visual Output, and Storyboarding Building on the Kling VIDEO O1 and Kling VIDEO 2.6, the Kling 3.0 Model Series leverage a deeply integrated unified model training framework, achieving more native multimodal input and output. It combines Native Audio with Element Consistency Control, and breaks through duration limits.
VIDEO 3.0 Omni: All-in-One Multimodal Input, Voice-Driven Characters, Direct Audio-Visual Output, and Storyboarding Building on the Kling VIDEO O1 and Kling VIDEO 2.6, the Kling 3.0 Model Series leverage a deeply integrated unified model training framework, achieving more native multimodal input and output. It combines Native Audio with Element Consistency Control, and breaks through duration limits.
On AIHubMix, kling-v3-omni costs $2 per million input tokens.
kling-v3-omni accepts text and image input.
kling-v3-omni is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to kling-v3-omni — no other code changes needed.
kling-v3-omni is developed by KLing. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Kling Video O1 is a major unified multimodal video model launched by Kuaishou. It…
Use kling-v3-omni via the AIHubMix unified API — one interface for every major LLM.