kling-v3-omni

by KLing

VIDEO 3.0 Omni: All-in-One Multimodal Input, Voice-Driven Characters, Direct Audio-Visual Output, and Storyboarding Building on the Kling VIDEO O1 and Kling VIDEO 2.6, the Kling 3.0 Model Series leverage a deeply integrated unified model training framework, achieving more native multimodal input and output. It combines Native Audio with Element Consistency Control, and breaks through duration limits.

API Pricing

Input$2 / 1M tokens

Specifications

Modalitiestext, image

FAQ

What is kling-v3-omni?

VIDEO 3.0 Omni: All-in-One Multimodal Input, Voice-Driven Characters, Direct Audio-Visual Output, and Storyboarding Building on the Kling VIDEO O1 and Kling VIDEO 2.6, the Kling 3.0 Model Series leverage a deeply integrated unified model training framework, achieving more native multimodal input and output. It combines Native Audio with Element Consistency Control, and breaks through duration limits.

How much does kling-v3-omni cost?

On AIHubMix, kling-v3-omni costs $2 per million input tokens.

What modalities does kling-v3-omni support?

kling-v3-omni accepts text and image input.

How do I call kling-v3-omni via API?

kling-v3-omni is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to kling-v3-omni — no other code changes needed.

Who develops kling-v3-omni?

kling-v3-omni is developed by KLing. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from KLing

kling-video-o1

by KLing

Kling Video O1 is a major unified multimodal video model launched by Kuaishou. It…

$2/1M in

Use kling-v3-omni via the AIHubMix unified API — one interface for every major LLM.