by Moonshot AI
The kimi-k2-turbo-preview model is a high-speed version of kimi-k2, with the same model parameters as kimi-k2, but the output speed has been increased from 10 tokens per second to 40 tokens per second.
The kimi-k2-turbo-preview model is a high-speed version of kimi-k2, with the same model parameters as kimi-k2, but the output speed has been increased from 10 tokens per second to 40 tokens per second.
kimi-k2-turbo-preview has a 262,144 token context window.
On AIHubMix, kimi-k2-turbo-preview costs $1.2 per million input tokens and $4.8 per million output tokens. Cached input reads are billed at $0.3 per million tokens.
kimi-k2-turbo-preview accepts text input.
kimi-k2-turbo-preview supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
kimi-k2-turbo-preview is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to kimi-k2-turbo-preview — no other code changes needed.
kimi-k2-turbo-preview is developed by Moonshot AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Kimi K3 is Kimi’s flagship model for long-horizon coding and end-to-end knowledge work…
coding-kimi-k3-free is the open and free version of coding-kimi-k3. To maintain reliable…
Kimi K2.7 Code is Kimi’s most intelligent Coding model, capable of completing programming…
High-Speed version of Kimi K2.7 Code model, with output speed of approximately 180…
Kimi K2.6 is Kimi's latest and most intelligent model, with stronger and more stable…
Use kimi-k2-turbo-preview via the AIHubMix unified API — one interface for every major LLM.