by DeepSeek
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading performance domestically and in the open-source domain in agent capabilities, world knowledge, and reasoning.
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading performance domestically and in the open-source domain in agent capabilities, world knowledge, and reasoning.
deepseek-v4-flash has a 1,000,000 token context window.
On AIHubMix, deepseek-v4-flash costs $0.15 per million input tokens and $0.31 per million output tokens. Cached input reads are billed at $0.0031 per million tokens.
deepseek-v4-flash accepts text input.
deepseek-v4-flash supports tool calling, function calling, structured outputs and thinking. Per-protocol parameter support is listed in the capability table on this page.
deepseek-v4-flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-v4-flash — no other code changes needed.
deepseek-v4-flash is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced…
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.1 non-thinking mode has now been updated to the DeepSeek-V3.1-Terminus…
Thinking mode of DeepSeek-V3.1; DeepSeek V3.1 is a text generation model provided by…
Use deepseek-v4-flash via the AIHubMix unified API — one interface for every major LLM.