by DeepSeek
DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.
DeepSeek-OCR is a vision-language model launched by DeepSeek AI, focusing on optical character recognition (OCR) and “contextual optical compression.” The model is designed to explore the limits of compressing contextual information from images, efficiently processing documents and converting them into structured text formats such as Markdown. The model requires an image as input.
DeepSeek-OCR has a 8,000 token context window.
On AIHubMix, DeepSeek-OCR costs $0.02 per million input tokens and $0.02 per million output tokens.
DeepSeek-OCR accepts text and image input.
DeepSeek-OCR is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to DeepSeek-OCR — no other code changes needed.
DeepSeek-OCR is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced…
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.1 non-thinking mode has now been updated to the DeepSeek-V3.1-Terminus…
Use DeepSeek-OCR via the AIHubMix unified API — one interface for every major LLM.