by DeepSeek
Provided by Groq, the DeepSeek-R1-Distill model is fine-tuned based on an open-source model, using samples generated by DeepSeek-R1. We have made slight modifications to their configurations and tokenizers. Please use our settings to run these models.
Provided by Groq, the DeepSeek-R1-Distill model is fine-tuned based on an open-source model, using samples generated by DeepSeek-R1. We have made slight modifications to their configurations and tokenizers. Please use our settings to run these models.
On AIHubMix, deepseek-r1-distill-llama-70b costs $0.8 per million input tokens and $1.6 per million output tokens.
deepseek-r1-distill-llama-70b accepts text input.
deepseek-r1-distill-llama-70b supports thinking. Per-protocol parameter support is listed in the capability table on this page.
deepseek-r1-distill-llama-70b is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-r1-distill-llama-70b — no other code changes needed.
deepseek-r1-distill-llama-70b is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek-V4-Flash official API release. Agent capabilities have been greatly enhanced…
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V4 features an ultra-long context of one million characters and achieves leading…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.2 is an efficient large language model equipped with DeepSeek Sparse…
DeepSeek-V3.1 non-thinking mode has now been updated to the DeepSeek-V3.1-Terminus…
Use deepseek-r1-distill-llama-70b via the AIHubMix unified API — one interface for every major LLM.