by Llama
On AIHubMix, deepinfra-llama-3.3-70b-instant-turbo costs $0.11 per million input tokens and $0.35 per million output tokens.
deepinfra-llama-3.3-70b-instant-turbo is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepinfra-llama-3.3-70b-instant-turbo — no other code changes needed.
deepinfra-llama-3.3-70b-instant-turbo is developed by Llama. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Llama 4 Maverick is a high-capacity Mixture-of-Experts (MoE) model from Meta, featuring…
Llama 4 Scout is a highly efficient Mixture-of-Experts (MoE) model from Meta, activating…
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and…
Use deepinfra-llama-3.3-70b-instant-turbo via the AIHubMix unified API — one interface for every major LLM.