by Llama
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model is optimized for multilingual dialogue use cases and outperforms many of the available open source and closed chat models on common industry benchmarks.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model is optimized for multilingual dialogue use cases and outperforms many of the available open source and closed chat models on common industry benchmarks.
llama-3.3-70b has a 65,536 token context window.
On AIHubMix, llama-3.3-70b costs $0.6 per million input tokens and $0.6 per million output tokens.
llama-3.3-70b is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to llama-3.3-70b — no other code changes needed.
llama-3.3-70b is developed by Llama. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Llama 4 Maverick is a high-capacity Mixture-of-Experts (MoE) model from Meta, featuring…
Llama 4 Scout is a highly efficient Mixture-of-Experts (MoE) model from Meta, activating…
Use llama-3.3-70b via the AIHubMix unified API — one interface for every major LLM.