deepinfra-llama-4-maverick-17b-128e-instruct

by Llama

API Pricing

Input$0.33 / 1M tokens
Output$1.32 / 1M tokens

FAQ

How much does deepinfra-llama-4-maverick-17b-128e-instruct cost?

On AIHubMix, deepinfra-llama-4-maverick-17b-128e-instruct costs $0.33 per million input tokens and $1.32 per million output tokens.

How do I call deepinfra-llama-4-maverick-17b-128e-instruct via API?

deepinfra-llama-4-maverick-17b-128e-instruct is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepinfra-llama-4-maverick-17b-128e-instruct — no other code changes needed.

Who develops deepinfra-llama-4-maverick-17b-128e-instruct?

deepinfra-llama-4-maverick-17b-128e-instruct is developed by Llama. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from Llama

llama-4-maverick

by Llama

Llama 4 Maverick is a high-capacity Mixture-of-Experts (MoE) model from Meta, featuring…

$0.2/1M in · $0.2/1M out
1,048,576 tokens context

llama-4-scout

by Llama

Llama 4 Scout is a highly efficient Mixture-of-Experts (MoE) model from Meta, activating…

$0.2/1M in · $0.2/1M out
131,000 tokens context

llama-3.3-70b

by Llama

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and…

$0.6/1M in · $0.6/1M out
65,536 tokens context

llama-3.1-70b

by Llama
$0.44/1M in · $0.44/1M out

llama3.1-8b

by Llama

cerebras

$0.3/1M in · $0.6/1M out

cerebras-llama-3.3-70b

by Llama
$0.6/1M in · $0.6/1M out

Use deepinfra-llama-4-maverick-17b-128e-instruct via the AIHubMix unified API — one interface for every major LLM.