gemma-4-26b-a4b-it-free

by Google

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model developed by Google DeepMind. Featuring an expansive context length of 262,144 tokens, it delivers near-31B quality with highly efficient inference. Despite its 25.2B total parameters, only 3.8B are activated per token, making it an incredibly fast and cost-effective solution.

Specifications

Context window262,144 tokens
Modalitiestext, image
Featuresreasoning, tool calling, long context
Endpointschat_completions

FAQ

What is gemma-4-26b-a4b-it-free?

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model developed by Google DeepMind. Featuring an expansive context length of 262,144 tokens, it delivers near-31B quality with highly efficient inference. Despite its 25.2B total parameters, only 3.8B are activated per token, making it an incredibly fast and cost-effective solution.

What is the context length of gemma-4-26b-a4b-it-free?

gemma-4-26b-a4b-it-free has a 262,144 token context window.

What modalities does gemma-4-26b-a4b-it-free support?

gemma-4-26b-a4b-it-free accepts text and image input.

What features does gemma-4-26b-a4b-it-free support?

gemma-4-26b-a4b-it-free supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call gemma-4-26b-a4b-it-free via API?

gemma-4-26b-a4b-it-free is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemma-4-26b-a4b-it-free — no other code changes needed.

Who develops gemma-4-26b-a4b-it-free?

gemma-4-26b-a4b-it-free is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

Paid version: gemma-4-26b-a4b-it

More from Google

gemini-3.6-flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

gemini-3.1-flash-lite-image

by Google

Google's newest, most compact, and most cost-effective image generation and editing…

$0.25/1M in · $1.5/1M out

gemini-3.5-flash-lite

by Google

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for…

$0.3/1M in · $2.5/1M out
1,048,576 tokens context

gemini-3.5-flash-lite-free

by Google

Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…

1,048,576 tokens context

gemini-3.6-flash-free

by Google

Gemini 3.6 Flash free version: fFree model resources are limited and provided only for…

1,000,000 tokens context

gemini-3.5-flash

by Google

Gemini 3.5 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $9/1M out
1,000,000 tokens context

Use gemma-4-26b-a4b-it-free via the AIHubMix unified API — one interface for every major LLM.