ling-3.0-flash-free

by Inclusionai

Developed by Inclusionai, ling-3.0-flash-free is a 124B-parameter Mixture-of-Experts (MoE) model with approximately 5.1B parameters activated per token. Featuring an expansive context length of 262,144 tokens, this model is built to handle extensive datasets and long-form content. It is designed with token efficiency and production-scale agentic inference as key priorities to enable seamless developer deployment.

Specifications

Context window262,144 tokens
Modalitiestext
Featuresreasoning, tool calling, long context
Endpointschat_completions

FAQ

What is ling-3.0-flash-free?

Developed by Inclusionai, ling-3.0-flash-free is a 124B-parameter Mixture-of-Experts (MoE) model with approximately 5.1B parameters activated per token. Featuring an expansive context length of 262,144 tokens, this model is built to handle extensive datasets and long-form content. It is designed with token efficiency and production-scale agentic inference as key priorities to enable seamless developer deployment.

What is the context length of ling-3.0-flash-free?

ling-3.0-flash-free has a 262,144 token context window.

What modalities does ling-3.0-flash-free support?

ling-3.0-flash-free accepts text input.

What features does ling-3.0-flash-free support?

ling-3.0-flash-free supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call ling-3.0-flash-free via API?

ling-3.0-flash-free is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to ling-3.0-flash-free — no other code changes needed.

Who develops ling-3.0-flash-free?

ling-3.0-flash-free is developed by Inclusionai. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

Use ling-3.0-flash-free via the AIHubMix unified API — one interface for every major LLM.