by Inclusionai
Developed by Inclusionai, ling-3.0-flash-free is a 124B-parameter Mixture-of-Experts (MoE) model with approximately 5.1B parameters activated per token. Featuring an expansive context length of 262,144 tokens, this model is built to handle extensive datasets and long-form content. It is designed with token efficiency and production-scale agentic inference as key priorities to enable seamless developer deployment.
Developed by Inclusionai, ling-3.0-flash-free is a 124B-parameter Mixture-of-Experts (MoE) model with approximately 5.1B parameters activated per token. Featuring an expansive context length of 262,144 tokens, this model is built to handle extensive datasets and long-form content. It is designed with token efficiency and production-scale agentic inference as key priorities to enable seamless developer deployment.
ling-3.0-flash-free has a 262,144 token context window.
ling-3.0-flash-free accepts text input.
ling-3.0-flash-free supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.
ling-3.0-flash-free is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to ling-3.0-flash-free — no other code changes needed.
ling-3.0-flash-free is developed by Inclusionai. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Use ling-3.0-flash-free via the AIHubMix unified API — one interface for every major LLM.