inclusionai/ling-3.0-flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
A dedicated public landing page for this synced catalog row.
Synced from the active catalog
Validated for gateway requests
Charged from your UAI balance
Same API key and billing flow
This page exists for direct linking, onboarding, and SEO. The same model row powers /v1/models, the public catalog, the app pricing screen, and request validation for /v1/chat/completions.
The same model route, but with one account, one balance, and one API shape.
inclusionAI: Ling 3.0 Flash exposes 262K tokens, which makes it a good fit for large documents, transcripts, and retrieval-heavy chat flows.
inclusionAI: Ling 3.0 Flash is positioned well for chat UIs, demos, and agent loops where faster turn time matters.
Tool and function-style parameters stay available through the same UAI request surface, so agent stacks do not need a separate provider integration.
Current synced numbers for this exact model row.
Need another option? Browse the full models catalog or compare rows in the signed-in pricing table.
Use this model with your UAI key over the standard OpenAI-compatible route.
Create a key in API Keys, then send requests to https://uai.sh/v1/chat/completions.
curl https://uai.sh/v1/chat/completions \
-H "Authorization: Bearer uai-YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "inclusionai/ling-3.0-flash",
"messages": [{"role": "user", "content": "Say hi in one short sentence."}],
"stream": false
}'
from openai import OpenAI
client = OpenAI(base_url="https://uai.sh/v1", api_key="uai-YOUR_API_KEY")
response = client.chat.completions.create(
model="inclusionai/ling-3.0-flash",
messages=[{"role": "user", "content": "Say hi in one short sentence."}],
stream=False
)
print(response.choices[0].message.content)
import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://uai.sh/v1",
apiKey: "uai-YOUR_API_KEY"
});
const response = await client.chat.completions.create({
model: "inclusionai/ling-3.0-flash",
messages: [{ role: "user", content: "Say hi in one short sentence." }],
stream: false
});
console.log(response.choices[0].message.content);
Other active rows to compare without leaving the public catalog.