Qwen model page

Qwen: Qwen3.7 Flash

qwen/qwen3.7-flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

TextVisionVideoStructured outputTool callingReasoning
Input / 1M
$0.04
Output / 1M
$0.17
Context
1000K
Markup
30%

Use this model through UAI

A dedicated public landing page for this synced catalog row.

Provider
Qwen

Synced from the active catalog

Context window
1M tokens

Validated for gateway requests

UAI input
$0.04

Charged from your UAI balance

UAI output
$0.17

Same API key and billing flow

This page exists for direct linking, onboarding, and SEO. The same model row powers /v1/models, the public catalog, the app pricing screen, and request validation for /v1/chat/completions.

Why use this model through UAI

The same model route, but with one account, one balance, and one API shape.

Reasoning workloads

Use Qwen: Qwen3.7 Flash for multi-step prompts, planning, and deeper analysis while keeping the same OpenAI-compatible UAI endpoint.

Multimodal prompts

The synced capabilities for this row indicate image or vision support, so you can route multimodal requests through the same UAI account and billing flow.

Interactive latency

Qwen: Qwen3.7 Flash is positioned well for chat UIs, demos, and agent loops where faster turn time matters.

Pricing and access

Current synced numbers for this exact model row.

UAI input / 1M tokens
$0.04
UAI output / 1M tokens
$0.17
Qwen input / 1M
$0.03
Qwen output / 1M
$0.13
Context window
1,000,000
Current markup
30%

Need another option? Browse the full models catalog or compare rows in the signed-in pricing table.

Quickstart

Use this model with your UAI key over the standard OpenAI-compatible route.

Create a key in API Keys, then send requests to https://uai.sh/v1/chat/completions.

curl
curl https://uai.sh/v1/chat/completions \
  -H "Authorization: Bearer uai-YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3.7-flash",
    "messages": [{"role": "user", "content": "Say hi in one short sentence."}],
    "stream": false
  }'
Python (OpenAI SDK)
from openai import OpenAI

client = OpenAI(base_url="https://uai.sh/v1", api_key="uai-YOUR_API_KEY")
response = client.chat.completions.create(
    model="qwen/qwen3.7-flash",
    messages=[{"role": "user", "content": "Say hi in one short sentence."}],
    stream=False
)
print(response.choices[0].message.content)
Node.js (OpenAI SDK)
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://uai.sh/v1",
  apiKey: "uai-YOUR_API_KEY"
});

const response = await client.chat.completions.create({
  model: "qwen/qwen3.7-flash",
  messages: [{ role: "user", content: "Say hi in one short sentence." }],
  stream: false
});

console.log(response.choices[0].message.content);

Related models

Other active rows to compare without leaving the public catalog.

qwen
Qwen 3 30B
qwen3:30b
Input / 1M$0.00
ContextContext pending
qwen
Qwen3.6 35B A3B (Lenovo)
shared-local/up_lenovo_rtx8000_ollama/qwen3.6:35b-a3b
Input / 1M$0.00
Context33K tokens
qwen
Qwen3.6 35B A3B (Mac Max)
shared-local/up_125e9804399140ab9fde90afda8cd12c/qwen/qwen3.6-35b-a3b
Input / 1M$0.00
Context33K tokens