Qwen3.7 Flash API model profile
Evidence snapshot:
Qwen3.7 Flash is a Alibaba Qwen model profile with a 1,000,000-token context window. The dated catalog snapshot records text, image, video modalities and interfaces for agentic, coding, prompt caching, reasoning, reasoning controls.
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world visual perception.
Qwen3.7 Flash is worth reaching for when you need a route priced in the cheapest quarter of the catalog. The conditions below are the ones its own numbers support.
Qwen3.7 Flash is available under the Kendr alias kc-qwen3.7-flash. Current account availability and customer credit quotes come from Kendr's live public model API and applicable account policy.
Qwen3.7 Flash accepts text, image, and video and returns text. The snapshot records agentic, coding, prompt caching, reasoning, reasoning controls, structured output, text generation, tool calling, tools, video input, vision, and web search as capabilities or interfaces.
The dated reference snapshot lists $0.03 input and $0.13 output per million tokens. Cached input is $0.006 per million tokens. The separately published provider-route reference is From $0.03 / $0.13 ($0.006/M cached; tiered above 32K and 256K input).
No comparable third-party benchmark value is present in the 2026-09-06 snapshot. Missing values are not estimated.
Qwen3.7 Flash ranked #55 in the cited trailing-7-day third-party catalog usage dataset as of 2026-09-06. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share.
Profile data was reviewed for the 2026-09-06 snapshot. Claims retain their source dates, and unavailable fields are shown as unavailable instead of estimated.
Profile facts
- Profile type
- Kendr API model
- Kendr API alias
- kc-qwen3.7-flash
- Model developer
- Alibaba Qwen
- Kendr route provider
- Alibaba Qwen
- Context window
- 1M
- Snapshot date
- 2026-09-06
- Knowledge cutoff
- Not disclosed
- Reference catalog ID
- qwen/qwen3.7-flash
When to pick Qwen3.7 Flash
Qwen3.7 Flash is worth reaching for when you need a route priced in the cheapest quarter of the catalog. The conditions below are the ones its own numbers support.
Each condition below is derived from this model's own values in the 2026-09-06 snapshot, compared only against models publishing the same field.
- Reach for it when cost is the binding constraint. At $0.03 in and $0.13 out per 1M tokens, combined token price sits in the cheapest quarter of the 146 models publishing both rates in this snapshot.
- Reach for it when the input includes images. The snapshot records image input, so screenshots, scans, and diagrams can be sent directly instead of described.
- Reach for it when the model has to call tools. Tool calling is recorded in the snapshot, so this route can drive an agent loop rather than only answer in prose.
Context, modalities, and identity
Qwen3.7 Flash accepts text, image, and video and returns text. The snapshot records agentic, coding, prompt caching, reasoning, reasoning controls, structured output, text generation, tool calling, tools, video input, vision, and web search as capabilities or interfaces.
- Model developer
- Alibaba Qwen
- Kendr route provider
- Alibaba Qwen
- Context window
- 1M
- Maximum output
- 65,536 tokens
- Input modalities
- text, image, and video
- Output modalities
- text
- Knowledge cutoff
- Not disclosed
Model overview
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world visual perception.
- Reference model ID
- qwen/qwen3.7-flash
- Canonical version
- qwen/qwen3.7-flash-20260727
- Hugging Face ID
- Not separately documented
Capabilities and supported controls
Capabilities and parameters are reported from the dated catalog and source records; their presence does not guarantee identical behavior across every provider route.
- Supported parameters
- include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, response_format, seed, temperature, tool_choice, tools, top_logprobs, top_p
- agentic
- coding
- prompt caching
- reasoning
- reasoning controls
- structured output
- text generation
- tool calling
- tools
- video input
- vision
- web search
Dated pricing snapshot
The dated reference snapshot lists $0.03 input and $0.13 output per million tokens. Cached input is $0.006 per million tokens. The separately published provider-route reference is From $0.03 / $0.13 ($0.006/M cached; tiered above 32K and 256K input).
Reference figures are dated 2026-09-06; the live Kendr quote can differ by selected provider route, context tier, caching, tools, region, and current rate card.
- Price date
- 2026-09-06
- Input
- $0.03 per 1M tokens
- Cached input
- $0.006 per 1M tokens
- Output
- $0.13 per 1M tokens
| Prompt threshold | Input | Cached input | Output |
|---|---|---|---|
| From 32,000 prompt tokens | $0.1 / 1M | $0.02 / 1M | $0.4 / 1M |
| From 256,000 prompt tokens | $0.2 / 1M | $0.04 / 1M | $0.8 / 1M |
Provider routes, performance, and uptime
The snapshot retains 1 provider endpoint. Provider prices, context limits, p50 performance, and uptime can differ by route and are not Kendr guarantees.
- Active provider endpoints
- 1
- Best p50 latency
- 0.78 s
- Best p50 throughput
- 46 tok/s
- Availability with routing
- 99.94% over the sampled window
- Availability without routing
- 99.94% over the sampled window
- Performance date
- 2026-09-06
| Provider | Quantization | Input / 1M | Output / 1M | Cache read / 1M | Context | p50 latency | p50 throughput | Uptime (1d) |
|---|---|---|---|---|---|---|---|---|
| Alibaba Cloud Int. | unknown | $0.03 / 1M | $0.13 / 1M | $0.006 / 1M | 1,000,000 tokens | 0.78 s | 46 tok/s | 100.00% |
Benchmark evidence and limitations
No comparable third-party benchmark value is present in the 2026-09-06 snapshot. Missing values are not estimated.
Third-party catalog fields normalized from the cited model dataset; benchmark methodology and coverage differ by source.
Popularity and market context
Qwen3.7 Flash ranked #55 in the cited trailing-7-day third-party catalog usage dataset as of 2026-09-06. Observed within the cited trailing-7-day third-party catalog usage dataset; this is not global AI market share.
- Third-party catalog rank
- #55
- Global market share
- Not inferred
- Observation date
- 2026-09-06
OpenAI-compatible API example
This example applies to the published Kendr alias on this hosted profile. Check the live public catalog before use.
curl https://api.kendr.org/v1/chat/completions \
-H "Authorization: Bearer $KENDR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"kc-qwen3.7-flash","messages":[{"role":"user","content":"Hello"}]}'Frequently asked questions
When should I use Qwen3.7 Flash?
Qwen3.7 Flash is worth reaching for when you need a route priced in the cheapest quarter of the catalog. The conditions below are the ones its own numbers support. Reach for it when cost is the binding constraint: At $0.03 in and $0.13 out per 1M tokens, combined token price sits in the cheapest quarter of the 146 models publishing both rates in this snapshot. Reach for it when the input includes images: The snapshot records image input, so screenshots, scans, and diagrams can be sent directly instead of described. Reach for it when the model has to call tools: Tool calling is recorded in the snapshot, so this route can drive an agent loop rather than only answer in prose.
What is Qwen3.7 Flash?
Qwen3.7 Flash is a Alibaba Qwen model profile with a 1,000,000-token context window. The dated catalog snapshot records text, image, video modalities and interfaces for agentic, coding, prompt caching, reasoning, reasoning controls.
What context window does Qwen3.7 Flash have?
The dated profile lists 1M of context and up to 65,536 output tokens. Provider routes, variants, and runtime configuration can impose lower effective limits.
What benchmark evidence is available for Qwen3.7 Flash?
No comparable third-party benchmark value is present in the 2026-09-06 snapshot. Missing values are not estimated. Third-party catalog fields normalized from the cited model dataset; benchmark methodology and coverage differ by source.
What is the Kendr API alias for Qwen3.7 Flash?
Use kc-qwen3.7-flash as the model value. The live public model API and signed-in catalog remain authoritative for route availability and customer credit quotes.
How is Qwen3.7 Flash priced on Kendr?
The dated reference snapshot lists $0.03 input and $0.13 output per million tokens. Cached input is $0.006 per million tokens. The separately published provider-route reference is From $0.03 / $0.13 ($0.006/M cached; tiered above 32K and 256K input). Kendr applies one 5% markup to configured provider model cost; the live quote and settled routing receipt are authoritative for a request.
Sources and evidence dates
- Qwen3.7 Flash third-party catalog record (catalog, checked 2026-09-06)
- Third-party model catalog methodology (methodology, checked 2026-09-06)
- Alibaba Qwen official model documentation (primary, checked 2026-09-06)
- Qwen3.7 Flash provider or route reference (primary, checked 2026-09-06)