Qwen logo

Qwen: Qwen3.7 Max API

qwen3.7-max
Playground

Qwen3.7 Max API is the flagship text model in the Qwen3.7 pair, designed for coding, office automation, and long-horizon agent execution. It trades image input for a much stronger top logic result than Qwen3.7 Plus. [1]

Its xhigh row ranks #5 with a 66.95 median and 74.56 best score. The standard row ranks #28 at a 18.53 median, showing that production quality depends heavily on the selected reasoning budget. [1][2]

MODALITIES
Price$0.875 / $2.625 / 1MCONTEXT1M

Playground

Test Qwen3.7 Max API with representative production prompts before using qwen3.7-max in a live workflow.

Providers

OpenRouter lists Qwen3.7 Max API with text input, text output, and a 1M context window. APINEED routing and prepaid rates are identified separately. [1]

API Need30% off
Uptime
Total Context
1M
Max Output
65.536K
Input
$0.875/ 1M
Output
$2.625/ 1M
Cache Read
$0.175/ 1M
Cache Write
$1.0938/ 1M

Discount

APINEED lowers official rates of $1.25 input, $3.75 output, $0.25 cache read, and $1.5625 cache create by 30%. Repeated system context can benefit from the lower read rate. The table compares official Qwen3.7 Max API pricing with the APINEED prepaid rate.

Official
Provider baseline
Input $1.25/ 1MOutput $3.75/ 1M
Baseline

Availability

APINEED continuously monitors Qwen3.7 Max API access and keeps requests on healthy capacity.

System statusLast 90 days
APINEED GATEWAY99.39% uptime

Qwen3.7 Max API Benchmarks

The xhigh configuration ranks #5, directly behind Claude Opus 5 xhigh, while the standard configuration ranks #28. The page preserves both benchmark rows. The complete LLM2014 logic 2026-07 table remains below, with Qwen3.7 Max API highlighted when a current row exists. [2]

1stGPT-5.5 (xhigh)83.8077.467.57%494s30,811$25.51$29.57
2ndKimi-K3 (max)82.9174.809.78%1095s39,912$16.76$15.00
3rdClaude Opus 4.8 (xhigh)82.6266.7019.27%791s41,833$28.86$24.64
4thGPT-5.6 Sol (xhigh)81.9472.9910.92%369s17,433$14.43$29.57
5thClaude Opus 5 (xhigh)78.3871.888.29%426s26,392$18.21$24.64
6thQwen3.7-Max (xhigh)74.5666.9510.21%374s54,207$7.81$5.14
7thGLM-5.2 (max)73.6858.2720.91%890s45,683$5.12$4.00
8thGemini 3.1 Pro (high)73.3658.4620.31%214s25,905$8.58$11.83
9thGPT-5.6 Luna (xhigh)69.4951.9725.21%242s38,083$6.31$5.91
10thDeepSeek V4 Pro (max)68.0049.9726.51%1468s60,558$1.45$0.86
11thGrok 4.5 (high)67.5956.7216.08%451s43,372$7.18$5.91
12thMuse Spark 1.166.1857.2313.52%542s42,416$4.98$4.19
13thGemini 3.5 Flash (high)65.9260.398.39%188s36,372$9.03$8.87
14thDoubao-Seed-2.1-pro (high)58.6949.3515.91%2003s85,059$10.21$4.29
15thQwen3.7-Plus (high)58.5044.0924.63%1237s57,152$1.83$1.14
16thGemini 3.6 Flash (high)56.8741.7426.6%206s26,329$5.45$7.39
17thTencent Hy3 (high)54.4244.6018.04%1035s52,037$0.83$0.57
18thClaude Sonnet 5 (xhigh)51.3743.1416.02%529s44,118$12.18$9.86
19thDeepSeek V4 Flash (max)50.6236.2428.41%611s49,593$0.40$0.29
20thMiniMax-M349.1840.9416.75%985s56,467$1.90$1.20
21stClaude Opus 542.0831.7324.6%154s10,180$7.02$24.64
22ndClaude Opus 4.636.8825.7930.07%65s4,387$3.03$24.64
23rdDoubao-Seed-2.0-lite 0428 (high)35.3226.7724.21%656s26,726$0.38$0.51
24thGemini 3.5 Flash (minimal)33.2422.2233.15%44s6,189$1.54$8.87
25thLing-3.0-flash32.5320.6236.61%688s86,204$0.00$0.00
26thGemini 3.5 Flash Lite (high)30.2322.8224.51%188s18,297$1.26$2.46
27thGPT-5.5 Instant28.8717.6638.83%25s1,673$1.39$29.57
28thGemma 4 31B27.9122.0421.03%770s15,050$0.17$0.39
29thMiMo-V2.5-Pro26.9113.0051.69%478s29,443$0.71$0.86
30thQwen3.7-Plus26.1916.2338.03%314s8,692$0.28$1.14
31stStep-3.7-Flash26.1813.7447.52%258s44,912$1.46$1.16
32ndQwen3.5-27B24.9617.9827.96%451s26,391$0.51$0.69
33rdGemini 3.1 Flash Lite (high)23.7814.3039.87%82s28,351$1.17$1.48
34thClaude Sonnet 523.6410.8654.06%109s5,962$1.65$9.86
35thQwen3.7-Max22.4518.5317.46%196s6,710$0.97$5.14
36thERNIE 5.120.4915.4924.4%457s24,093$1.73$2.57
37thopenPangu-2.0-Flash19.8310.7645.74%571s28,199$0.18$0.23
38thLongCat-2.019.409.7449.79%395s14,886$0.48$1.14
39thDeepSeek V4 Flash19.3913.3731.05%63s5,286$0.04$0.29
40thMistral Medium 3.518.4413.8724.78%194s27,854$5.77$7.39
41stDoubao-Seed-2.1-pro17.499.7444.31%305s6,009$0.72$4.29
42ndGLM-5.214.8010.4329.53%23s953$0.11$4.00
43rdLing-2.6-1T13.058.3935.71%229s4,770$0.31$2.29
44thiFLYTEK Spark X212.396.2149.88%458s11,285$0.09$0.29
45thGemini 3.1 Flash Lite11.7910.2812.81%9s773$0.03$1.48
46thMistral Medium 3.510.837.1334.16%17s2,546$0.53$7.39
47thLing-2.6-flash10.074.8551.84%25s4,997$0.04$0.30

Quick Start

Connect Qwen3.7 Max API without changing the OpenAI-style request shape. Choose xhigh selectively for complex repository work or autonomous planning, and keep standard mode for faster routine steps. Tool calling and structured outputs use the same OpenAI-style request.

1

Get your API key

Create an APINEED key for Qwen3.7 Max API and keep it in an environment variable.

export API_NEED_API_KEY=sk-apineed-v1-...
2

Make your first request

Use qwen3.7-max for Qwen3.7 Max API with the APINEED API. The request shape is compatible with OpenAI chat completions, so most SDKs only need a base URL change.

TypeScript SDKPythoncURLOpenAI SDK
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.API_NEED_API_KEY,
  baseURL: "https://apineed.com/v1"
});

const result = await client.chat.completions.create({
  model: "qwen3.7-max",
  messages: [
    { role: "user", content: "Why is the sky blue?" }
  ]
});

console.log(result.choices[0].message.content);
3

Enable streaming and fallbacks

Add stream: true when Qwen3.7 Max API should return server-sent events. APINEED keeps routing, provider health, and fallback handling behind the same endpoint.

curl https://apineed.com/v1/chat/completions \
  -H "Authorization: Bearer $API_NEED_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3.7-max",
    "stream": true,
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

Qwen3.7 Max API Endpoint

Qwen3.7 Max API accepts chat conversations here with streaming or non-streaming text output; the supported controls are listed on this page.

POST/v1/chat/completions
Authorization
Bearer $API_NEED_API_KEY
Content-Type
application/json
HTTP-Referer
optional - your site URL, for rankings
X-Title
optional - your site name, for rankings
Model
qwen3.7-max

Parameters

Qwen3.7 Max API parameters currently listed for qwen3.7-max by the OpenRouter Models API. [1]

NameTypeStatusDescription
frequency_penaltynumberSupportedPenalizes repeated token frequency.
include_reasoningbooleanSupportedIncludes reasoning content in the response when available.
logprobsbooleanSupportedReturns token log probabilities.
max_tokensintegerSupportedLimits generated output tokens.
presence_penaltynumberSupportedPenalizes tokens already present in the output.
reasoningobjectSupportedControls reasoning behavior and token allocation.
response_formatobjectSupportedRequests a specific response format.
seedintegerSupportedRequests deterministic sampling when supported by the provider.
stopstring or arraySupportedStops generation at the supplied sequence.
structured_outputsbooleanSupportedEnables schema-constrained structured output.
temperaturenumberSupportedControls sampling randomness.
tool_choicestring or objectSupportedControls which tool the model may call.
toolsarraySupportedDefines tools available to the model.
top_kintegerSupportedRestricts sampling to the highest-probability tokens.
top_logprobsintegerSupportedSets how many top-token log probabilities are returned.
top_pnumberSupportedControls nucleus sampling.

Qwen3.7 Max API Q&A

Model-specific answers for teams comparing Qwen3.7 Max API on capability, benchmark evidence, integration, and cost.

How highly does Qwen3.7 Max API rank?

The xhigh configuration ranks #5 with a 66.95 median score in the current LLM2014 logic table.

Does Qwen3.7 Max API accept images?

No. The catalog lists text input and text output; choose Qwen3.7 Plus when image input is required.

How much does Qwen3.7 Max API cache creation cost?

See the live pricing table above for the current API Need input and output prices for qwen3.7-max.

Why are two Qwen3.7 Max rows shown?

For Qwen3.7 Max API, lLM2014 tested xhigh and standard configurations separately, so APINEED does not combine their different reasoning budgets.

Which endpoint serves Qwen3.7 Max API?

Call Qwen3.7 Max API through https://apineed.com/v1/chat/completions with qwen3.7-max as the model value. Existing OpenAI SDK clients usually need only the APINEED base URL and key.

How do I validate Qwen3.7 Max API before release?

Evaluate Qwen3.7 Max API on representative prompts, record quality, latency, and token use, then choose reasoning settings and fallbacks from those results.

How to Deploy the Qwen3.7 Max API on apineed.com

Deploy Qwen3.7 Max API through APINEED after validating its model-specific trade-offs above. The API key, credit, and endpoint flow stays consistent across the catalog.

1

Add credits for Qwen3.7 Max API

Add credits on APINEED before deploying the Qwen3.7 Max API. Pay-as-you-go billing lets usage start small and scale with production traffic.

2

Get your API key

Create an APINEED API key for the Qwen3.7 Max API. The same key can call Qwen3.7 Max and other AI APIs through apineed.com.

3

Set the model slug

Use qwen3.7-max as the model value when you deploy the Qwen3.7 Max API. Keep the APINEED base URL at https://apineed.com/v1.

4

Send a request to Qwen3.7 Max API

Send chat completions or responses to the Qwen3.7 Max API from Claude Code, Codex, or any custom agent. The request keeps its OpenAI-compatible shape.

More models from Qwen