Qwen3.7-Max

Qwen3.7 系列规模最大的旗舰模型, 面向智能体场景, 在编程 / 长周期任务 / 推理上有强势表现. 1M context.

Model ID: qwen3.7-max · Type: chat · Provider: Alibaba

Endpoints: /v1/chat/completions · /v1/messages

Pricing

Input (per 1M tokens)$1.75 USD
Output (per 1M tokens)$5.25 USD
Cache read (per 1M tokens)$0.175 USD
Cache write 5m (per 1M tokens)$2.1875 USD

Capabilities

from openai import OpenAI

client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
    model="qwen3.7-max",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)

FAQ

How much does Qwen3.7-Max cost on J8API?

Qwen3.7-Max (`qwen3.7-max`) is billed per usage at $1.75/1M in · $5.25/1M out, in USD. Current pricing is always listed at https://j8api.com/models/qwen3.7-max.

How do I call Qwen3.7-Max through J8API?

Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "qwen3.7-max"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.

Which endpoints does Qwen3.7-Max support?

Qwen3.7-Max can be called on: /v1/chat/completions; /v1/messages.

What is Qwen3.7-Max's context window?

Qwen3.7-Max accepts up to 991,808 input tokens and can return up to 65,536 output tokens. Requests exceeding the input limit are rejected before reaching the model.

What can Qwen3.7-Max do?

Qwen3.7-Max supports: function_calling, prompt_caching, reasoning, thinking, streaming.

Who makes Qwen3.7-Max?

Qwen3.7-Max is a chat model from Alibaba, available through the J8API gateway with the same API key as every other model.

Call it through the J8API OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.

API reference · All models