Qwen
qwen3.7-flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial unde
Aggregated modalities
Channel
Each channel is compared at its lowest input-price tier; output breaks ties. A dash means /api/pricing does not provide that rate. Selecting a channel updates the endpoints and call example below.
ChannelUSD per 1M tokensInputOutputCache readCache write
qwen/qwen3.7-flashInput
$0.03 USD per 1M tokens
Output
$0.13 USD per 1M tokens
Cache read
$0.006 USD per 1M tokens
Cache write
$0.0375 USD per 1M tokens
Endpoints
POST
/v1/chat/completionsChat CompletionsCall it
Using qwen/qwen3.7-flash
1from openai import OpenAI23client = OpenAI(4 base_url="https://www.realrelay.ai/v1",5 api_key="sk-***",6)78response = client.chat.completions.create(9 model="qwen/qwen3.7-flash",10 messages=[{"role": "user", "content": "Hello"}],11)1213print(response.choices[0].message.content)1import OpenAI from "openai";23const client = new OpenAI({4 baseURL: "https://www.realrelay.ai/v1",5 apiKey: process.env.REALRELAY_API_KEY,6});78const response = await client.chat.completions.create({9 model: "qwen/qwen3.7-flash",10 messages: [{ role: "user", content: "Hello" }],11});1213console.log(response.choices[0].message.content);1curl https://www.realrelay.ai/v1/chat/completions \2 -H "Authorization: Bearer $REALRELAY_API_KEY" \3 -H "Content-Type: application/json" \4 -d '{5 "model": "qwen/qwen3.7-flash",6 "messages": [{ "role": "user", "content": "Hello" }]7 }'Prices and availability as published on Oct 4, 2026.
Capabilities
Built-in toolsReasoningStructured outputTool use
Input
TextImageVideo
Output
Text
