Vision / Qwen
Qwen3.7-Plus
API name: qwen3.7-plus
Overview
Among the Qwen3.7 series, the cost-effective Plus model builds on its robust text capabilities while delivering a comprehensive upgrade to its vision‑language abilities, all while preserving its full‑stack agent‑level intelligence for coding, tool use, and productivity workflows. Its key distinguishing feature is multi‑modal interactive hybrid agent capabilities, enabling it to perceive real‑world scenes, read screens and interact with GUIs, generate code based on visual references, and perform end‑to‑end navigation within mobile apps. This version is functionally equivalent to snapshotqwen3.7-plus-2026-05-26
- Model type: Vision
- Input: Text / Image / Video
- Output: Text
- Endpoints: chat / chat_completion
| Tier | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|
| ≤256K input tokens | $0.0003/1K | $0.0011/1K | $0.0001/1K | $0.0003/1K |
| >256K input tokens | $0.0008/1K | $0.0033/1K | $0.0002/1K | $0.0010/1K |
Billed bylen (input context token count). Coefficient is the $ / 1M tokens price. Default group price (other groups converted by discount).
Code example
Call with qwen3.7-plus :
curl https://www.starunion.net/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3.7-plus",
"messages": [{
"role": "user",
"content": [
{ "type": "text", "text": "Describe this image" },
{ "type": "image_url", "image_url": { "url": "https://example.com/photo.jpg" } }
]
}]
}'