Qwen3.7 Plus

4.05
Text + image input, text output. Upgraded vision-language capabilities, with full agentic strength in coding and tool use retained.
Advertisement 728 × 90
CompanyAlibaba
Context1M
Released2026-06
Updated2026-08-04

Qwen3.7 Plus Overview

Qwen3.7-Plus is a multimodal agent model released by Alibaba on June 2, 2026. Positioned as a unified vision-language agent foundation, it ranks among the top 5 globally and No. 1 in China on the prestigious Vision Arena benchmark.

Qwen3.7 Plus Pricing

PlanPriceDescription
API International (Standard Context) Input: $0.40 / Output: $1.60 per 1M tokens Supports a 1M-token context window with multimodal capabilities for text, image, and video inputs. Includes Implicit Cache support with cached input pricing at $0.08 per 1M tokens. Ideal for everyday AI applications and multimodal workloads.
API International (Long Context) Input: $1.20 / Output: $4.80 per 1M tokens Designed for extended context workloads from 256K to 1M tokens, using tiered pricing. Suitable for long-document analysis, complex reasoning, and large-scale knowledge processing.
New User Free Credits Free (1M tokens) New Alibaba Cloud users receive 1M free token credits, available for 90 days after registration. Ideal for testing, prototyping, and early development.

Qwen3.7 Plus Key Features

• GUI Agent
Understands graphical interfaces on mobile and desktop devices, and performs actions like clicking and typing just like a human. For example, it can open a stock trading app, observe its interface, and then replicate a fully functional clone from scratch.

• Visual Programming
Generates runnable web pages, SVG animations, or application code directly from images, sketches, or even videos. Simply provide a design mockup, and it can generate a complete frontend page for you.

• Visual Agent
Combines tools to solve complex problems. It can invoke a code interpreter to analyze images for "spot the difference" games or maze solving, and leverage web search to accurately identify equipment functions and parameters from a mechanical drawing.

• Multimodal Reasoning & Understanding
Demonstrates strong comprehension of images and video, scoring above Gemini 3.1 Pro on visual reasoning benchmarks. It also understands complex real-world scenes such as autonomous driving scenarios.

• Long-Horizon Task Autonomy
Integrates "see, think, write, do, and verify" into a unified workflow. In official demonstrations, it ran autonomously for over 11 hours, completing the entire development lifecycle of an English learning app—from requirements analysis and coding to testing and release.

Summary

• Individual Developers & Programming Enthusiasts: Low-cost API access for daily coding, prototyping, and learning new technologies.

• SMEs & Startups: Affordable AI integration for automation, customer support, and content production.

• Business Professionals & Content Creators: Boosting daily productivity in writing, report processing, and multimedia material handling.

• Students & Researchers: Assisting with studies, paper polishing, and data organization for moderately complex logical tasks.

Comments (0)

Leave a comment

Advertisement 728 × 90

Alibaba Model Comparison

Model Context Pricing API Released Global Heat
Qwen3.7 Plus
1M Paid YES 2026-06
81/100
1M YES 2026-08
88/100
262K YES 2026-08
90/100
262K YES 2026-08
60/100
1M YES 2026-07
79/100

Similar Models

Qwen3-235B
90
Alibaba MoE model achieving high-quality reasoning at low cost with 22B active parameters.
Alibaba
Qwen3.8 2.4T A95B
90
What stands out to me isn’t how well it answers a single question—it’s whether it can keep a complex task going. In my own experience, hitting a wall usually isn’t about missing the basics. It’s about finding something that can pick up where you left off and move forward.
Alibaba
Qwen3.8 Max
88
Alibaba released Qwen3.8-Max on August 3. It has 2.4 trillion total parameters, activates around 95 billion parameters per inference step, and supports a 1 million token context window. The pitch is not just “better coding.” The model is supposed to handle the full process: break down requirements, write code, debug, iterate, and eventually deliver a working project.
Alibaba
Qwen3.7 Max
87
China's strongest AI of 2026, deep reasoning + autonomous execution, from conversation to getting things done.
Alibaba
Qwen3.6 Flash
81
Alibaba’s Qwen3.6 Flash has been getting a fair amount of attention among developers. After using it for a few days, my impression is pretty straightforward: it is not trying to beat the Plus model on raw capability.
Alibaba
Qwen3.7 Flash
79
Alibaba’s Qwen3.7-Flash is a high-performance lightweight AI model optimized for multimodal understanding, AI Agents, coding, and fast, cost-efficient reasoning.
Alibaba
Qwen3.8 27B
60
Alibaba’s Qwen team finally released the model the open-source community had been waiting for: Qwen3.8-27B.What caught my attention was the size: 27B parameters. That is still small enough to be realistic for local deployment on high-end consumer hardware, especially when quantized.
Alibaba
Claude Opus 5
100
Anthropic’s flagship Claude model built for complex reasoning, AI Agents, software development, and enterprise knowledge workflows.
Anthropic

Related Tools

Related News