Bitcoin.com AI logo
Provider logo

Qwen: QvQ Max

qvq-max
Provider logo

Qwen: QvQ Max

qvq-max

QvQ Max is the top model of the Qwen series. QvQ Max is capable of thinking and reasoning, can achieve significantly enhanced performance especially on hard problems.

Added Mar 28, 2025

Context Window

128.0K

Max Output

8.2K

Input Price (Auto)

$1.20/1M

Output Price (Auto)

$4.80/1M

Cache Read (Auto)

$0.60/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Qwen: QvQ Max with similar models from the same provider or model family.

Qwen3.8 2.4T A95B (Max)

qwen/qwen3.8-2.4t-a95b

This is the same underlying model as Qwen3.8 Max, exposed under its architecture-based 2.4T A95B name for easier discovery. It uses the identical routing, pricing, capabilities, and non-thinking mode.

Qwen3 Max

qwen/qwen3-max

Qwen3 Max. The latest Qwen 3 model (5 september 2025). Higher accuracy in coding and science, better instruction following, and optimized for tool calling.

Qwen 2.5 Max

qwen-max

Qwen 2.5 Max is the upgraded version of Qwen Max, beating GPT-4o, Deepseek V3 and Claude 3.5 Sonnet in benchmarks.

Qwen3.8 Max 0902

alibaba/qwen3.8-max-0902

Qwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.

Qwen3.8 Max

qwen3.8-max

Qwen3.8 Max is Qwen's 2.4T-parameter flagship model for coding, knowledge work, full-stack development, data analysis, and long-running agent workflows in non-thinking mode. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.

Qwen3.8 Max Thinking

qwen3.8-max:thinking

Qwen3.8 Max Thinking enables generation-time reasoning for deeper coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.