GitHub[208K]Try OpenCode

Compare AI models: Qwen3.7 Max vs Grok 4.6

  1. Qwen3.7 MaxAlibabaHollow points use fallbacks · hover for details
  2. Grok 4.6xAIHollow points use fallbacks · hover for details
ReasoningCodingCost efficiencyContext windowMultimodalTool use
Normalized model capability scores
ModelReasoningCodingCost efficiencyContext windowMultimodalTool use
Qwen3.7 Max56/10093/10026/10080/1000/10050/100 — Tool calling supported; no comparable benchmark
Grok 4.650/100 — Reasoning supported; no comparable benchmark38/10030/10073/10033/10050/100 — Tool calling supported; no comparable benchmark
Overview
Author
Context length
1M
500K
Reasoning
True
True
Input modalities
Text
Text, Image
Output modalities
Text
Text
Providers
Pricing
Input
$2.50/ 1M
$2.00/ 1M
Output
$7.50/ 1M
$6.00/ 1M
Cached input
$0.50/ 1M
$0.50/ 1M
MomentumLast 2 mo
Unique users
82K
53K
Completed sessions
87,138
110,065
Token share
0.01%
0.01%
Tokens
89B-79%
74B+100%
AUG 8OCT 2
AUG 8OCT 2
RetentionWeek 1
Returning users
45.4%2.6K user-weeks
42.7%2.6K user-weeks

Related comparisons. Other model pairs to check.