Model
Model explorer
Qwen3-235B-A22B-Thinking-2507
OPENAlibaba · Qwen3 family · released Jul 25, 2025
Paired always-on reasoning ('Thinking-only') sibling of Qwen3-235B-A22B-Instruct-2507, released 4 days later. Enforces <think> mode, extends max output to 81,920 tokens, native 262,144-token context. Same 235B/22B-active MoE backbone as the Instruct-2507 release already in corpus, but a distinct fine-tune (not a shared-weights effort tier), so tracked as its own standalone entry.
ReasoningCodingVisionFunction callingTool useAgentic
1827.4
Elo · rank #96
Parameters
235B
Active params
22B (MoE)
Context
256K tokens
Architecture
MoE, 128 experts / 8 active
License
Apache 2.0
Languages
119+
API price (in/out)
$0.23 / $2.3
Modalities
text
Benchmark results
Bar shows position within the tracked field; marker = field best
best: GPT-5.2 · 100.0%
best: Qwen3-Max · 86.1%
best: Claude Fable 5 · 1505
best: Hunyuan-A13B · 78.3%
best: DeepSeek-V4-Pro (Think Max) · 3206
best: GPT-6 Astra · 96.0%
best: Claude Opus 5 · 64.7%
best: Gemma 4 26B A4B · 98.5%
best: DeepSeek-V4-Pro (Think Max) · 93.5%
best: Claude Fable 5 · 91.5%
best: Qwen3.7-Max · 95.0%
best: Qwen3.7-Max · 73.6%
best: Claude Opus 4.6 · 99.3%
Run it locally
VRAM @ Q4
—
VRAM @ FP16
—
Fits on (Q4)
Multi-node cluster required
Throughput data unavailable.
Quantizations
—
Fine-tune it
PermissiveQLoRA159.8 GB1× B200 192GB
LoRA500.6 GB4× H200 141GB
Full fine-tune3771.8 GBbeyond 8× B200
QLoRA SFT on ~10k samples ≈ $7.87 (1× B200 192GB)
Qwen3 family
Elo progression across releases
Qwen3-14B (Non-Thinking)Apr 20251060.2Qwen3-14B (Thinking)Apr 20251382.9Qwen3-235B-A22B (Non-Thinking)Apr 20251213.0Qwen3-235B-A22B (Thinking)Apr 20251516.1Qwen3-30B-A3B (Non-Thinking)Apr 20251070.4Qwen3-30B-A3B (Thinking)Apr 20251452.3Qwen3-32B (Non-Thinking)Apr 20251097.1Qwen3-32B (Thinking)Apr 20251484.0Qwen3-8B (Non-Thinking)Apr 2025900.6Qwen3-8B (Thinking)Apr 20251311.4Qwen3-235B-A22B-Instruct-2507Jul 20251559.3Qwen3-235B-A22B-Thinking-2507Jul 20251827.4
API price $0.23/$2.3 · each benchmark row carries its own source badge (see methodology)