Qwen 3 14b is a 14b model. Supports switching between thinking and non thinking: trigger thinking with /think and /no_think anywhere in a prompt or system message to toggle chain-of-thought reasoning.
Context Window
41.0K
Max Output
32.8K
Input Price (Auto)
$0.080/1M
Output Price (Auto)
$0.24/1M
Cache Read (Auto)
$0.040/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
6.8
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
47.0%
Better than 22% of models compared
HLE
Humanity's Last Exam
4.1%
Better than 16% of models compared
IFBench
Instruction-following benchmark
23.9%
Better than 6% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
32.2%
Better than 40% of models compared
AA-LCR
Long context reasoning evaluation
0.0%
Better than 6% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
SciCode
Python programming for scientific computing
26.5%
Better than 31% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
5.3%
Better than 32% of models compared
LiveCodeBench
Contamination-free coding benchmark
28.0%
Better than 31% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
58.0%
Better than 55% of models compared
AIME
American Invitational Mathematics Examination
28.0%
Better than 54% of models compared
Math-500
Diverse mathematical problem solving benchmark
87.1%
Better than 57% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
67.5%
Better than 30% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
13.3%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
92.3%
Last updated Aug 18, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Qwen 3 14b with similar models from the same provider or model family.
Qwen 3.8 27B Fable
qwen/qwen3.8-27b-fableQwen 3.8 27B Fable is an open-weight multimodal creative finetune for expressive dialogue, long-form storytelling, character work, and roleplay.
Qwen3.8 2.4T A95B (Max)
qwen/qwen3.8-2.4t-a95bThis is the same underlying model as Qwen3.8 Max, exposed under its architecture-based 2.4T A95B name for easier discovery. It uses the identical routing, pricing, capabilities, and non-thinking mode.
Qwen 3.8 27B Uncensored Thinking
qwen/qwen3.8-27b-uncensored:thinkingQwen 3.8 27B Uncensored with thinking enabled for more deliberate creative work, coding, multimodal analysis, tool use, and long-context problem solving.
Qwen 3.6 35B A3B Uncensored Thinking
qwen/qwen3.6-35b-a3b-uncensored:thinkingQwen 3.6 35B A3B Uncensored with thinking enabled for more deliberate coding, multimodal analysis, tool use, and complex chat tasks.
Qwen 3.8 27B Obliterated
qwen/qwen3.8-27b-obliteratedQwen 3.8 27B Obliterated is an open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.
Qwen 3.8 27B Obliterated Thinking
qwen/qwen3.8-27b-obliterated:thinkingQwen 3.8 27B Obliterated with thinking enabled for more deliberate creative work, coding, multimodal analysis, tool use, and long-context problem solving.
