Qwen2.5-7B
by Alibaba Qwen · ChinaOne of the most widely fine-tuned small models in the open ecosystem.
Specification
The numbers
- Maker
- Alibaba Qwen
- Released
- 2024-09
- Parameters
- 7B
- Architecture
- Dense transformer
- Context window
- 131,072 tokens
- Max output
- 8,192 tokens
- Input
- Text
- Output
- Text
- Reasoning
- No
- Tool calling
- Yes
- Knowledge cutoff
- Not reported
- Licence
- Apache 2.0
- Availability
- Not reported
- Weights
- Downloadable
Cost
Price per million tokens
- Input
- $0.025 / M tokens
- Output
- $0.05 / M tokens
- Blended 3:1
- $0.031
This model has open weights, so there is no first-party price. The figures above are a representative third-party hosting rate — you can also run it yourself for the cost of the hardware.
Published scores
Benchmarks
Figures published by Alibaba Qwen or taken from a public leaderboard. Row last checked 2026-08.
| Category | Score | Rank |
|---|---|---|
| ReasoningMulti-step logic on problems that cannot be looked up | Not reported | — |
| MathsCompetition mathematics, graded on the final answer | Not reported | — |
| CodingHumanEval | 84.8% | 22 of 46 reporting the same tests |
| KnowledgeMMLU-Pro | 56.3% | 47 of 77 reporting the same tests |
| MultimodalReading charts, diagrams and photographs | Not reported | — |
| Instruction followingObeying an exact, checkable format | Not reported | — |
| Human preferenceWhich answer people pick, blind | Not reported | — |
A category averages every benchmark in it that Qwen2.5-7B reports. The rank counts only models that report the same tests, so it never compares an average over three benchmarks against an average over one.
Every reported test
| MMLU-ProKnowledge | Rank 47 of 77 models reporting | |
|---|---|---|
| GPQA DiamondReasoning | Not reported | No figure published |
| AIME 2025Maths | Not reported | No figure published |
| MATH-500Maths | Not reported | No figure published |
| SWE-bench VerifiedCoding | Not reported | No figure published |
| SWE-bench ProCoding | Not reported | No figure published |
| Terminal-Bench 2.1Coding | Not reported | No figure published |
| Frontier-Bench v0.1Reasoning | Not reported | No figure published |
| Terminal-Bench 4.0Coding | Not reported | No figure published |
| LiveCodeBenchCoding | Not reported | No figure published |
| HumanEvalCoding | Rank 22 of 46 models reporting | |
| MMMUMultimodal | Not reported | No figure published |
| IFEvalInstruction following | Not reported | No figure published |
| LMArena EloHuman preference | Not reported | No figure published |
Where these numbers come from
Every score on this page is a published figure, taken from the model's own card, system card, technical report or release post, or from a public leaderboard. CorX Labs did not run these evaluations. Most are self-reported by the lab that built the model, which means they were produced under that lab's own choice of prompt, scaffold and number of attempts — so treat them as a starting point for a shortlist, not as a settled ranking.
A score someone other than the model's maker measured is marked Independent and names its measurer. Those are the stronger numbers on this page — an outside harness has no reason to flatter anyone — and there are not many of them.
Where a figure has not been published, the cell reads Not reported rather than an estimate. Nothing here is inferred, interpolated or guessed. Each model records the month its row was last checked. Full method and caveats.
Same maker
Other models from Alibaba Qwen
| Qwen3-MaxAlibaba Qwen | Alibaba Qwen | 262K | $1.20 | $6.00 | — | 85.4% | 92.3% | 69.6% | Proprietary |
| Qwen3-Coder-480B-A35BAlibaba Qwen | Alibaba Qwen | 262K | $0.30 | $1.20 | — | — | — | 69.6% | Apache 2.0 |
| Qwen3-235B-A22BAlibaba Qwen | Alibaba Qwen | 131K | $0.20 | $0.60 | 68.2% | 71.1% | 85.7% | — | Apache 2.0 |
| Qwen3-32BAlibaba Qwen | Alibaba Qwen | 131K | $0.10 | $0.30 | — | 66.8% | 81.4% | — | Apache 2.0 |
| Qwen3-30B-A3BAlibaba Qwen | Alibaba Qwen | 131K | $0.08 | $0.290 | — | 65.8% | 80.4% | — | Apache 2.0 |
| Qwen3-14BAlibaba Qwen | Alibaba Qwen | 131K | $0.06 | $0.24 | — | 64% | 79.3% | — | Apache 2.0 |
| Qwen3-8BAlibaba Qwen | Alibaba Qwen | 131K | $0.035 | $0.138 | — | 62% | 76% | — | Apache 2.0 |
| Qwen3-4BAlibaba Qwen | Alibaba Qwen | 131K | $0.02 | $0.06 | — | 55.9% | 73.8% | — | Apache 2.0 |
| QwQ-32BAlibaba Qwen | Alibaba Qwen | 131K | $0.15 | $0.20 | — | 65.2% | 79.5% | — | Apache 2.0 |
| Qwen2.5-VL-72BAlibaba Qwen | Alibaba Qwen | 131K | $0.40 | $0.40 | — | — | — | — | Qwen License |
| Qwen2.5-Coder-32BAlibaba Qwen | Alibaba Qwen | 131K | $0.070 | $0.16 | — | — | — | — | Apache 2.0 |
| Qwen2.5-72BAlibaba Qwen | Alibaba Qwen | 131K | $0.35 | $0.40 | 71.1% | 49% | — | — | Qwen License |
| Qwen2.5-32BAlibaba Qwen | Alibaba Qwen | 131K | $0.08 | $0.20 | 69% | — | — | — | Apache 2.0 |