# Qwen 3 14B

> Canonical: https://www.overmindlab.ai/models/qwen-3-14b

Mid-tier dense Qwen 3 for parallel experiments against compact and large.

| Property | Value |
| --- | --- |
| Model ID | Qwen/Qwen3-14B |
| Provider | Qwen |
| Group | Qwen 3 |
| Parameters | 14B |
| Tier | Mid |
| Context window | 128K |
| Max training context | 40K |
| Tool calling | Yes |
| Training methods | — |
| Training cost | from $1.25 |
| Serving cost (per 1M output tokens) | from $10.00 |

## Overview

Qwen 3 gives you the widest size ladder in the catalogue: 0.6B through 32B, plus a code-specialised mixture-of-experts model. Hold the family constant, vary only size across parallel experiments, and let the benchmark pick the winner against your production model.

The widest size ladder in the catalogue, 0.6B through 32B, plus a code-specialised mixture-of-experts model. Useful when you want to hold the family constant and vary only size across parallel experiments.

## Good for

- Side-by-side size sweeps in one family
- Code agents (via Qwen 3 Coder 30B)
- Compact bases when tool calling is optional

Train this model on your agent's data: https://docs.overmindlab.ai/models/training.md

## Related models

- [Qwen 3 4B](https://www.overmindlab.ai/models/qwen-3-4b): 4B, 128K context
- [Qwen 3 1.7B](https://www.overmindlab.ai/models/qwen-3-1-7b): 1.7B, 128K context
- [Qwen 3 0.6B](https://www.overmindlab.ai/models/qwen-3-0-6b): 0.6B, 128K context
- [Qwen 3 8B](https://www.overmindlab.ai/models/qwen-3-8b): 8B, 128K context