Licensing · Open source

Open source LLM pricing

Every open-source LLM with license, cheapest API provider, and rough self-host cost. Weights on HuggingFace, API on OpenRouter, DeepInfra, Together, Fireworks, or self-host.

Models243
Cheapest API$0.00
LicensesApache · MIT · Custom
What this page is
This page lists every open-source LLM with priced API access. For each, we show the license, cheapest API on the market, and a rough self-host cost estimate based on parameter count. Weights are downloadable from HuggingFace. APIs are offered by OpenRouter, DeepInfra, Together, Fireworks, Groq, Cerebras, and others. For very high volume, self-hosting on reserved GPUs is cheaper; for anything under 100M tokens per day, API access usually wins on total cost.

Cheapest input price first.

#ModelCheapest API
1Gemma 3 12B (free)$0.00/M
2Gemma 3 27B (free)$0.00/M
3Gemma 3 4B (free)$0.00/M
4Gemma 3n 2B (free)$0.00/M
5Gemma 3n 4B (free)$0.00/M
6Gemma 4 26B A4B (free)$0.00/M
7Gemma 4 31B (free)$0.00/M
8GLM 4.5 Air (free)$0.00/M
9gpt-oss-120b (free)$0.00/M
10gpt-oss-20b (free)$0.00/M
11Hermes 3 405B Instruct (free)$0.00/M
12Laguna M.1 (free)$0.00/M
13Laguna XS.2 (free)$0.00/M
14LFM2.5-1.2B-Instruct (free)$0.00/M
15LFM2.5-1.2B-Thinking (free)$0.00/M
16Llama 3.2 3B Instruct (free)$0.00/M
17Llama 3.3 70B Instruct (free)$0.00/M
18MiniMax M2.5 (free)$0.00/M
19Mistral Small 3.1 24B (free)$0.00/M
20Nemotron 3 Nano 30B A3B (free)$0.00/M
21Nemotron 3 Nano Omni (free)$0.00/M
22Nemotron 3 Super (free)$0.00/M
23Nemotron 3 Ultra (free)$0.00/M
24Nemotron 3.5 Content Safety (free)$0.00/M
25Nemotron Nano 12B 2 VL (free)$0.00/M
26Nemotron Nano 9B V2 (free)$0.00/M
27Nex-N2-Pro (free)$0.00/M
28North Mini Code (free)$0.00/M
29Qwen3 4B (free)$0.00/M
30Qwen3 Coder 480B A35B (free)$0.00/M
31Qwen3 Next 80B A3B Instruct (free)$0.00/M
32Qwen3.6 Plus Preview (free)$0.00/M
33Step 3.5 Flash (free)$0.00/M
34Trinity Large Preview (free)$0.00/M
35Trinity Mini (free)$0.00/M
36Uncensored (free)$0.00/M
37LFM2-2.6B$0.01/M
38LFM2-8B-A1B$0.01/M
39Granite 4.0 Micro$0.02/M
40Llama 3.1 8B Instruct$0.02/M
41Mistral Nemo$0.02/M
42Llama 3.2 1B Instruct$0.03/M
43gpt-oss-20b$0.03/M
44Gemma 2 9B$0.03/M
45LFM2-24B-A2B$0.03/M
46Qwen2.5 Coder 7B Instruct$0.03/M
47Qwen-Turbo$0.03/M
48gpt-oss-120b$0.04/M
49Llama 3 8B Lunaris$0.04/M
50Nemotron Nano 9B V2$0.04/M
51Qwen2.5 7B Instruct$0.04/M
52Trinity Mini$0.04/M
53Qwen3 30B A3B Instruct 2507$0.05/M
54Gemma 3 12B$0.05/M
55Gemma 3 4B$0.05/M
56Granite 4.1 8B$0.05/M
57Llama 3.2 3B Instruct$0.05/M
58Mistral Small 3$0.05/M
59Nemotron 3 Nano 30B A3B$0.05/M
60Olmo 2 32B Instruct$0.05/M
61Gemma 3n 4B$0.06/M
62Gemma 4 26B A4B $0.06/M
63GLM 4.7 Flash$0.06/M
64MythoMax 13B$0.06/M
65Hy3 preview$0.06/M
66Qwen3.5-Flash$0.07/M
67ERNIE 4.5 21B A3B$0.07/M
68ERNIE 4.5 21B A3B Thinking$0.07/M
69Phi 4$0.07/M
70Qwen3 Coder 30B A3B Instruct$0.07/M
71gpt-oss-safeguard-20b$0.07/M
72Mistral Small 3.2 24B$0.07/M
73Gemma 3 27B$0.08/M
74Nemotron 3 Super$0.08/M
75Phi 4 Mini Instruct$0.08/M
76Qwen3 32B$0.08/M
77DeepSeek V4 Flash$0.08/M
78MiMo-V2-Flash$0.09/M
79Qwen3 235B A22B Instruct 2507$0.09/M
80Qwen3 Next 80B A3B Instruct$0.09/M
81Tongyi DeepResearch 30B A3B$0.09/M
82Qwen3 Next 80B A3B Thinking$0.10/M
83Devstral Small 1.1$0.10/M
84Laguna XS.2$0.10/M
85Llama 3.3 70B Instruct$0.10/M
86Llama 4 Scout$0.10/M
87Ministral 3 3B 2512$0.10/M
88Mistral Small Creative$0.10/M
89Qwen3 14B$0.10/M
90Qwen3.5-9B$0.10/M
91Reka Edge$0.10/M
92Reka Flash 3$0.10/M
93Step 3.5 Flash$0.10/M
94UI-TARS 7B $0.10/M
95Voxtral Small 24B 2507$0.10/M
96Qwen3 VL 32B Instruct$0.10/M
97MiMo-V2.5$0.10/M
98Mistral 7B Instruct v0.1$0.11/M
99Qwen3 Coder Next$0.11/M
100Qwen3 8B$0.12/M
101Qwen3 VL 8B Instruct$0.12/M
102Qwen3 VL 8B Thinking$0.12/M
103Gemma 4 31B$0.12/M
104Qwen3 30B A3B$0.12/M
105GLM 4.5 Air$0.13/M
106Hermes 4 70B$0.13/M
107Qwen3 30B A3B Thinking 2507$0.13/M
108Qwen3 VL 30B A3B Instruct$0.13/M
109Qwen3 VL 30B A3B Thinking$0.13/M
110DeepSeek V3.1 Nex N1$0.14/M
111Qwen VL Plus$0.14/M
112ERNIE 4.5 VL 28B A3B$0.14/M
113Hermes 2 Pro - Llama-3 8B$0.14/M
114Hunyuan A13B Instruct$0.14/M
115Llama 3 8B Instruct$0.14/M
116Qwen3.5-35B-A3B$0.14/M
117Qwen3.6 35B A3B$0.14/M
118Qwen3 235B A22B Thinking 2507$0.15/M
119Llama 4 Maverick$0.15/M
120MiniMax M2.5$0.15/M
121Ministral 3 8B 2512$0.15/M
122Mistral Small 4$0.15/M
123Olmo 3 32B Think$0.15/M
124Olmo 3.1 32B Think$0.15/M
125QwQ 32B$0.15/M
126Rnj 1 Instruct$0.15/M
127Llama Guard 4 12B$0.18/M
128Qwen3.6 Flash$0.19/M
129Qwen3 Coder Flash$0.20/M
130Qwen3.5-27B$0.20/M
131INTELLECT-3$0.20/M
132Laguna M.1$0.20/M
133LongCat Flash Chat$0.20/M
134MiniMax-01$0.20/M
135Ministral 3 14B 2512$0.20/M
136Nemotron Nano 12B 2 VL$0.20/M
137Olmo 3.1 32B Instruct$0.20/M
138Qwen2.5 VL 32B Instruct$0.20/M
139Qwen3 VL 235B A22B Instruct$0.20/M
140Saba$0.20/M
141Step 3.7 Flash$0.20/M
142DeepSeek V3$0.20/M
143DeepSeek V3.1$0.21/M
144DeepSeek V3.2$0.21/M
145Qwen3 Coder 480B A35B$0.22/M
146DeepSeek V3 0324$0.24/M
147MiniMax M2.7$0.24/M
148Rocinante 12B$0.25/M
149Trinity Large Thinking$0.25/M
150MiniMax M2$0.26/M
151Qwen Plus 0728$0.26/M
152Qwen Plus 0728 (thinking)$0.26/M
153Qwen-Plus$0.26/M
154Qwen3 VL 235B A22B Thinking$0.26/M
155Qwen3.5 Plus 2026-02-15$0.26/M
156Qwen3.5-122B-A10B$0.26/M
157DeepSeek V3.1 Terminus$0.27/M
158DeepSeek V3.2 Exp$0.27/M
159ERNIE 4.5 300B A47B $0.28/M
160Qwen3.6 27B$0.28/M
161R1 Distill Qwen 32B$0.29/M
162Codestral 2508$0.30/M
163Cydonia 24B V4.1$0.30/M
164DeepSeek R1T2 Chimera$0.30/M
165GLM 4.6V$0.30/M
166MiniMax M2.1$0.30/M
167MiniMax M3$0.30/M
168Qwen3.5 Plus 2026-04-20$0.30/M
169Qwen3.7 Plus$0.32/M
170Qwen3.6 Plus$0.33/M
171Llama 3.2 11B Vision Instruct$0.34/M
172Mistral Small 3.1 24B$0.35/M
173Qwen2.5 72B Instruct$0.36/M
174Kimi K2.5$0.38/M
175Qwen3.5 397B A17B$0.39/M
176DeepSeek V3.2 Speciale$0.40/M
177Devstral 2 2512$0.40/M
178Devstral Medium$0.40/M
179GLM 4.7$0.40/M
180Llama 3.1 70B Instruct$0.40/M
181Llama 3.3 Nemotron Super 49B V1.5$0.40/M
182Mistral Medium 3$0.40/M
183Mistral Medium 3.1$0.40/M
184UnslopNemo 12B$0.40/M
185ERNIE 4.5 VL 424B A47B $0.42/M
186GLM 4.6$0.43/M
187DeepSeek V4 Pro$0.43/M
188MiMo-V2.5-Pro$0.43/M
189ReMM SLERP 13B$0.45/M
190Qwen3 235B A22B$0.46/M
191Llama Guard 3 8B$0.48/M
192Mistral Large 3 2512$0.50/M
193Nemotron 3 Ultra$0.50/M
194Nex-N2-Pro$0.50/M
195R1 0528$0.50/M
196Llama 3 70B Instruct$0.51/M
197Qwen VL Max$0.52/M
198Mixtral 8x7B Instruct$0.54/M
199Skyfall 36B V2$0.55/M
200Kimi K2 0711$0.57/M
201GLM 4.5$0.60/M
202GLM 4.5V$0.60/M
203GLM 5$0.60/M
204Kimi K2 0905$0.60/M
205Kimi K2 Thinking$0.60/M
206Llama 3.1 Nemotron Ultra 253B v1$0.60/M
207Kimi K2.7 Code$0.61/M
208WizardLM-2 8x22B$0.62/M
209Gemma 2 27B$0.65/M
210Llama 3.3 Euryale 70B$0.65/M
211Qwen3 Coder Plus$0.65/M
212Kimi K2.6$0.66/M
213Qwen2.5 Coder 32B Instruct$0.66/M
214Aion-1.0-Mini$0.70/M
215Hermes 3 70B Instruct$0.70/M
216R1$0.70/M
217Qwen3 Max$0.78/M
218Qwen3 Max Thinking$0.78/M
219CodeLLaMa 7B Instruct Solidity$0.80/M
220Llemma 7b$0.80/M
221Qwen2.5 VL 72B Instruct$0.80/M
222R1 Distill Llama 70B$0.80/M
223Llama 3.1 Euryale 70B v2.2$0.85/M
224GLM 5.1$0.97/M
225GLM 5.2$1.00/M
226Hermes 3 405B Instruct$1.00/M
227Hermes 4 405B$1.00/M
228Qwen-Max $1.04/M
229Qwen3.6 Max Preview$1.04/M
230Llama 3.1 Nemotron 70B Instruct$1.20/M
231Qwen3.7 Max$1.25/M
232Llama 3 Euryale 70B v2.1$1.48/M
233Mistral Medium 3.5$1.50/M
234Jamba Large 1.7$2.00/M
235Mistral Large$2.00/M
236Mistral Large 2407$2.00/M
237Mistral Large 2411$2.00/M
238Mixtral 8x22B Instruct$2.00/M
239Pixtral Large 2411$2.00/M
240Command A$2.50/M
241Llama 3.1 70B Hanami x1$3.00/M
242Magnum v4 72B$3.00/M
243Goliath 120B$3.75/M
Low volume
Under 1M tokens/day

Always pick API. A single GPU hour wipes out weeks of API spend at this volume.

Mid volume
10M to 100M tokens/day

Depends on model size. Small models (under 30B) are cheaper via API. Large models (200B+) favor dedicated GPUs.

High volume
1B+ tokens/day

Self-host wins, if utilization stays near 100%. Use reserved instances and bundle across workloads.

Cheapest
Gemma 3 12B (free)
$0.00/M
$ per 1M input tokens
Why the gap

Within OSS, the price range is driven by parameter count (bigger = more expensive to serve) and provider margins. Smaller dense models and heavily quantized serves sit at the bottom.

Most expensive
Goliath 120B
$3.75/M
$ per 1M input tokens
The weights are publicly downloadable under some license (Apache 2.0, MIT, Llama Community, Qwen License, DeepSeek License, etc.). Not every "open" license is actually OSI-compliant · always read the license before commercial use.