How Many AI Tokens Can I Get for My Budget?

Enter a monthly dollar amount and instantly see how many tokens each AI model gives you. Last updated:

How far does my AI budget go? Different models have wildly different per-token rates. A $50 monthly budget buys you about 2,500 million input tokens on Llama 3.1 8B Instruct but only about 2 million on GPT-5.4 Pro. This calculator divides your budget by each model's input and output rates and ranks every model by total token allowance, so you can find the best bang for your buck.

Your Budget

$25.00 input / $25.00 output

Token Allowance per Model — $50.00/mo Budget

Sorted by total tokens (most to least). Split: 50% input / 50% output.

#ModelProviderInput TokensOutput TokensTotal Tokens
1Embed v3 Englishcohere250.0MInfinityBInfinityB
2Embed v3 Multilingualcohere250.0MInfinityBInfinityB
3Rerank v3cohere12.5MInfinityBInfinityB
4Llama 3.1 8B Instructnovita1.3B500.0M1.8B
5DeepSeek-OCR 2novita833.3M833.3M1.7B
6LFM2 24B A2Btogether833.3M208.3M1.0B
7Qwen3.5 4Bcastform833.3M166.7M1.0B
8AutoGLM-Phone-9B-Multilingualnovita714.3M181.2M895.4M
9Command R7Bcohere666.7M166.7M833.3M
10Llama 3.1 8B Instantgroq500.0M312.5M812.5M
11Mistral NeMonovita625.0M147.1M772.1M
12Gemma 3n E4B Instructtogether416.7M208.3M625.0M
13GPT-OSS 20Btogether500.0M125.0M625.0M
14GPT-5 nanoopenai500.0M62.5M562.5M
15Ministral 3Bmistral250.0M250.0M500.0M
16Qwen3 Coder 30B A3B Instructnovita357.1M92.6M449.7M
17GLM 5.3 Flashnovita333.3M100.0M433.3M
18GLM-4.7-Flashnovita357.1M62.5M419.6M
19Gemini 2.0 Flash-Litegoogle333.3M83.3M416.7M
20GPT-OSS 20Bgroq333.3M83.3M416.7M
21GPT-OSS Safeguard 20Bgroq333.3M83.3M416.7M
22Llama 3 8B Instruct Litetogether178.6M178.6M357.1M
23Devstral Small 2mistral250.0M83.3M333.3M
24Llama 4 Scouttogether250.0M83.3M333.3M
25Ministral 8Bmistral166.7M166.7M333.3M
26Mistral NeMomistral166.7M166.7M333.3M
27Pixtral 12Bmistral166.7M166.7M333.3M
28RNJ-1 Instructtogether166.7M166.7M333.3M
29Qwen3 235B A22B Instruct 2507novita277.8M43.1M320.9M
30Gemini 2.0 Flashgoogle250.0M62.5M312.5M
31Gemini 2.5 Flash-Litegoogle250.0M62.5M312.5M
32GPT-4.1 nanoopenai250.0M62.5M312.5M
33Llama 4 Scout 17B 16E Instructgroq227.3M73.5M300.8M
34DeepSeek V4 Flashnovita178.6M89.3M267.9M
35Ministral 14Bmistral125.0M125.0M250.0M
36Llama 3.3 70B Instructnovita185.2M62.5M247.7M
37Qwen3.5 9Btogether147.1M100.0M247.1M
38GLM-4.5 Airnovita192.3M29.4M221.7M
39Qwen3.8 Flashnovita166.7M53.2M219.9M
40Qwen3.8-Flashalibaba156.3M53.2M209.4M
41Command R 08-2024cohere166.7M41.7M208.3M
42GPT-4o miniopenai166.7M41.7M208.3M
43GPT-OSS 120Bgroq166.7M41.7M208.3M
44GPT-OSS 120Btogether166.7M41.7M208.3M
45Llama 4 Mavericktogether166.7M41.7M208.3M
46Mistral Small 4mistral166.7M41.7M208.3M
47Mistral 7Bmistral100.0M100.0M200.0M
48Qwen3 Next 80B A3B Instructnovita166.7M16.7M183.3M
49Llama 4 Scout Instructnovita138.9M42.4M181.3M
50Grok 4.1 Fast Non-Reasoningxai125.0M50.0M175.0M
51Grok 4.1 Fast Reasoningxai125.0M50.0M175.0M
52Qwen2.5 7B Instruct Turbotogether83.3M83.3M166.7M
53Qwen3 235B A22B Instruct 2507together125.0M41.7M166.7M
54Qwen3 VL 30B A3B Instructnovita125.0M35.7M160.7M
55Qwen3 235B A22Bnovita125.0M31.3M156.3M
56DeepSeek V3.2novita92.9M62.5M155.4M
57DeepSeek V3.2 Expnovita92.6M61.0M153.6M
58DeepSeek V4 Flash 0731deepseek113.6M37.9M151.5M
59GPT-5.6 Lunaopenai125.0M20.8M145.8M
60GPT-5.4 nanoopenai125.0M20.0M145.0M
61Qwen3 Coder Nextnovita125.0M16.7M141.7M
62Qwen MT Plusnovita100.0M33.3M133.3M
63Qwen3 32Bgroq86.2M42.4M128.6M
64Qwen 2.5 72B Instructnovita65.8M62.5M128.3M
65Llama 4 Maverick Instructnovita92.6M29.4M122.0M
66Gemma 4 31B IT Pearltogether89.3M29.1M118.4M
67Qwen3.6-35B-A3Bnovita100.8M16.8M117.6M
68DeepSeek V3.1novita92.6M25.0M117.6M
69DeepSeek V3.1 Terminusnovita92.6M25.0M117.6M
70Gemini 3.1 Flash-Litegoogle100.0M16.7M116.7M
71Qwen3.6-Flashalibaba100.0M16.7M116.7M
72DeepSeek V3 0324novita92.6M22.3M114.9M
73GPT-5 miniopenai100.0M12.5M112.5M
74Qwen3.5-35B-A3Bnovita100.0M12.5M112.5M
75Codestralmistral83.3M27.8M111.1M
76GLM-4.6Vnovita83.3M27.8M111.1M
77MiniMax M2novita83.3M20.8M104.2M
78MiniMax M2.1novita83.3M20.8M104.2M
79MiniMax M2.5together83.3M20.8M104.2M
80MiniMax M2.5novita83.3M20.8M104.2M
81MiniMax M2.7together83.3M20.8M104.2M
82MiniMax M2.7novita83.3M20.8M104.2M
83MiniMax M3minimax83.3M20.8M104.2M
84MiniMax M3together83.3M20.8M104.2M
85Qwen3 VL 235B A22B Instructnovita83.3M16.7M100.0M
86Qwen3.7-Plusalibaba78.1M19.5M97.7M
87Qwen3.5-27Bnovita83.3M10.4M93.8M
88Gemini 2.5 Flashgoogle83.3M10.0M93.3M
89Gemini 3.5 Flash-Litegoogle83.3M10.0M93.3M
90Qwen3 235B A22B Thinking 2507novita83.3M8.3M91.7M
91Gemma 4 31B ITtogether64.1M25.8M89.9M
92Qwen3 Coder 480B A35B Instructnovita65.8M16.1M81.9M
93DeepSeek V3 (Turbo)novita62.5M19.2M81.7M
94GPT-4.1 miniopenai62.5M15.6M78.1M
95DeepSeek V4 Flash 0731novita56.8M18.9M75.8M
96Devstral Medium 2mistral62.5M12.5M75.0M
97Mistral Medium 3mistral62.5M12.5M75.0M
98Llama 3.3 70B Versatilegroq42.4M31.6M74.0M
99Mixtral 8x7Bmistral35.7M35.7M71.4M
100Qwen3.5-122B-A10Bnovita62.5M7.8M70.3M
101Qwen3.8 27Bnovita59.5M8.3M67.9M
102gpt-3.5-turboopenai50.0M16.7M66.7M
103gpt-3.5-turbo-0125openai50.0M16.7M66.7M
104Magistral Smallmistral50.0M16.7M66.7M
105Mistral Large 3mistral50.0M16.7M66.7M
106DeepSeek R1 Distill Llama 70Bnovita31.3M31.3M62.5M
107Gemini 3 Flashgoogle50.0M8.3M58.3M
108GLM-4.6novita45.5M11.4M56.8M
109MiniMax M1novita45.5M11.4M56.8M
110DeepSeek V3.1together41.7M14.7M56.4M
111GLM-4.5Vnovita41.7M13.9M55.6M
112Kimi K2 Instructnovita43.9M10.9M54.7M
113GLM-4.7novita41.7M11.4M53.0M
114MiniMax M2.5 Highspeednovita41.7M10.4M52.1M
115Kimi K2 0905novita41.7M10.0M51.7M
116Kimi K2 Thinkingnovita41.7M10.0M51.7M
117DeepSeek V4 Pro 0813deepseek37.9M12.6M50.5M
118Kimi K2.5novita41.7M8.3M50.0M
119Sonarperplexity25.0M25.0M50.0M
120Nemotron 3 Ultra 550B A55Btogether41.7M6.9M48.6M
121Qwen3.5 397B A17Btogether41.7M6.9M48.6M
122Qwen3.5-397B-A17Bnovita41.7M6.9M48.6M
123Qwen3.6-27Bnovita41.7M6.9M48.6M
124Llama 3.3 70Btogether24.0M24.0M48.1M
125DeepSeek R1 (Turbo)novita35.7M10.0M45.7M
126DeepSeek R1 0528novita35.7M10.0M45.7M
127Mixtral 8x22Btogether20.8M20.8M41.7M
128Qwen 2.5 72Btogether20.8M20.8M41.7M
129Cogito v2.1 671Btogether20.0M20.0M40.0M
130DeepSeek V3together20.0M20.0M40.0M
131Gemini 3.6 Flashgoogle33.3M6.7M40.0M
132Gemini 3.7 Flashgoogle33.3M6.7M40.0M
133GPT-5.4 miniopenai33.3M5.6M38.9M
134Kimi K2.6novita31.3M7.4M38.6M
135gpt-3.5-turbo-1106openai25.0M12.5M37.5M
136Grok Build 0.1xai25.0M12.5M37.5M
137GLM-5together25.0M7.8M32.8M
138GLM-5novita25.0M7.8M32.8M
139Kimi K2.7 Codetogether26.3M6.3M32.6M
140Kimi K2.7 Codenovita26.3M6.3M32.6M
141Qwen3 VL 235B A22B Thinkingnovita25.5M6.3M31.8M
142Claude Haiku 4.5anthropic25.0M5.0M30.0M
143Grok 4.20 0309 Non-Reasoningxai20.0M10.0M30.0M
144Grok 4.20 0309 Reasoningxai20.0M10.0M30.0M
145Grok 4.20 Multi-Agent 0309xai20.0M10.0M30.0M
146Grok 4.3xai20.0M10.0M30.0M
147gpt-3.5-turbo-instructopenai16.7M12.5M29.2M
148o3-miniopenai22.7M5.7M28.4M
149o4-miniopenai22.7M5.7M28.4M
150Qwen3.7-Maxalibaba20.0M6.7M26.7M
151Qwen3.7-Maxnovita20.0M6.7M26.7M
152Kimi K2.6together20.8M5.6M26.4M
153DeepSeek V4 Pro 0813novita18.9M6.3M25.3M
154GLM-5.1novita18.1M5.7M23.8M
155GLM 5.3novita17.9M5.7M23.5M
156GLM-5.1together17.9M5.7M23.5M
157GLM-5.2together17.9M5.7M23.5M
158GLM-5.2zai17.9M5.7M23.5M
159GLM-5.2novita17.9M5.7M23.5M
160GLM-5.3zai17.9M5.7M23.5M
161DeepSeek V4 Pronovita15.6M7.8M23.4M
162Gemini 2.5 Progoogle20.0M2.5M22.5M
163GPT-5openai20.0M2.5M22.5M
164GPT-5.1openai20.0M2.5M22.5M
165DeepSeek V4 Protogether14.4M7.2M21.6M
166Mistral Medium 3.5mistral16.7M3.3M20.0M
167Gemini 3.5 Flashgoogle16.7M2.8M19.4M
168Magistral Mediummistral12.5M5.0M17.5M
169Grok 4.5xai12.5M4.2M16.7M
170Grok 4.6xai12.5M4.2M16.7M
171Mixtral 8x22Bmistral12.5M4.2M16.7M
172Pixtral Largemistral12.5M4.2M16.7M
173Qwen3.8 2.4T A95Bnovita12.5M4.2M16.7M
174Qwen3.8 Maxnovita12.5M4.2M16.7M
175GPT-5.2openai14.3M1.8M16.1M
176GPT-4.1openai12.5M3.1M15.6M
177o3openai12.5M3.1M15.6M
178Sonar Deep Researchperplexity12.5M3.1M15.6M
179Sonar Reasoning Properplexity12.5M3.1M15.6M
180Claude Sonnet 5anthropic12.5M2.5M15.0M
181Gemini 3 Progoogle12.5M2.1M14.6M
182Gemini 3.1 Progoogle12.5M2.1M14.6M
183GPT-5.6 Terraopenai12.5M2.1M14.6M
184Llama 3.1 405Btogether7.1M7.1M14.3M
185Command Acohere10.0M2.5M12.5M
186Command R+ 08-2024cohere10.0M2.5M12.5M
187GPT-4oopenai10.0M2.5M12.5M
188DeepSeek R1together8.3M3.6M11.9M
189Gemini 2.5 Pro (>200k tokens)google10.0M1.7M11.7M
190GPT-5.4openai10.0M1.7M11.7M
191Kimi K3telnyx9.3M1.9M11.1M
192Mistral Largetogether8.3M2.8M11.1M
193Claude Sonnet 4anthropic8.3M1.7M10.0M
194Claude Sonnet 4.5anthropic8.3M1.7M10.0M
195Claude Sonnet 4.6anthropic8.3M1.7M10.0M
196Kimi K3moonshot8.3M1.7M10.0M
197Kimi K3novita8.3M1.7M10.0M
198Sonar Properplexity8.3M1.7M10.0M
199GPT-5.6 Solopenai6.3M1.3M7.5M
200gpt-4o-2024-05-13openai5.0M1.7M6.7M
201Claude Opus 4.5anthropic5.0M1.0M6.0M
202Claude Opus 4.6anthropic5.0M1.0M6.0M
203Claude Opus 4.7anthropic5.0M1.0M6.0M
204Claude Opus 4.8anthropic5.0M1.0M6.0M
205Claude Opus 5anthropic5.0M1.0M6.0M
206GPT-5.5openai5.0M833.3K5.8M
207gpt-4-turbo-2024-04-09openai2.5M833.3K3.3M
208Claude Fable 5anthropic2.5M500.0K3.0M
209Claude Mythos 5anthropic2.5M500.0K3.0M
210GPT-5.5 Cyberopenai2.0M333.3K2.3M
211GPT-5.6 Cyberopenai2.0M333.3K2.3M
212o1openai1.7M416.7K2.1M
213Claude Opus 4anthropic1.7M333.3K2.0M
214Claude Opus 4.1anthropic1.7M333.3K2.0M
215GPT-5 Proopenai1.7M208.3K1.9M
216o3-proopenai1.3M312.5K1.6M
217GPT-5.2 Proopenai1.2M148.8K1.3M
218gpt-4-0613openai833.3K416.7K1.3M
219GPT-5.4 Proopenai833.3K138.9K972.2K
220GPT-5.5 Proopenai833.3K138.9K972.2K
220 models compared at $50.00/mo budget

How does this calculator work?

  1. Enter your monthly budget — the total dollar amount you want to spend on AI API calls per month.
  2. Optionally select a focus model — highlights that model in the table and shows a detailed breakdown.
  3. Choose a budget split — decide how to allocate between input and output tokens (50/50, 80/20, or 20/80).
  4. Compare the table — models are ranked from most tokens to least, so the best-value models appear first.

Methodology

Token allowance is calculated by inverting the standard cost formula:

input_tokens = (budget × split_ratio / input_rate_per_M) × 1,000,000

output_tokens = (budget × (1 - split_ratio) / output_rate_per_M) × 1,000,000

The budget split controls what fraction of your monthly spend goes toward input vs output tokens. For most chat use cases, 50/50 is a reasonable default. If you send long prompts with short replies, use 80/20. If you request long-form content generation, use 20/80.

All rates come from our daily-updated pricing database. Models with $0 rates (free tiers) are excluded from ranking.

Does the budget only work at huge volume? If your workload is steady enough to keep GPUs busy, compare API spend with the GPU break-even guide, then test RunPod pricing against your real throughput.

The RunPod route may show a $5 referral credit after the first $10 added, but verify current terms and normal hourly pricing before treating it as part of your budget.

Affiliate disclosure: this link may earn us a commission at no extra cost to you. It does not affect token allowance rankings.