| GPT-5 mini (flex), input |
125 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5 mini (flex), output |
1,000 per million tokens |
Tokens the model writes in its answer. |
| GPT-5 mini (flex), read from cache |
12.5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5 mini (standard), input |
250 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5 mini (standard), output |
2,000 per million tokens |
Tokens the model writes in its answer. |
| GPT-5 mini (standard), read from cache |
25 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5 nano (flex), input |
25 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5 nano (flex), output |
200 per million tokens |
Tokens the model writes in its answer. |
| GPT-5 nano (flex), read from cache |
2.5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5 nano (standard), input |
50 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5 nano (standard), output |
400 per million tokens |
Tokens the model writes in its answer. |
| GPT-5 nano (standard), read from cache |
5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.2 (flex), input |
875 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.2 (flex), output |
7,000 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.2 (flex), read from cache |
87.5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.2 (standard), input |
1,750 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.2 (standard), output |
14,000 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.2 (standard), read from cache |
175 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.4 (flex), input |
2,500 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.4 (flex), output |
11,250 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.4 (flex), read from cache |
250 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.4 (standard), input |
5,000 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.4 (standard), output |
22,500 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.4 (standard), read from cache |
500 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.4 mini (flex), input |
375 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.4 mini (flex), output |
2,250 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.4 mini (flex), read from cache |
37.5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.4 mini (standard), input |
750 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.4 mini (standard), output |
4,500 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.4 mini (standard), read from cache |
75 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.4 nano (flex), input |
100 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.4 nano (flex), output |
625 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.4 nano (flex), read from cache |
10 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.4 nano (standard), input |
200 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.4 nano (standard), output |
1,250 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.4 nano (standard), read from cache |
20 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.5 (flex), input |
5,000 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.5 (flex), output |
22,500 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.5 (flex), read from cache |
500 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| GPT-5.5 (standard), input |
10,000 per million tokens |
Tokens the model reads from the prompt you send. |
| GPT-5.5 (standard), output |
45,000 per million tokens |
Tokens the model writes in its answer. |
| GPT-5.5 (standard), read from cache |
1,000 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3 Flash (batch), input |
250 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3 Flash (batch), output |
1,500 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3 Flash (batch), read from cache |
50 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3 Flash (flex), input |
250 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3 Flash (flex), output |
1,500 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3 Flash (flex), read from cache |
50 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3 Flash (standard), input |
500 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3 Flash (standard), output |
3,000 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3 Flash (standard), read from cache |
50 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.1 Flash Lite (batch), input |
125 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.1 Flash Lite (batch), output |
750 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.1 Flash Lite (batch), read from cache |
12.5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.1 Flash Lite (flex), input |
125 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.1 Flash Lite (flex), output |
750 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.1 Flash Lite (flex), read from cache |
12.5 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.1 Flash Lite (standard), input |
250 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.1 Flash Lite (standard), output |
1,500 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.1 Flash Lite (standard), read from cache |
25 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.1 Pro (flex), input |
2,000 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.1 Pro (flex), output |
9,000 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.1 Pro (flex), read from cache |
400 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.1 Pro (standard), input |
4,000 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.1 Pro (standard), output |
18,000 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.1 Pro (standard), read from cache |
400 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.5 Flash (batch), input |
750 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.5 Flash (batch), output |
4,500 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.5 Flash (batch), read from cache |
75 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.5 Flash (flex), input |
750 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.5 Flash (flex), output |
4,500 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.5 Flash (flex), read from cache |
80 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |
| Gemini 3.5 Flash (standard), input |
1,500 per million tokens |
Tokens the model reads from the prompt you send. |
| Gemini 3.5 Flash (standard), output |
9,000 per million tokens |
Tokens the model writes in its answer. |
| Gemini 3.5 Flash (standard), read from cache |
150 per million tokens |
Tokens the model reads from a cached prompt. Cheaper than a fresh read. |