| model name | pricing | Notes |
|---|---|---|
| gpt-4o gpt-5 gpt-5.2 | 50% of OpenAI's official pricing | 50% of OpenAI's official pricing |
| gpt-4o-mini gpt-4.1 gpt-4.1-mini gpt-4.1-nano | 75% of OpenAI's official pricing | 75% of OpenAI's official pricing |
| claude-opus-4-6 | input 1M tokens = $3.75 output 1M tokens = $18.75 | 75% of Anthropic's official pricing |
| claude-sonnet-4-6 | input 1M tokens = $2.25 output 1M tokens = $11.25 | 75% of Anthropic's official pricing |
| gpt-image-1 | same as openai official | see openai-official-pricing |
| gemini-2.5-flash-nothinking | input 1M tokens = $0.075 output 1M tokens = $0.625 | 25% of Gemini's official pricing |
| gpt-5.6-sol | input 1M tokens = $0.50 cached input 1M tokens = $0.05 output 1M tokens = $3.00 | Limited-time promotional pricing (10% of OpenAI's official pricing) |
| gpt-5.6-terra | input 1M tokens = $0.20 cached input 1M tokens = $0.02 output 1M tokens = $1.20 | Limited-time promotional pricing (10% of OpenAI's official pricing) |
| gpt-5.6-luna | input 1M tokens = $0.02 cached input 1M tokens = $0.002 output 1M tokens = $0.12 | Limited-time promotional pricing (10% of OpenAI's official pricing) |
| gpt-6-astra | input 1M tokens = $2.00 cached input 1M tokens = $0.20 output 1M tokens = $10.00 | Limited-time promotional pricing (20% of OpenAI's official pricing) |
gpt-6-astra: when the input context of a request exceeds 272,000 tokens, the entire request is billed at 2x the input and cached-input price and 1.5x the output price - input 1M tokens = $4.00, cached input 1M tokens = $0.40, output 1M tokens = $15.00. This mirrors OpenAI's own long-context pricing.