GPT-5.6 Luna vs Gemini 3.5 Flash Lite — LLM API Cost Comparison
Compare GPT-5.6 Luna (OpenAI) vs Gemini 3.5 Flash Lite (Google) on cost per million tokens, context window, and monthly spend.
Prices verified 2026-08-03 · Pricing may change — use the calculator for current estimates
- Input
- $0.2/1M tokens
- Output
- $1.2/1M tokens
- Context
- 1M tokens
- Released
- 2026-07
OpenAI cost-sensitive GPT-5.6 tier; 1.05M context
- Input
- $0.3/1M tokens
- Output
- $2.5/1M tokens
- Context
- 1M tokens
- Released
- 2026-07
Fastest low-cost Gemini 3.5 model for high-throughput tasks; 1M context
Monthly Cost by Usage Tier (70% input / 30% output ratio)
| Usage | GPT-5.6 Luna | Gemini 3.5 Flash Lite | Cheaper by |
|---|---|---|---|
| Light (1M tokens) | $0.500 | $0.960 | GPT-5.6 Luna (48%) |
| Moderate (10M tokens) | $5.00 | $9.60 | GPT-5.6 Luna (48%) |
| Heavy (100M tokens) | $50.00 | $96.00 | GPT-5.6 Luna (48%) |
| Very Heavy (1B tokens) | $500 | $960 | GPT-5.6 Luna (48%) |
Frequently Asked Questions
Which is cheaper — GPT-5.6 Luna or Gemini 3.5 Flash Lite?
For input tokens, GPT-5.6 Luna is cheaper at $0.2/1M tokens — 1.5× less than $0.3/1M. For output tokens, GPT-5.6 Luna wins at $1.2/1M vs $2.5/1M. At heavy workloads (100M tokens/month), the cost difference can be significant.
What is the context window difference between GPT-5.6 Luna and Gemini 3.5 Flash Lite?
GPT-5.6 Luna supports 1,050,000 tokens per request; Gemini 3.5 Flash Lite supports 1,048,576 tokens. GPT-5.6 Luna wins on context length, making it better for long documents, large codebases, or extended conversations without chunking.
When should I choose GPT-5.6 Luna over Gemini 3.5 Flash Lite?
Choose GPT-5.6 Luna (OpenAI) if you prefer OpenAI's ecosystem, tooling, or reliability track record. OpenAI cost-sensitive GPT-5.6 tier; 1.05M context. Choose Gemini 3.5 Flash Lite (Google) if Fastest low-cost Gemini 3.5 model for high-throughput tasks; 1M context. the price/performance fits your workload better. Use this calculator to find the break-even point for your exact token volume.
How much does 1 billion tokens cost on GPT-5.6 Luna vs Gemini 3.5 Flash Lite?
At 700M input + 300M output tokens (1B total): GPT-5.6 Luna costs $500; Gemini 3.5 Flash Lite costs $960. The difference is $460/billion tokens at this 70/30 input/output ratio.