Gemini 3.1 Flash-Lite Preview vs GPT-5.6 Luna
Side-by-side specs, pricing and capabilities. Both models run on Clade under one subscription, so you can switch between them in the same conversation.
Gemini 3.1 Flash-Lite Preview comes from Google and GPT-5.6 Luna from OpenAI. The practical difference for most people is context, price, and which input types each one accepts.
GPT-5.6 Luna holds more in a single conversation — 1.1M tokens against 1.0M — which matters for long documents and large codebases.
GPT-5.6 Luna is the cheaper of the two on input tokens at $0.20 / 1M tokens.
| Specification | Gemini 3.1 Flash-Lite Preview | GPT-5.6 Luna |
|---|---|---|
| Provider | OpenAI | |
| Model ID | gemini-3.1-flash-lite-preview | gpt-5.6-luna |
| Context window | 1.0M | 1.1M |
| Max output | 66K | 128K |
| Input price | $0.25 / 1M tokens | $0.20 / 1M tokens |
| Output price | $1.50 / 1M tokens | $0.02 / 1M tokens |
| Knowledge cutoff | 2025-01 | — |
| Input types | text, image, video, audio, file | text, image, file, audio |
| Output types | text | text |
| Plan | Free | Free |
Which should you use?
Pick Gemini 3.1 Flash-Lite Preview when…
- You need video input, which the other model does not accept.
Pick GPT-5.6 Luna when…
- You need the larger context window — 1.1M against 1.0M.
- Cost matters: input runs at $0.20 / 1M tokens versus $0.25 / 1M tokens.
- You want longer single responses — up to 128K output tokens.
Related comparisons
Try both on Clade
You do not have to choose. One Clade subscription gives you Gemini 3.1 Flash-Lite Preview, GPT-5.6 Luna, and every other model on the platform.
