Head-to-head

GPT-5.6 Luna vs GPT-5.3 Codex

Two OpenAI models with different jobs — a fast, low-cost general tier versus an agentic coding specialist.

Updated August 28, 2026 · Independent comparison — GPT-5.6 Luna, GPT-5.3 Codex are separate products.

G

GPT-5.6 Luna

OpenAI's cheapest, fastest GPT-5.6 tier

GPT-5.6 Luna is the fastest, most affordable tier of OpenAI's GPT-5.6 family, sitting below the Sol flagship and the Terra mid-tier. It's a reasoning model with a ~1.05M-token context and a 128K output ceiling, and after an 80% price cut on 2026-07-30 it costs just $0.20/$1.20 per million input/output tokens ($0.02 cached input; $0.40/$1.80 in long-context mode). OpenAI positions it for cost-sensitive, high-volume work such as classification, summarization, routing, and real-time apps, and it became the default chat model on the Free and Go plans in early August 2026. It accepts text and image input, supports reasoning effort from none through max, and carries a February 2026 knowledge cutoff.

G

GPT-5.3 Codex

OpenAI's agentic coding-specialized model

GPT-5.3 Codex is OpenAI's agentic coding specialist, tuned for long-horizon software tasks inside the Codex CLI, the Codex IDE extension, and Codex Cloud. It's priced at $1.75/$14.00 per million tokens with a 90% cache discount ($0.175 cached input); OpenAI has not published its context window or max output. It supports low/medium/high/xhigh reasoning effort, and on Artificial Analysis it posts an intelligence index around 46 and streams at roughly 123 tokens per second, fast for a reasoning coder. It accepts text, images, and files, and carries an August 2025 knowledge cutoff.

Bottom line

Two OpenAI models with different jobs — a fast, low-cost general tier versus an agentic coding specialist. Choose GPT-5.6 Luna if you need the cheapest GPT-5.6 tier for high-volume classification, summarization, routing, or real-time features; choose GPT-5.3 Codex if you're doing agentic, long-horizon coding in the Codex CLI, the IDE extension, or Codex Cloud.

Context & Output

FeatureGPT-5.6 LunaGPT-5.3 Codex
Context window~1,050,000 tokensNot published
Max output tokens128,000Not published
Input modalitiesText, imageText, image, files (PDF)
Knowledge cutoffFebruary 16, 2026August 31, 2025

Pricing (per 1M tokens)

FeatureGPT-5.6 LunaGPT-5.3 Codex
Input$0.20$1.75
Output$1.20$14.00
Cached input$0.02$0.175 (90% off)
Long-context rate$0.40 / $1.80Not published

Coding & Reasoning

FeatureGPT-5.6 LunaGPT-5.3 Codex
Reasoning model
Reasoning effort levelsnone-max (6 levels)low / medium / high / xhigh
Coding-optimized
Artificial Analysis intelligence indexNot published~46
FocusGeneral-purpose / high-volumeCodex agentic coding

Speed & Latency

FeatureGPT-5.6 LunaGPT-5.3 Codex
Speed tierCheapest / fastest GPT-5.6 tierFast for a reasoning coder
Output speedNot published~123 tokens/sec (Artificial Analysis)
Best forHigh-volume, low-latency tasksLong-horizon coding tasks

Availability & Access

FeatureGPT-5.6 LunaGPT-5.3 Codex
Open weights
OpenAI API
Bundled inChatGPT Free/Go default, OpenCode GoCodex CLI, Codex IDE, Codex Cloud
Other surfacesAzure AI, Amazon Bedrock, CursorOpenCode Zen, Cursor

The Verdict

Choose GPT-5.6 Luna if...

  • You need the cheapest GPT-5.6 tier for high-volume classification, summarization, routing, or real-time features.
  • You want a large ~1M-token context at $0.20/$1.20 per million tokens.
  • Latency and cost per token matter more to you than top-end coding accuracy.
  • You want one fast model with adjustable reasoning effort from none to max.
  • A general-purpose model is fine — you don't need a coding specialist.

Choose GPT-5.3 Codex if...

  • You're doing agentic, long-horizon coding in the Codex CLI, the IDE extension, or Codex Cloud.
  • You want a model tuned for tool use and multi-step software tasks, not just chat.
  • You value the 90% cache discount on repeated large-context coding sessions.
  • You want fast token streaming (~123 t/s) with high reasoning effort available.
  • Coding accuracy justifies the higher $1.75/$14 per-million pricing.

Frequently Asked Questions

What is the difference between GPT-5.6 Luna and GPT-5.3 Codex?

GPT-5.6 Luna is the fastest, most affordable tier of OpenAI's GPT-5.6 family, sitting below the Sol flagship and the Terra mid-tier. GPT-5.3 Codex is OpenAI's agentic coding specialist, tuned for long-horizon software tasks inside the Codex CLI, the Codex IDE extension, and Codex Cloud.

How much do GPT-5.6 Luna and GPT-5.3 Codex cost?

GPT-5.6 Luna: $0.20. GPT-5.3 Codex: $1.75. See the pricing table above for full plan details.

Should I choose GPT-5.6 Luna or GPT-5.3 Codex?

Choose GPT-5.6 Luna if you need the cheapest GPT-5.6 tier for high-volume classification, summarization, routing, or real-time features. Choose GPT-5.3 Codex if you're doing agentic, long-horizon coding in the Codex CLI, the IDE extension, or Codex Cloud.

1DevTool1DevTool

Run OpenAI's models on your terms

Both GPT-5.6 Luna and GPT-5.3 Codex ship through OpenAI's API and Codex tooling. 1DevTool lets you drive them from the Codex CLI — alongside Claude Code, Gemini CLI, Cline, and OpenCode — in persistent terminals: bring your own OpenAI key or ChatGPT plan, switch models per task, and use the built-in HTTP client, 13-engine database client, and embedded browser. One-time $29, no subscription.