Models/OpenAI/GPT-6 Luna

GPT-6 Luna

GPT-6 Luna is OpenAI’s most efficient GPT-6 model, positioned for focused, high-volume tasks. It costs $0.10 input and $0.50 output per million. Continuum hosts gpt-6-luna.

01

Identity

1,050,000Context in
128,000Context out
2026-09-22Released
ProprietaryWeights
Canonical name
GPT-6 Luna
Aliases
gpt-6-luna, continuum/gpt-6-luna
Continuum hosted id
gpt-6-luna
OpenRouter slug
Not on the 2026-08-19 OpenRouter catalog
Hugging Face
None (closed or unpublished)
Modalities
text, image
Plan mapping
Codex with ChatGPT sign-in serves it from Codex CLI 0.155.0. Pin with codex -m gpt-6-luna. Continuum also serves it on the hosted lane. See Codex models.
Closed weights Hosted on Continuum Vendor lab

The API id is gpt-6-luna. Continuum uses the bare id at the gateway and continuum/gpt-6-luna in cross-provider pickers.

The input context is 1,050,000 tokens and maximum output is 128,000 tokens. Input is text and images; output is text. The knowledge cutoff is 18 May 2026.

Reasoning effort on the API is none, low, medium, high, xhigh, or max, with medium as the default. Continuum exposes low through max.

OpenAI positions it as: “Our most efficient model for focused, high-volume tasks.”

02

Should I use this for coding agents

Use GPT-6 Luna for mechanical, well-specified work at volume: renames, formatting, extraction, and short focused edits.

No independent row for this SKU

No independent board lists this exact SKU as of 2026-08-19.

Closest benchmarked sibling: GPT-5.6 Sol, 73% ±3% Pass@1 at max on the DeepSWE official board. That is a score for the sibling, not for this SKU.

Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysis

No Intelligence Index integer fetched for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
03

When not to use it

Trust this before you buy

  • Do not hand it open-ended architecture or long investigations. It is positioned for focused tasks, not for the hardest reasoning.
  • Requests over 272K prompt tokens bill the whole request at 2x input and cache rates and 1.5x output.
  • Closed weights. No local checkpoint and no official GGUF.

Get Plus to use the Continuum-hosted GPT-6 Luna rail. Not a Mac download.

04

Artifact

Weights

OpenAI does not publish GPT-6 Luna weights. No HF Files button. Get Plus is hosted inference.

  • Official OpenAI model page, read 22 Sep 2026: $0.10 input / $0.01 cached input / $0.125 cache write / $0.50 output per million.
  • Long-context rule: above 272K prompt-input tokens the whole request bills at 2x input and cache rates and 1.5x output, which is $0.20 input and $0.75 output per million.

What you can open

Lab card and OpenRouter listing only. No Hugging Face Files button, because there is no official repo to point at.

Fetched OpenRouter catalog on 2026-08-19. hugging_face_id was empty.

05

Continuum serving

Continuum hosts gpt-6-luna on the hosted lane, and it is available through your own Codex subscription (Codex CLI 0.155.0 or later on a ChatGPT account).

Continuum hosted id gpt-6-luna

click to select

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

Official pricing per million tokens: $0.10 input, $0.50 output, $0.01 cached input, and $0.125 cache write.

Price this model Opens the pricing calculator preloaded with GPT-6 Luna.
07

Provenance

ClaimSourceAs of
API id, $0.10 / $0.01 cached / $0.125 cache write / $0.50 out, 272K request-wide surcharge, 1.05M context, 128K output, effort levels, modalities, knowledge cutoffOpenAI GPT-6 Luna model page2026-09-22
No base OpenRouter slug on the 2026-08-19 catalog; Hugging Face unlistedOpenRouter /api/v1/models2026-08-19
08

Compare, FAQ, and Get Plus

Same official DeepSWE harness (mini-swe-agent, public 113 tasks), fetched 2026-08-19. Different effort labels are the lab's own setting on that board, shown here rather than normalized.

Continue through OpenAI's model family with GPT-5.6 Luna.

FAQ

Is GPT-6 Luna the same model as GPT-5.6 Luna?

No. They are separate ids with separate prices. gpt-6-luna and gpt-5.6-luna are both listed, and the bare luna shorthand on this site still means GPT-5.6 Luna.

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

Verified against the live public allowlist on 2026-08-19. Base URL is https://continuumcode.ai/v1. Keys are cont_sk_ from Settings, Account, Inference API. Personal keys need Plus or above.

curl https://continuumcode.ai/v1/chat/completions \
  -H "Authorization: Bearer $CONTINUUM_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gpt-6-luna","messages":[{"role":"user","content":"Review this diff."}]}'

Get Plus to use the Continuum-hosted GPT-6 Luna rail. Not a Mac download.