OpenAI/gpt-6-luna/Released Sep 2026

GPT 6 Luna

Low-cost high-volume reasoning, 1M context

Commercial use
Text

About GPT 6 Luna

GPT 6 Luna is OpenAI's most efficient GPT-6 generation model, built for focused, high-volume tasks where cost and speed matter more than maximum reasoning depth.

It accepts text and images and supports configurable reasoning effort (none, low, medium, high, xhigh, and max), streaming, structured outputs, and function calling, with the web search, file search, image generation, code interpreter, hosted shell, apply patch, skills, computer use, MCP, and tool search tools. Built-in tools and function calling belong on the Responses API, and Chat Completions supports function calling only with reasoning effort set to none. Fine-tuning is not available.

MetricValue
Context Length1,050,000 tokens
Max Output Tokens128,000 tokens
Knowledge CutoffMay 18, 2026
MultilingualYes

Ready to build with GPT 6 Luna?

Try GPT 6 Luna in the Workbench to prompt it, compare outputs, and iterate on prompts without writing any code. When you're ready to ship, call the same model from our API and build your own apps on top of it.

curl -sSf -X POST https://hub.oxen.ai/api/ai/chat/completions \
    -H "Content-Type: application/json" \
    -H "Authorization: Bearer $OXEN_API_KEY" \
    -d '{
  "model": "gpt-6-luna",
  "messages": [
    {
      "role": "user",
      "content": "Try sending a message."
    }
  ]
}'

API endpoint

See the API reference for request and response formats.

POSThttps://hub.oxen.ai/api/ai/chat/completions