About GPT 6 Luna
GPT 6 Luna is OpenAI's most efficient GPT-6 generation model, built for focused, high-volume tasks where cost and speed matter more than maximum reasoning depth.
It accepts text and images and supports configurable reasoning effort (none, low, medium, high, xhigh, and max), streaming, structured outputs, and function calling, with the web search, file search, image generation, code interpreter, hosted shell, apply patch, skills, computer use, MCP, and tool search tools. Built-in tools and function calling belong on the Responses API, and Chat Completions supports function calling only with reasoning effort set to none. Fine-tuning is not available.
| Metric | Value |
|---|---|
| Context Length | 1,050,000 tokens |
| Max Output Tokens | 128,000 tokens |
| Knowledge Cutoff | May 18, 2026 |
| Multilingual | Yes |
Ready to build with GPT 6 Luna?
Try GPT 6 Luna in the Workbench to prompt it, compare outputs, and iterate on prompts without writing any code. When you're ready to ship, call the same model from our API and build your own apps on top of it.
curl -sSf -X POST https://hub.oxen.ai/api/ai/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $OXEN_API_KEY" \
-d '{
"model": "gpt-6-luna",
"messages": [
{
"role": "user",
"content": "Try sending a message."
}
]
}'API endpoint
See the API reference for request and response formats.
https://hub.oxen.ai/api/ai/chat/completions