20% off all GPT ๐ŸŽ‰ 30% off DeepSeekLearn more

OpenAI: GPT-6 Luna

Chat-20%
openai/gpt-6-luna

GPT-6 Luna is OpenAI's fastest and most cost-efficient GPT-6 model, built for high-volume, latency-sensitive tasks such as classification, structured extraction, and lightweight agents. It accepts text and image inputs and produces text, with adjustable reasoning effort. It supports vision, function calling, web search, and prompt caching. Context window: 1.05M tokens, output: 128K, allowing large inputs and extended responses within one request. Available through Ofox via OpenAI and Anthropic protocols, including OpenAI-compatible Chat Completions and Responses endpoints.

Context Window
1M
Max Output Tokens
128K
Released
2026-09-22
Capabilities
VisionFunction CallingReasoningPrompt CachingWeb Search
Available Providers
OpenAIAzure
Supported Protocols
openaianthropic

Providers

Azureprovider.type: "azure_foundry"
-20%
Input Tokens
$0.08/M
$0.1/M
Output Tokens
$0.4/M
$0.5/M
Cache Read
$0.008/M
$0.01/M
Cache Write
$0.1/M
$0.125/M
Web Search
$0.01/R
Protocols
openai/v1/chat/completions/v1/responses
anthropic
OpenAIprovider.type: "openai"
-20%
Input Tokens
$0.08/M
$0.1/M
Output Tokens
$0.4/M
$0.5/M
Cache Read
$0.008/M
$0.01/M
Cache Write
$0.1/M
$0.125/M
Web Search
$0.01/R
Protocols
openai/v1/chat/completions/v1/responses
anthropic

Code Examples

from openai import OpenAI
client = OpenAI(
base_url="https://api.ofox.run/v1",
api_key="YOUR_OFOX_API_KEY",
)
response = client.chat.completions.create(
model="openai/gpt-6-luna",
messages=[
{"role": "user", "content": "Hello!"}
],
)
print(response.choices[0].message.content)

Uptime & Status

Frequently Asked Questions

OpenAI: GPT-6 Luna on Ofox.ai costs $0.1/M per million input tokens and $0.5/M per million output tokens. Pay-as-you-go, no monthly fees.