OpenAI: GPT-5.6 Luna
Chat-80%openai/gpt-5.6-lunaGPT-5.6 Luna is the fast, cost-efficient tier of OpenAI's GPT-5.6 series, released on 2026-07-09. It is built for high-volume, latency-sensitive workloads such as chat, classification, and lightweight agentic workflows, providing capable reasoning for its price tier. It supports reasoning, vision, tool use (function calling), prompt caching, and web search. Context window: 1M tokens, output: 128K. Pricing is $1/M input tokens and $6/M output tokens, with cached input reads billed at $0.10/M. Available via OpenAI and Anthropic protocols through Ofox.
Context Window
1M
Max Output Tokens
128K
Released
2026-07-09
Capabilities
VisionFunction CallingReasoningPrompt CachingWeb Search
Available Providers
Azure
Supported Protocols
openaianthropic
Providers
Azure
-80%Input Tokens
$0.2/M
$1/M
Output Tokens
$1.2/M
$6/M
Cache Read
$0.02/M
$0.1/M
Cache Write
$0.25/M
$1.25/M
Output Image
$32/M
Web Search
$0.035/R
Protocols
openai
/v1/chat/completions/v1/responsesanthropic
Code Examples
from openai import OpenAIclient = OpenAI(base_url="https://api.ofox.run/v1",api_key="YOUR_OFOX_API_KEY",)response = client.chat.completions.create(model="openai/gpt-5.6-luna",messages=[{"role": "user", "content": "Hello!"}],)print(response.choices[0].message.content)
Uptime & Status
Related Models
Frequently Asked Questions
OpenAI: GPT-5.6 Luna on Ofox.ai costs $1/M per million input tokens and $6/M per million output tokens. Pay-as-you-go, no monthly fees.
OpenAI: GPT-5.6 Luna supports a context window of 1M tokens with max output of 128K tokens, allowing you to process large documents and maintain long conversations.
Simply set your base URL to https://api.ofox.run/v1 and use your Ofox API key. The API is OpenAI-compatible — just change the base URL and API key in your existing code.
OpenAI: GPT-5.6 Luna supports the following capabilities: Vision, Function Calling, Reasoning, Prompt Caching, Web Search. Access all features through the Ofox.ai unified API.