OpenAI
Latency—
Throughput—
Upload rate—
Context length—
Max output—
Input price—1M output tokens
Output price$1.21M output tokens
Cache read—1M output tokens
Cache write—1M output tokens
/openai/gpt-5.6-luna
OpenAI began the limited preview of GPT-5.6 Luna on June 26, 2026 and later included it in the generally available GPT-5.6 family. GPT-5.6 Luna is the lowest-cost model in the family and is used for workloads where lower inference cost and fast response are priorities.
Replace <YOUR_API_KEY> with the API key generated on your token management page.
All requests must include the Authorization: Bearer <TOKEN> header.
RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.