OpenAI
Latency—
Throughput—
Upload rate—
Context length—
Max output—
Input price$0.5per 1M tokens
Output price—per 1M tokens
Cache read$0per 1M tokens
Cache write$0per 1M tokens
/openai/gpt-3.5-turbo
GPT-3.5 Turbo is OpenAI's long-running chat model, still used for lightweight text generation and legacy integrations.
Replace <YOUR_API_KEY> with the API key generated on your token management page.
All requests must include the Authorization: Bearer <TOKEN> header.
RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.