Models
Enterprise
Subscribe
Resource
Documentation
Console

openai/gpt-3.5-turbo

OpenAI

OpenAI: gpt-3.5-turbo

GPT-3.5 Turbo is OpenAI's long-running chat model, still used for lightweight text generation and legacy integrations.

Modetext → text
Input price$0.5 per M tokens
Output price
Context length
Weekly usage0 tokens
Listed at

OpenAI

Latency
Throughput
Upload rate
Context length
Max output
Input price$0.5per 1M tokens
Output priceper 1M tokens
Cache read$0per 1M tokens
Cache write$0per 1M tokens

Performance

Avg TPS
Avg latency
Avg success rate
TPS
TTFT
Latency
Success rate

Success rate

Speed

Code Example

Replace <YOUR_API_KEY> with the API key generated on your token management page.

Authentication

All requests must include the Authorization: Bearer <TOKEN> header.

Endpoint Path

POST

Supported Parameters

ParameterTypeDefaultDescription

Rate Limits

GroupRPMTPMRPD

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply by token group.

Pricing

GroupBilling typePrice summary