x-ai
// openai
GPT-3.5 Turbo (older v0613)
openai/gpt-3.5-turbo-0613
GPT-3.5 Turbo (older v0613) costs $1.00 per 1M input tokens and $2.00 per 1M output tokens with a 4K-token context window.
A tool-calling model from openai, released Jan 25, 2024. LiveBench hasn't published a run for it, so this page carries measured price and specs only — no benchmark numbers.
- Input / 1M
- $1.00
- Output / 1M
- $2.00
- Context
- 4K
- LiveBench overall
- Not evaluated
Specs and pricing
| Input price USD per 1M prompt tokens | $1.00 |
|---|---|
| Output price USD per 1M completion tokens, reasoning included | $2.00 |
| Blended price 3:1 input:output — #120 most expensive of 343 listed models | $1.25 |
| Cached input read Blank means not priced separately — not free | — |
| Cache write | — |
| Context window | 4,095 tokens |
| Max output | 3,685 tokens |
| Input modalities | text |
| Tool calling | Yes |
| Extended reasoning | No |
| Knowledge cutoff | 2021-09-30 |
| Listed | Jan 25, 2024 |
What GPT-3.5 Turbo (older v0613) costs to run
Monthly list cost across five workload shapes, using the published cached-input rate where there is one. A floor, not a quote — batch discounts and cache writes aren't included.
| Workload | Per month |
|---|---|
| Support chatbot 1.2K in / 400 out × 200K requests | $400.00/mo |
| RAG assistant 8K in / 600 out × 100K requests | $920.00/mo |
| Coding agent 40K in / 4K out × 20K requests | $960.00/mo |
| Document extraction 20K in / 1.5K out × 50K requests | $1,150/mo |
| Bulk classification 500 in / 20 out × 5M requests | $2,700/mo |
Compare GPT-3.5 Turbo (older v0613) with
Benchmarked models at a similar price.
moonshotai
GPT-3.5 Turbo (older v0613) vs Kimi K2.7 Code
deepseek
GPT-3.5 Turbo (older v0613) vs DeepSeek V4 Pro 0423
qwen
GPT-3.5 Turbo (older v0613) vs Qwen3.8 27B
nvidia
GPT-3.5 Turbo (older v0613) vs Nemotron 3 Ultra
GPT-3.5 Turbo (older v0613) vs Gemini 3.8 Flash
GPT-3.5 Turbo (older v0613) FAQ
How much does GPT-3.5 Turbo (older v0613) cost?
GPT-3.5 Turbo (older v0613) lists at $1.00 per 1M input tokens and $2.00 per 1M output tokens. A RAG assistant handling 100K requests a month (8K tokens in, 600 out) comes to about $920.00 at list price.
What is the context window of GPT-3.5 Turbo (older v0613)?
GPT-3.5 Turbo (older v0613) accepts up to 4,095 tokens of context and can return up to 3,685 output tokens per response.
Does GPT-3.5 Turbo (older v0613) support tool calling?
Yes — GPT-3.5 Turbo (older v0613) supports tool (function) calling. It accepts text input.
How does GPT-3.5 Turbo (older v0613) score on benchmarks?
LiveBench has not published a run for GPT-3.5 Turbo (older v0613) yet. Until it does, compare it on price and specs, or run your own evaluation on your own tasks.
What is the API model ID for GPT-3.5 Turbo (older v0613)?
On OpenRouter the model ID is "openai/gpt-3.5-turbo-0613". It was listed on Jan 25, 2024.
More from openai
How these numbers are produced
- Price and specs — provider list data from OpenRouter, refreshed every 15 minutes. “Blended” is a 3:1 input:output mix.
- Scores and ranks — LiveBench release 2026-06-25. Ranks count only models that have a published run; a blank means “not evaluated”, never “bad”.
- Cost per point — the measured dollars LiveBench spent on the run, divided by the score it earned.
Published benchmarks rank models on someone else's tasks. Before committing, see LLM & agent evaluation for building an eval on your own.