// openai

GPT-3.5 Turbo 16k

openai/gpt-3.5-turbo-16k

tool calling
OpenRouter + LiveBenchAll modelsFull leaderboard

GPT-3.5 Turbo 16k costs $3.00 per 1M input tokens and $4.00 per 1M output tokens with a 16K-token context window.

A tool-calling model from openai, released Aug 28, 2023. LiveBench hasn't published a run for it, so this page carries measured price and specs only — no benchmark numbers.

Input / 1M
$3.00
Output / 1M
$4.00
Context
16K
LiveBench overall
Not evaluated

Specs and pricing

Input price

USD per 1M prompt tokens

$3.00
Output price

USD per 1M completion tokens, reasoning included

$4.00
Blended price

3:1 input:output — #71 most expensive of 343 listed models

$3.25
Cached input read

Blank means not priced separately — not free

Cache write
Context window16,385 tokens
Max output4,096 tokens
Input modalitiestext
Tool callingYes
Extended reasoningNo
Knowledge cutoff2021-09-30
ListedAug 28, 2023
All OpenAI API pricing

What GPT-3.5 Turbo 16k costs to run

Monthly list cost across five workload shapes, using the published cached-input rate where there is one. A floor, not a quote — batch discounts and cache writes aren't included.

WorkloadPer month
Support chatbot

1.2K in / 400 out × 200K requests

$1,040/mo
RAG assistant

8K in / 600 out × 100K requests

$2,640/mo
Coding agent

40K in / 4K out × 20K requests

$2,720/mo
Document extraction

20K in / 1.5K out × 50K requests

$3,300/mo
Bulk classification

500 in / 20 out × 5M requests

$7,900/mo
Price your own workload

Compare GPT-3.5 Turbo 16k with

Benchmarked models at a similar price.

GPT-3.5 Turbo 16k FAQ

How much does GPT-3.5 Turbo 16k cost?

GPT-3.5 Turbo 16k lists at $3.00 per 1M input tokens and $4.00 per 1M output tokens. A RAG assistant handling 100K requests a month (8K tokens in, 600 out) comes to about $2,640 at list price.

What is the context window of GPT-3.5 Turbo 16k?

GPT-3.5 Turbo 16k accepts up to 16,385 tokens of context and can return up to 4,096 output tokens per response.

Does GPT-3.5 Turbo 16k support tool calling?

Yes — GPT-3.5 Turbo 16k supports tool (function) calling. It accepts text input.

How does GPT-3.5 Turbo 16k score on benchmarks?

LiveBench has not published a run for GPT-3.5 Turbo 16k yet. Until it does, compare it on price and specs, or run your own evaluation on your own tasks.

What is the API model ID for GPT-3.5 Turbo 16k?

On OpenRouter the model ID is "openai/gpt-3.5-turbo-16k". It was listed on Aug 28, 2023.

More from openai

How these numbers are produced

  • Price and specs — provider list data from OpenRouter, refreshed every 15 minutes. “Blended” is a 3:1 input:output mix.
  • Scores and ranksLiveBench release 2026-06-25. Ranks count only models that have a published run; a blank means “not evaluated”, never “bad”.
  • Cost per point — the measured dollars LiveBench spent on the run, divided by the score it earned.

Published benchmarks rank models on someone else's tasks. Before committing, see LLM & agent evaluation for building an eval on your own.