OpenAI

GPT-3.5 Turbo (older v0613)

Generally available4K context

Summary

GPT-3.5 Turbo (older v0613) is a language model from OpenAI, released on 25 Jan 2024. It accepts text and generates text, with a 4K-token context window and up to 4K output tokens per response. As of 4 Oct 2026, the lowest listed API price is $1 per million input tokens and $2 per million output tokens (OpenRouter), across 1 provider we track.

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

Last verified 4 Oct 2026. Every figure on this page links to its source.

Specifications

DeveloperOpenAI
Release date25 Jan 2024 [OpenRouter]
StatusGenerally available
Context window4K (4,095 tokens) [OpenRouter]
Max output4K tokens [OpenRouter]
Knowledge cutoff30 Sept 2021 [OpenRouter]
InputText [OpenRouter]
OutputText [OpenRouter]
Reasoning modeNo [OpenRouter]
Tool callingYes [OpenRouter]
Structured outputYes [OpenRouter]

GPT-3.5 Turbo (older v0613) API pricing by provider

USD per million tokens. Sorted by input price.

ProviderInputOutputCached inputContextSource
OpenRouter
openai/gpt-3.5-turbo-0613
$1$2–4KOpenRouter, 4 Oct 2026

Frequently asked questions

What is the context window of GPT-3.5 Turbo (older v0613)?

GPT-3.5 Turbo (older v0613) has a context window of 4,095 tokens (4K) and can generate up to 3,685 tokens in a single response, according to OpenRouter.

How much does the GPT-3.5 Turbo (older v0613) API cost?

GPT-3.5 Turbo (older v0613) is listed from $1 per million input tokens and $2 per million output tokens (OpenRouter). Prices last verified 4 Oct 2026.

When was GPT-3.5 Turbo (older v0613) released?

OpenAI released GPT-3.5 Turbo (older v0613) on 25 Jan 2024, according to OpenRouter.

Change history

No changes detected since we started tracking this model. We re-check every source daily.

Sources