GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
Specifications
| Provider | OpenAI |
|---|---|
| Context window | 4,095 tokens (4K) |
| Input price | $1 / 1M tokens |
| Output price | $2 / 1M tokens |
| Input modalities | text |
| Output modalities | text |
| Model ID | openai/gpt-3.5-turbo-0613 |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does GPT-3.5 Turbo (older v0613) cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $1 per 1M input tokens and $2 per 1M output tokens. A provider's own list price may differ. On Vincony, GPT-3.5 Turbo (older v0613) is available on credit-based pricing alongside 800+ other models.
What is GPT-3.5 Turbo (older v0613)'s context window?
GPT-3.5 Turbo (older v0613) supports a context window of up to 4,095 tokens (4K), which is how much text it can consider at once.
Who makes GPT-3.5 Turbo (older v0613)?
GPT-3.5 Turbo (older v0613) is built by OpenAI. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use GPT-3.5 Turbo (older v0613)?
You can use GPT-3.5 Turbo (older v0613) directly through OpenAI, or via Vincony, which aggregates 800+ models (including GPT-3.5 Turbo (older v0613)) into one interface with side-by-side comparison and credit-based pricing.