---
source_url: "https://developers.openai.com/api/docs/models/gpt-realtime-translate"
title: "GPT-Realtime-Translate Model | OpenAI API"
mirrored_at: 2026-08-16T03:03:14.629Z
host: developers.openai.com
cited_in_42a: true
mirror_canonical: "https://index.42a.ai/developers.openai.com/api/docs/models/gpt-realtime-translate"
---

> **Original source:** https://developers.openai.com/api/docs/models/gpt-realtime-translate

[Models](https://developers.openai.com/api/docs/models)

![gpt-realtime-translate](https://developers.openai.com/images/api/models/icons/gpt-realtime-translate.png)

GPT-Realtime-Translate

Default

Streaming speech-to-speech translation model

Streaming speech-to-speech translation model

Performance

Highest

Speed

Very fast

Price

$0.034

Price

Input

Audio

Output

Audio, text

GPT-Realtime-Translate is a streaming speech-to-speech translation model for live multilingual audio experiences. It uses a dedicated realtime translation endpoint and returns translated audio plus transcript deltas while source audio is still arriving. GPT-Realtime-Translate is priced by audio duration rather than text tokens.

16,000

context window

2,000

max output tokens

Sep 30, 2024 knowledge cutoff

Pricing

Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the

[pricing page](https://developers.openai.com/api/docs/pricing).

Realtime audio duration

Per

minute

Price

$0.034

GPT-Realtime-Translate is priced by audio duration rather than text tokens.

Modalities

Text

Output only

Image

Not supported

Audio

Input and output

Video

Not supported

Endpoints

Chat Completions

v1/chat/completions

Responses

v1/responses

Realtime

v1/realtime

Realtime translation

v1/realtime/translations

Realtime transcription

v1/realtime/transcription\_sessions

Assistants

v1/assistants

Batch

v1/batch

Fine-tuning

v1/fine-tuning

Embeddings

v1/embeddings

Image generation

v1/images/generations

Videos

v1/videos

Image edit

v1/images/edits

Speech generation

v1/audio/speech

Transcription

v1/audio/transcriptions

Translation

v1/audio/translations

Moderation

v1/moderations

Completions (legacy)

v1/completions

Features

Streaming

Supported

Function calling

Not supported

Structured outputs

Not supported

Fine-tuning

Not supported

Predicted outputs

Not supported

Snapshots

Snapshots let you lock in a specific version of the model so that performance and behavior remain consistent. Below is a list of all available snapshots and aliases for

GPT-Realtime-Translate

.

![gpt-realtime-translate](https://developers.openai.com/images/api/models/icons/gpt-realtime-translate.png)

gpt-realtime-translate

gpt-realtime-translate

gpt-realtime-translate

Rate limits

Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.

Tier

Minutes-of-audio per minute

Free

Not supported

Tier 1

50

Tier 2

200

Tier 3

400

Tier 4

650

Tier 5

850