BenchLeader

GPT Realtime 2

Reasoning effort

GPT Realtime 2 is an OpenAI proprietary model. It is not yet ranked: no independent benchmark results so far, only pricing and metadata. It has been measured at 2 reasoning-effort settings; this summary describes the best-scoring one, and the tabs above switch between them.

Blended price
Output speed
First answer
Context
Released
date not published by our sources

Versions

OpenAI has shipped 3 models under this name. Each is ranked on its own results; a newer version often has fewer results so far, which holds its index nearer the average until more arrive.

ModelReleasedIndexRank
GPT Realtime
GPT Realtime 1.5
GPT Realtime 2xhighthis page

Reasoning-effort configurations

The same model behaves differently depending on how much it is allowed to think. Each row is one setting, scored only on the benchmarks that were run at that setting. “Not stated” collects results from publishers that did not say which setting they used; for a reasoning model that is usually its thinking mode, but we do not assume it. Pick a setting here or at the top of the page to see its price, speed and category scores.

EffortIndexRankSpeedFirst answerChat reply costCategories
xhighbest
not stated

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Multimodal

Benchmarkxhighnot statedSource
AudioMCnot in index48.5%#437.6%#8Scale AI SEAL
AudioMC (audio output)not in index48.5%#137.6%#3Scale AI SEAL

See also

Data as of 19 Sept 2026. Compare these configurations.

Cite as: BenchLeader, “GPT Realtime 2: benchmarks, pricing, speed and rank”, https://www.benchleader.com/models/gpt-realtime-2, data as of 19 Sept 2026.