The Circuitry
THE CIRCUITRYYour one-stop source for all tech news
HOMETODAYNEWSFEEDEVENTS
BOOKMARKS
RSS
© 2026 The Circuitry
About UsSourcesContactCorrectionsPrivacy
  • Today
  • Feed
  • Events
  • Saved
Scroll for more
Verification
VERIFIEDConfidence: HIGH
Source identified
Claims cross-referenced
No discrepancies found
Sourcing
1source

via 9to5Mac

9to5Mac · track record
68Stories
100%Verified
530d
All sources →
Home/Tech/OpenAI Releases Three New Realtime Voice Models for Developers
VERIFIEDBy Xavier Rivera· ·1 min read

OpenAI Releases Three New Realtime Voice Models for Developers

OpenAI has launched three realtime voice models focused on reasoning, translation, and transcription. The models are now accessible through the company's Realtime API with specific per-token and per-minute pricing.

Source:9to5Mac
Post
OpenAI Releases Three New Realtime Voice Models for Developers
TL;DRAI · 60 sec read

OpenAI releases three realtime voice models via its Realtime API: GPT-Realtime-2 with GPT-5-class reasoning for handling complex voice requests, interruptions, and tools; GPT-Realtime-Translate for live speech translation from 70+ input languages to 13 outputs; and GPT-Realtime-Whisper for low-latency streaming transcription. These enable new voice applications for developers.

OpenAI has released three new realtime voice models designed to support a new class of voice applications for developers. Each model targets a distinct function in live audio interactions.

The first model, GPT-Realtime-2, introduces GPT-5-class reasoning capabilities to voice conversations. It manages complex requests, maintains natural dialogue flow, calls tools, processes corrections or interruptions, and generates contextually appropriate responses during live exchanges.

GPT-Realtime-Translate provides live speech translation across more than 70 input languages into 13 output languages. The model maintains pace with the speaker to deliver continuous translation during ongoing speech.

GPT-Realtime-Whisper offers low-latency streaming speech-to-text transcription. It converts spoken audio into text in real time, enabling immediate captions and meeting notes that update as conversation progresses.
From The CircuitryThe Feed — live briefs across tech, all day.See what’s happening →

All three models are available through OpenAI's Realtime API. Developers can access them via the Playground for testing or integrate GPT-Realtime-2 into existing applications using Codex.

Pricing for the models varies by type. GPT-Realtime-2 costs $32 per million audio input tokens, with cached input tokens at $0.40 per million, and $64 per million audio output tokens. GPT-Realtime-Translate is priced at $0.034 per minute, while GPT-Realtime-Whisper is priced at $0.017 per minute.
Additional details on the models and current developer usage are available through OpenAI resources.
Why this mattersAI · ~100 words

Tap a lens to see what this story means for you.

Reader-supported
DonateBuy me a coffee →Follow@thecircuitry_ →Follow@thecircuitry.to →

Reader-supported · The Brief

Liked this? The Brief brings you the whole day in tech, verified, every morning. Two minutes, free forever.

HELP US IMPROVE
From The Circuitry

See what’s happening right now

The Feed runs all day — short, verified briefs the moment they break.

Open the Feed →
From The Circuitry

Follow @thecircuitry_

Every story we publish, as it happens. No noise between.

Follow on X ↗On Bluesky ↗

Reader-supported

The Circuitry is a passion project I've always wanted to build, and I love the work behind it.

Running it costs real money. APIs, hosting, time. To keep improving the site and growing this into something useful for everyone, those costs have to be covered.

Any contribution is appreciated. If not, no pressure. Thanks for reading.

Buy me a coffee
OpenAIAIVoice
More from9to5Mac
  • Tim Cook Marks End of His Tenure Leading Apple Earnings Calls

    Tech · 8d
  • Apple tests AI for Genius Bar, sends letters to ex-staff at OpenAI

    Tech · 17d
  • Apple to launch Upgrade leasing program on July 28

    Tech · 17d
More inTech
  • Tesla and SpaceX confirm Terafab chip fab in Texas

    Tech · 1d
  • OpenAI Urges Federal Judge to Throw Out Apple's Trade Secrets Complaint

    Tech · 2d
  • Meta introduces Muse Code, its terminal-based coding agent

    Tech · 2d
SupportThe Work

The Circuitry is reader-supported. If you find the daily brief useful, you can buy me a coffee to keep it going.

Buy a coffee →
SubscribeCircuitry Brief

Liked this? The Brief brings you the whole day in tech, verified, every morning. Free forever.

MORE IN TECH

Tesla and SpaceX confirm Terafab chip fab in Texas

Tesla and SpaceX have confirmed Grimes County, Texas as the site for their Terafab semiconductor megafactory, with the first phase costing roughly $16.8 billion. The project targets the largest chip manufacturing facility on the planet to supply over 1 terawatt of compute per year that exceeds current and future global production capacity.

OpenAI Urges Federal Judge to Throw Out Apple's Trade Secrets Complaint

OpenAI has filed a motion asking a federal judge to dismiss Apple's trade secrets lawsuit, describing the claims as meritless. The dispute, which follows a July suit and this week's injunction request from Apple, highlights tensions after their prior partnership on Siri and OpenAI's hardware push.

Meta introduces Muse Code, its terminal-based coding agent

Meta has released an early beta of Muse Code, a terminal-based coding agent driven by the updated Muse Spark 1.2 model and positioned against Anthropic's Claude Code and OpenAI's Codex. Substantially lower rates, including a contributor plan at $0.10 for every million tokens received, may encourage migration away from higher-priced options such as Anthropic's Sonnet 5.