AIToday
AI Business & IndustryTechCrunch AIMay 8, 2026

OpenAI launches new voice intelligence features in its API, including GPT-Realtime-2 with reasoning capabilities, real-time translation across more than 70 input languages, and live transcription.

OpenAI launches new voice intelligence features in its API, including GPT-Realtime-2 with reasoning capabilities, real-time translation across more than 70 input languages, and live transcription.

Summaries like this, in your inbox every morning.

Sign up free →

3 Key Points

  1. OpenAI released three new voice models through its Realtime API: GPT-Realtime-2 (a voice model built with GPT-5-class reasoning for handling complex requests), GPT-Realtime-Translate (real-time translation supporting more than 70 input languages and 13 output languages), and GPT-Realtime-Whisper (live speech-to-text capabilities).

  2. GPT-Realtime-2 differs from its predecessor GPT-Realtime-1.5 by incorporating GPT-5-class reasoning. Translate and Whisper are billed by the minute, while GPT-Realtime-2 is billed by token consumption.

  3. OpenAI built guardrails into the system so that conversations can be halted if detected as violating its harmful content guidelines, designed to prevent misuse for spam, fraud, or online abuse. The company notes the features are intended for customer service, education, media, events, and creator platforms.

Get the latest AI Business & Industry news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Discussion

No comments yet. Be the first to share your thoughts!

Log in to join the discussion

Related Articles

Stay ahead with AI news

Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.

Get Started Free

Free · takes 30 seconds · unsubscribe anytime