OpenAI has expanded its API with three new voice intelligence models for real time voice applications. The release includes GPT-Realtime-2, a conversational model designed for realistic spoken interactions with GPT-5 class reasoning, allowing it to handle more complex live conversations than earlier versions. The company also introduced GPT-Realtime-Translate for real time spoken translation, with support for more than 70 input languages and 13 output languages. Following these, GPT-Realtime-Whisper adds live speech to text transcription, letting applications capture spoken interactions as they happen. GPT-Realtime-2 is billed by token usage, while translation and transcription are billed by the minute. All three models are available through OpenAI’s Realtime API and can support use cases such as customer service, education, media, events, and creator platforms. OpenAI also noted potential risks around spam, fraud, and online abuse, adding safety triggers that can halt conversations if...
Related
Microsoft introduces its first cybersecurity AI model at half the cost of competing models
Microsoft has introduced MAI-Cyber-1-Flash, its first AI model built specifically for cybersecurity. Designed to identify complex vulnerabilities in large codebases, the model work...
FireDragon 13 revamps web browser core, adds XDG Base Directories, and new welcome dialog
Garuda Linux has released FireDragon 13, the latest version of its customized web browser derived from Floorp. This update marks a complete re-implementation, delivering a new code...
YouTube introduces custom thumbnails for Shorts and Ask Studio AI thumbnail generation
YouTube is launching several updates to simplify the thumbnail creation process. Partner Program creators can now upload custom thumbnails for Shorts, addressing a highly requested...