OpenAI released three new realtime voice models designed to enable developers to build a new class of voice applications. The models include GPT-Realtime-2 with GPT-5-class reasoning, GPT-Realtime-Whisper for transcription, and GPT-Realtime-Translate for multilingual support.
GPT-Realtime-2 brings advanced reasoning capabilities to voice interactions, matching GPT-5 performance in real-time conversations. The model enables developers to build applications that can understand and respond to complex spoken queries without the traditional latency associated with voice processing.
GPT-Realtime-Whisper handles live speech transcription, converting audio input to text in real time. This component allows for immediate processing of spoken input across applications.
GPT-Realtime-Translate supports translation across 70+ languages, enabling developers to build multilingual voice applications. The feature allows real-time conversation across language barriers without separate processing steps.
OpenAI stated the models will "unlock a new class of voice apps for developers." The release targets the API, making the technology available to third-party developers building commercial and internal applications.
In parallel, OpenAI launched Trusted Contact, an optional safety feature for ChatGPT. The feature allows users over 18 to assign an emergency contact for mental health and safety concerns. The capability expands existing teenage safety options to adult users, enabling designated contacts to be notified when the platform detects potential safety issues.
The Trusted Contact feature is optional and gives users control over their safety settings. When activated, designated contacts can receive alerts if OpenAI's systems detect concerning behavior patterns.
Both announcements reflect OpenAI's focus on expanding voice capabilities while strengthening user safety features. The voice models target developers seeking to integrate advanced audio processing into applications, while the Trusted Contact feature addresses growing concerns about digital safety and mental health support.
StemDeck is a new open-source AI stem separator that runs locally on your machine without cloud dependencies. The free tool splits audio into individual instrument tracks.
A firsthand look at China's AI development reveals both countries pursuing remarkably similar technological paths, despite geopolitical tensions. The competition appears less ideological and more focused on matching capabilities.
Anthropic will permanently increase Claude Code's weekly limits by 25% starting September 14, but users will see a 17% net reduction once a current 50% temporary boost expires.
A developer who scraped artwork for AI training is now collaborating with Cara, a creator platform designed to prevent unauthorized AI data collection, as the service faces ongoing attacks from trolls attempting to breach and publish its data.