578 related articles

OpenAI's upgraded voice assistant can speak dialects, do real-time simultaneous interpretation, teach English, and even get flustered. Here's what changed.

OpenAI's GPT-Live voice model family brings full-duplex interaction, task delegation, GPT-5-level intelligence, real-time translation, and image understanding to voice AI.

OpenAI's new voice model delivers near-zero-latency bidirectional conversation with real-time multilingual simultaneous interpretation across Cantonese, Spanish, and English.

OpenAI's GPT Live brings full-duplex voice AI with simultaneous listening and speaking, real-time interruption, dual-model delegation, and semantic-level live translation. A deep dive into the technology.

OpenAI's new voice model GPT-Live-1 focuses on fewer interruptions, recognizing pauses, and respecting conversational rhythm. A deep dive into its technical advances.

OpenAI unveils GPT-Live, a full-duplex voice model with real-time interruption, tiered compute routing, and dynamic UI rendering—surpassing Siri and targeting the OS-level voice gateway.

OpenAI unveils the GPT-Live voice model family, with full-duplex interaction enabling AI to listen and speak simultaneously and delegate complex reasoning to GPT-5.5. GPQA benchmark jumps from 45% to 80%.

OpenAI launches GPT Live, a voice AI model family supporting full-duplex real-time conversation, deep task delegation, multimodal interaction, and instant translation, with reasoning near GPT-5 level.

OpenAI launches GPT-5.6, ChatGPT Work, upgraded Codex super-app, and GPT Live voice AI — a deep dive into all four products and their impact on the AI landscape.

Hands-on test of OpenAI's new voice model: real-time interruption, simultaneous translation, emotion switching, code review, and comparison with Doubao.

OpenAI launches GPT-Live, a full-duplex voice AI supporting continuous interaction and intelligent task delegation for real-time translation, language correction, and parallel search.
Tech FrontiersOpenAI hosted Voice Hack Night where teams built 4 real-time voice agent projects in 6 hours. Deep analysis of technical challenges, use cases, and developer ecosystem trends in real-time voice AI.
Tech FrontiersOpenAI hosts a Realtime Voice Demo event on May 27 in San Francisco for developers. Learn about judging criteria, rewards, and what this means for voice AI ecosystems.
Tutorialsruby-openai is an open-source library with 3,200+ GitHub stars, supporting GPT-5 and WebRTC real-time voice. Learn how to integrate OpenAI API into Ruby on Rails for AI-powered apps.

OpenAI CEO Sam Altman demos unreleased Astra model to Washington policymakers, revealing proactive regulatory engagement trends and their implications for AI governance.

A complete guide to building a local private AI assistant with Ollama and Qwen-Agent. Covers RAG knowledge integration, voice interaction, and permission isolation for a secure local AI Agent architecture.

New EU regulations require mandatory labeling of realistic AI-generated content, covering deepfake videos, AI images, and voice clones. Analysis of the rules, challenges, and industry impact.

New EU rules mandate labeling for realistic AI-generated content including deepfakes, AI images, and voice clones. Analysis of enforcement challenges and industry impact.

Deep dive into MiniMax H3 multimodal model: 2K video generation, native stereo audio-visual integration, and precise text rendering designed for motion design and brand marketing.

In-depth review of Laxis AI meeting tool: bot-free recording, 100+ language real-time translation, voice dictation 4x faster than typing. Features, competitors & value analysis.