Observed Signal · Aug 29, 2025 · Technical Release · Source: OnlineMarketing.de · Impact: 5/5 · Sentiment: Positive

OpenAI Unveils gpt-realtime Real-Time API

Executive Signal Summary

OpenAI introduced gpt-realtime, a real-time speech/text AI release featuring a generally available Real-Time API and the gpt-realtime model. The update adds Remote MCP Server connectivity, Image Input for describing visuals, SIP-Telephony integration for linking phone systems, and Reusable Prompts to store standardized conversations. Early tests by Zillow, T-Mobile, StubHub, and Oscar Health show real-world exploration beyond the lab. OpenAI also trims pricing by about 20% versus the previous gpt-4o-realtime-preview, with audio input tokens at $32 per million and audio output at $64 per million tokens. The platform offers finer control over conversation context, including per-session token limits to shorten long sessions. Data privacy is emphasized with EU Data Residency, active content filters, mandatory KI-dialog labeling, and an Agents SDK for additional safeguards.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

Major platform (OpenAI) technical release with real-time API, pricing changes, and EU data residency implications.

SIGNAL RADAR

Track TIME Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • OpenAI released the generally available Realtime API and new gpt-realtime model.
  • New features include Remote MCP Server, Image Input, SIP-Telephony, and Reusable Prompts.
  • Zillow, T-Mobile, StubHub and Oscar Health are testing gpt-realtime.
  • Pricing: ~20% reduction vs gpt-4o-realtime-preview; audio input tokens $32 per million; audio output $64 per million.
  • EU Data Residency is supported; multi-layer security includes content filters and an Agents SDK.
Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: OnlineMarketing.de•Published: Aug 29, 2025
Original Coverage Title: “OpenAI startet gpt-realtime: So menschlich klang KI noch nie | OnlineMarketing.de”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

Conversational AI & ChatbotsMay 7, 2026

OpenAI launches three Realtime voice models

OpenAI announced the addition of three realtime voice-intelligence models to its Realtime API on May 7, 2026: GPT‑Realtime‑2, a GPT‑5‑class reasoning voice model for realistic conversational and agentic workflows; GPT‑Realtime‑Translate, which provides live translation with support for more than 70 input languages and 13 output languages; and GPT‑Realtime‑Whisper, a low-latency streaming speech-to-text transcription capability. The features are intended for customer service, education, media, events and creator platforms. Translate and Whisper are billed by the minute while GPT‑Realtime‑2 is billed by token consumption. OpenAI said it has embedded safety guardrails and active classifiers to halt conversations that violate harmful-content policies. The announcement follows OpenAI’s published evaluations showing gains versus prior realtime models and details pricing and safety controls in the Realtime API documentation.

Read assessment
Conversational AI & Voice ModelsMay 8, 2026

OpenAI Releases GPT‑Realtime‑2, Translate, Whisper

OpenAI introduced a Chrome extension for Codex that lets the coding agent run in the browser background across tabs (extension currently in the Codex app but not yet available in the UK/EU) and reported that Codex sees over four million weekly users. Separately, OpenAI published three Realtime API voice models — GPT‑Realtime‑2, GPT‑Realtime‑Translate and GPT‑Realtime‑Whisper — bringing GPT‑5‑class reasoning and low‑latency streaming to voice agents. GPT‑Realtime‑2 supports interruption recovery and tool use; Translate covers 70+ input to 13 output languages; Whisper delivers streaming transcription. Pricing published: GPT‑Realtime‑2 — $32 per 1M audio‑input tokens and $64 per 1M audio‑output tokens; GPT‑Realtime‑Translate ~$0.00034/min; GPT‑Realtime‑Whisper ~$0.00017/min. OpenAI says ChatGPT will receive related voice updates in future releases.

Read assessment
InfrastructureAug 3, 2026

OpenAI launches GPT‑Live real‑time voice system

OpenAI describes GPT‑Live, a full‑duplex realtime voice system that streams audio into a voice model capable of listening and speaking simultaneously, removing prior turn‑detector architectures. The system uses stateful, streaming inference and an asynchronous delegation path to consult larger frontier models (e.g., GPT‑5.5) without blocking the live media loop. Engineering changes include rewriting the media frontend and inference logic in Go, using WebRTC for transport, new handshake/transport optimizations (WARP and Instant Connect), and production shadow testing of real ChatGPT Voice sessions. The architecture separates the live media path from application logic, supports seamless model instance handoffs and context compaction, improves observability and rollout controls, and will underpin a forthcoming GPT‑Live API and expanded ChatGPT Voice capabilities.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.