OpenAI Releases GPT-Realtime-2.1 and ...
Quick Answer
OpenAI has launched GPT-Realtime-2.1 and GPT-Realtime-2.1-mini, enhancing low-latency voice agents with integrated reasoning capabilities at the same cost as previous models.
Quick Take
The mini version offers improved performance with a 25% reduction in p95 latency and significantly lower audio output costs, making reasoning the default feature in voice applications.
Key Points
- GPT-Realtime-2.1-mini offers reasoning at the same price as the old model, $0.60/1M text.
- P95 latency reduced by at least 25% due to improved caching across voice models.
- Audio output for mini is $20/1M, compared to $64/1M for the full model.
- Reasoning effort can be adjusted from minimal to high, optimizing latency and depth.
- Full model enhances alphanumeric recognition and noise handling capabilities.
Source Excerpt
OpenAI Releases GPT-Realtime-2.1 and GPT-Realtime-2.1-mini for Low-Latency Voice Agents in the API
For years, the cheap voice model was the "fast but not smart" tier. You picked reasoning or you picked low cost — never both. OpenAI just deleted that trade-off.
They shipped two
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from WebSearch (Tavily)
See more →It's true that Anthropic is the fastest growing company ...
Anthropic is rapidly emerging as a leader in AI with its Claude model, which outperforms competitors in various benchmarks. Their coding agents are also recognized as the best in the industry, contributing to their unprecedented revenue growth and solidifying their position in the enterprise market.
WSJ: OpenAI is considering deep price reductions as competition ...
OpenAI is contemplating significant price cuts in response to competitive pressure from Anthropic, particularly due to the success of Claude Code in developer and coding workflows. This shift could affect pricing strategies in the AI market as companies vie for dominance in coding solutions.