OpenAI 推出 GPT-Realtime-2.1 与 GPT-Realtime-2.1-mini,强化低延迟语音 Agent
Decision Brief
What changedOpenAI 在 API 中新增 GPT-Realtime-2.1 与 GPT-Realtime-2.1-mini 两款即时模型。
Why it mattersGPT-Realtime-2.1-mini 是专为语音设计的小型推理模型,定价与前代相同,且改进缓存后 p95 延迟降低至少 25%,对建置低延迟语音应用的开发者极具吸引力。
Who should careTeams building on model APIs
Affected stackOpenAI
Builder actionWorth evaluating: can OpenAI replace or complement your current model
Source confidenceMedium · Reliable media or first-hand reporting
OpenAI 在 API 中新增了两款即时模型:GPT-Realtime-2.1 与 GPT-Realtime-2.1-mini。其中 GPT-Realtime-2.1-mini 是一款专为语音设计的小型推理模型,定价与先前的 gpt-realtime-mini 相同。此外,OpenAI 透过改善缓存机制,将 p95 延迟降低了至少 25%。这些模型支援透过 WebRTC 连接。 对于使用 OpenAI API 开发即时语音代理的开发者而言,GPT-Realtime-2.1-mini 提供了更低的延迟和更高效的推理能力,特别适合需要即时回应的语音互动场景。同时,定价维持不变,有助于控制成本。
Summary basis: official / RSS sourceCompiled from the source scope noted above; the original remains authoritative.
Sources
- MarkTechPost
Fast research-paper and ML tooling summaries, useful for infra and agent updates.
- MarkTechPost
留言
登入后即可留言,和其他 builder 交换实测心得。
还没有留言,抢头香。