# OpenAI 发布 GPT Transcribe 非流式语音转录模型

- 来源：Artificial Analysis (@ArtificialAnlys)
- 发布时间：2026-07-29 10:01
- AIHOT 分数：69
- AIHOT 链接：https://aihot.virxact.com/items/cms5gt1yq007urolw4sc0v3mz
- 原文链接：https://x.com/ArtificialAnlys/status/2082285338509418727

## AI 摘要

OpenAI 推出非流式语音转录模型 GPT Transcribe，在 AA-WER 基准上词错误率为 3.31%（排名第 9），较前代提升 0.7 个百分点。该模型支持文本提示、关键词及多语言提示三种上下文，处理速度约达实时 34 倍，定价降至每千分钟音频 $4.50。

## 正文

OpenAI has released GPT Transcribe： a Speech to Text model scoring 3.31% on AA-WER （#9）， improving 0.7 p.p. over its predecessor GPT-4o Transcribe while lowering price 25% to $4.50 per 1，000 minutes of audio

GPT Transcribe is OpenAI's latest non-streaming （batch） speech transcription model， now accepting three kinds of context to improve transcription quality： a text prompt describing the recording's topic or setting， keywords for literal terms that may appear in the audio （such as product names or acronyms）， and multiple language hints for multilingual and code-switching audio. The model processes audio at ~34× real-time and is available at $4.50 per 1，000 minutes of audio （$0.0045/min） via the OpenAI API Platform.

OpenAI has also released GPT-Live-Transcribe， a streaming Speech to Text model. We are currently benchmarking this model and plan to share results on our Streaming Speech to Text leaderboard.

See more details below ⬇️
