header-langage
简体中文
繁體中文
English
Tiếng Việt
한국어
日本語
ภาษาไทย
Türkçe
Scan to Download the APP

OpenAI has released two new speech-to-text models, with prices reduced by 25%.

According to TrendForce Monitoring, OpenAI has released two speech-to-text models, GPT Transcribe and GPT Live Transcribe. The former handles audio files and batch tasks, while the latter is used for live captions in scenarios like live broadcasts and phone calls.

Both models can incorporate audio topics, keywords, and language cues, focusing on improving recognition accuracy for short sentences, numbers, technical terms, multiple accents, and noisy environments.

In a study by Artificial Analysis, GPT Transcribe has a word error rate of 3.31%, which is 0.7 percentage points lower than the previous generation GPT-4o Transcribe. The price has also been reduced by 25%, charging $4.5 per 1000 minutes of audio.

举报 Correction/Report
Correction/Report
Submit
Add Library
Visible to myself only
Public
Save
Choose Library
Add Library
Cancel
Finish