OpenAIの新音声認識モデル、前世代より精度は改善も誤り率3.31%でElevenLabs・Google・Mistralに届かず
DRANK

7月29日、The Decoderが「GPT Transcribe improves on its predecessor but can't catch ElevenLabs, Google, or Mistral on error rates」と題した記事を公開した。OpenAIが新たにリリースした音声認識モデル「GPT Transcribe」と「GPT Live Transcribe」の性能を、競合他社のモデルと比較したベンチマーク結果を伝えている。

by @tf_official
Related Topics: AI Machine Learning AI Music/Voice Generator