2026年6月28日

W

A general-purpose speech recognition model by OpenAI.
AI WritingContent GenerationResearchEmail WritingSummarizationRewritingAcademic Research瀏覽器擴充功能

總覽
W 是什麼?

Whisper is a general-purpose speech recognition model developed by OpenAI. It is trained on a large dataset of diverse audio and is also a multi-task model that can perform multilingual speech recognition as well as speech translation and language identification. Whisper uses a Transformer sequence-to-sequence model trained on various speech processing tasks, including multilingual speech recognition, speech translation, spoken language identification, and voice activity detection. These tasks are jointly represented as a sequence of tokens to be predicted by the decoder, allowing a single model to replace many stages of a traditional speech-processing pipeline. The multitask training format uses a set of special tokens that serve as task specifiers or classification targets.

A general-purpose speech recognition model by OpenAI.

核心功能
Multilingual speech recognition
Speech translation
Language identification
Voice activity detection
熱門使用情境
  • Transcribing audio files to text
  • Translating speech from one language to another
  • Identifying the language spoken in an audio file
如何使用
  • Whisper can be used via command-line or within Python. For command-line usage
  • you can transcribe speech in audio files by specifying the audio file and model size. For Python usage
  • you can load the model and use the transcribe() method to process audio files.
產品時間線
待核實
定價
W 採用 Free 定價模式,價格與功能可能隨時間調整。
Free
$0
待核實
Pro
待核實
待核實
Team
待核實
待核實
Enterprise
待核實
待核實
優惠 / 優惠碼
目前沒有優惠碼。
驗證資訊
工具狀態
待核實
定價已核驗
待核實
創辦人已認領
否 / 待核實
來源
官網 / 社群提交
相關標籤
AI WritingContent GenerationResearchEmail WritingSummarizationRewritingAcademic Research瀏覽器擴充功能Freemium
你是這個工具的官方團隊嗎?
認領這個資料頁後,你可以更新產品資訊、定價與官方回覆。