← 模型库

openai/openai/whisper-large-v3-turbo

持续维护
发布日期 未标注·最后更新 08-09·数据来源 社区补充·待核实
这是你发布的模型吗?认领后可标注官方身份并长期维护主页信息。
入选场景榜🎧 语音 30.8
任务类型多模态
参数规模暂无可靠数据
上下文长度暂无可靠数据
开源协议暂无可靠数据
商用情况暂无可靠数据
推荐显存暂无可靠数据
主要语言暂无可靠数据
本地部署未标注

模型介绍

中文速览

发布方 openai

资料来源:DeepInfra 模型库,以下正文为官方原始说明。

本条目的官方说明为英文,以下内容保留原文,关键章节标题已中文标注。

openai/whisper-large-v3-turbo

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper “Robust Speech Recognition via Large-Scale Weak Supervision” by Alec Radford et al. from OpenAI. Trained on >5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting. Whisper large-v3-turbo is a finetuned version of a pruned Whisper large-v3. In other words, it’s the exact same model, except that the number of decoding layers have reduced from 32 to 4. As a result, the model is way faster, at the expense of a minor quality degradation.

  • 上下文长度:未知
  • 最大输出:未知
  • 标签:stt
  • 提供方:openai
  • 来源:DeepInfra

适合场景

暂无可靠数据。

已知限制

  • 本站未登记确切发布日期,版本时间线可能不完整。
  • 本站未登记上下文长度,长文本能力请以官方文档为准。
  • 本页评测与硬件数据为区间估算,未做统一环境实测,不能替代官方评测报告。

评测

暂无可靠数据。本站只收录标注了「评测来源、模型版本、是否官方数据、测试日期」的成绩, 不混合不同设置下的分数,也不把单一 Benchmark 解释为整体能力。