Whisper large-v3-turbo

OpenAI

Tikai katalogā Atvērti svari

Suverenitāte un mitināšana

Tikai katalogā

Šis modelis ir iekļauts katalogā uzziņai. Tas vēl nav mitināts LLM EU, un neviens pirmās puses galapunkts to šeit neapkalpo. Pieprasījumi tam tiek nosūtīti tieši nodrošinātājam saskaņā ar šī nodrošinātāja nosacījumiem.

Kopsavilkums

Whisper large-v3-turbo is a distilled version of Whisper large-v3 with 809M parameters, released under the MIT licence. It keeps the same 30-second receptive field and the multilingual language coverage of large-v3 but decodes faster, which is the reason to prefer it where latency matters. The model card is the shared Whisper card and publishes no benchmark table specific to this variant.

Specifikācijas

Nodrošinātājs
OpenAI (US)
Konteksta logs
Nav verificēts: nodrošinātājs nav dokumentējis šo apgalvojumu.
Parametri
0.81B
Modalitātes
audio
Licence
mit · licence
Atvērti svari
yes
Izlaiduma datums
2024. g. 01. okt.
Uzdevumi
translation, multilingual-support, cheap

Valodas

Spēcīga: en, de, fr, es, it, pt, nl, pl

Atbilstoša: sv, da, fi, cs, ro, el, uk, ru, tr, ar, zh, ja, ko, hi

Valodas

Spēcīga

  • en
  • de
  • fr
  • es
  • it
  • pt
  • nl
  • pl

Atbilstoša

  • sv
  • da
  • fi
  • cs
  • ro
  • el
  • uk
  • ru
  • tr
  • ar
  • zh
  • ja
  • ko
  • hi

Nav verificēts: nodrošinātājs nav dokumentējis šo apgalvojumu. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.

Kad to izmantot

Stiprās puses

  • roughly half the parameters of large-v3 at lower latency
  • MIT licence, more permissive than the Apache-2.0 parent
  • same multilingual transcription coverage as large-v3

Ierobežojumi

  • a distillation of large-v3, so transcription accuracy is below the parent model
  • no benchmark table specific to this variant is published on the model card
  • fixed 30-second receptive field, so longer audio needs chunking logic