whisper
reliable Speech Recognition via Large-Scale Weak Supervision. The multitask training format uses a set of special tokens that serve as task specifiers or classification targets.
What it solves
reliable Speech Recognition via Large-Scale Weak Supervision. The multitask training format uses a set of special tokens that serve as task specifiers or classification targets.
The project context on this page is free to read. The original GitHub repository remains the source of truth; sign in only when you want to save or join the discussion.
Community notes