# #whisper — MDRSS hashtag feed

> Public MDRSS cards tagged #whisper.
> Canonical feed: https://mdrss.com/feeds/whisper

## Cards (3)

### [Whisper](https://mdrss.com/multimodal/speech-and-audio/2474/2474.md)

It is trained on a large dataset of diverse audio and is also a multitasking model that can perform multilingual speech recognition, speech translation, and language identification. A Transformer sequence-to-sequence model is trained on various speech processing tasks, including multilingual speech recognition, speech translation, spoken language identification, and voice activity detection.

Classification: multimodal/speech-and-audio · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [Faster Whisper transcription with CTranslate2](https://mdrss.com/multimodal/speech-and-audio/1709/1709.md)

faster-whisper is a reimplementation of OpenAI's Whisper model using CTranslate2, which is a fast inference engine for Transformer models. This implementation is up to 4 times faster than openai/whisper for the same accuracy while using less memory.

Classification: multimodal/speech-and-audio · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1

### [What is Voicebox?](https://mdrss.com/multimodal/speech-and-audio/853/853.md)

The full voice I/O stack, running locally on your machine. voicebox.sh • Docs • Download • Features • API • Troubleshooting Click the image above to watch the demo video on voicebox.sh Voicebox is a local-first AI voice studio — a free and open-source alternative to ElevenLabs and WisprFlow in one app.

Classification: multimodal/speech-and-audio · Feed: multimodal · Updated: 2026-08-04T12:22:38.168Z · Version: 1
