Skip to main content
FluidVoice ships with support for eight distinct speech model families, ranging from the zero-download Apple Speech engine built into macOS to large multilingual models like Whisper Large. Each model runs entirely on your device — no audio leaves your Mac during transcription. The right choice depends on your language, how much disk space you can spare, and how fast you need words to appear on screen.
All models download to local storage on your Mac and run entirely on-device. Your voice and transcribed text are never sent to the cloud.

Model lineup


Model details

Nemotron Speech 3.5 — Ultra Fast Low Latency

Use this model when you want streaming-style multilingual dictation — words appear on screen as you speak with minimal lag. It covers roughly the same ~40-language range as the Multilingual variant but prioritises real-time throughput over peak accuracy. It requires Apple Silicon.

Nemotron 3.5 Multilingual

Choose this model when accuracy matters more than raw speed. It processes audio in a slightly less aggressive streaming mode than the Ultra Fast variant, producing cleaner transcripts across the same ~40-language range. It requires Apple Silicon.
Nemotron Speech 3.5 supports approximately 40 languages. Refer to the NVIDIA model card for the full and up-to-date list.

Parakeet Flash (Beta)

Parakeet Flash delivers the lowest perceived latency of any model in FluidVoice — text lands on screen fast enough to feel instantaneous. It is English-only and is the best pick for developers, power users, or anyone who dictates in English and wants speed above all else. It requires Apple Silicon.

Parakeet TDT v3

Parakeet TDT v3 is the recommended starting point for most multilingual users. It balances speed and accuracy across 25 European languages and is the default model FluidVoice suggests during onboarding. It requires Apple Silicon.
Bulgarian, Croatian, Czech, Danish, Dutch, English, Estonian, Finnish, French, German, Greek, Hungarian, Italian, Latvian, Lithuanian, Maltese, Polish, Portuguese, Romanian, Russian, Slovak, Slovenian, Spanish, Swedish, Ukrainian.

Parakeet TDT v2

Use Parakeet TDT v2 when you dictate exclusively in English and want the fastest possible throughput. It is functionally similar to Parakeet Flash but draws on an earlier generation of the model architecture. It requires Apple Silicon.

Cohere Transcribe

Cohere Transcribe prioritises transcription accuracy over raw speed and covers 14 languages including Mandarin, Japanese, Korean, Vietnamese, and Arabic — languages not available in the Parakeet family. At ~1.4 GB it is the largest download in the lineup. It requires Apple Silicon.
English, French, German, Italian, Spanish, Portuguese, Greek, Dutch, Polish, Mandarin, Japanese, Korean, Vietnamese, Arabic.

Apple Speech

Apple Speech uses the speech recognition engine built into macOS — no download required. It works on both Apple Silicon and Intel Macs and supports whatever languages your macOS installation has available. Use this model if you are on an Intel Mac and want the simplest possible setup, or if you are on Apple Silicon but have no disk space to spare.

Whisper Tiny / Base / Small / Medium / Large

The Whisper family offers the broadest language coverage (up to 99 languages) and is the only third-party option that runs on Intel Macs. Smaller variants (Tiny, Base) are fast and lightweight; larger variants (Medium, Large) trade speed for accuracy. Choose Whisper if you need a language that no other model covers, or if you are on an Intel Mac.
Whisper supports up to 99 languages depending on the model size you choose. Larger models (Medium, Large) recognise a wider range of languages more accurately than smaller ones. Refer to the OpenAI Whisper model card for the complete language list.

How to download a model

1

Open FluidVoice for the first time

The onboarding flow prompts you to select and download a model before you start dictating. Follow the on-screen steps to choose your language and latency preference, then let the download finish.
2

Switch or add a model later

Open Voice Engine, select the model you want, and click Download. You can switch between any downloaded models at any time by selecting the model, and clicking Activate.