For two decades, Dragon NaturallySpeaking was the gold standard for dictation accuracy. That position is changing. OpenAI's Whisper engine, particularly the Large v3 Turbo variant released in 2024, has reached parity with Dragon for general dictation and exceeds it for many specific cases. This article walks through what we actually know about the comparison, with sources.

What "accuracy" means in dictation

Dictation accuracy is usually measured as Word Error Rate (WER), which is the percentage of words that come out wrong in the transcript. Lower is better. A WER of 5% means 95 words out of 100 are correct.

WER varies enormously based on conditions: the speaker's voice, the microphone quality, background noise, the vocabulary domain, the dictation style. A benchmark number is only meaningful when you know the conditions it was measured under.

Whisper Large v3 Turbo: the published numbers

OpenAI published Whisper Large v3 Turbo in 2024 as a faster variant of Whisper Large v3 with comparable accuracy. The reported word error rates on standard benchmarks:

  • Whisper Large v3 (original): 13.20% WER averaged across multilingual datasets
  • Whisper Large v3 Turbo: 13.40% WER — within 0.2% of the original
  • English-only clean conditions: 3-5% WER (97-95% accuracy)

The Turbo variant runs about 4× faster than v3 with essentially the same accuracy. This is the engine CaringDictate uses.

For the typical CaringDictate use case (English, clear room, decent microphone), real-world accuracy averages 96-97% on continuous dictation. This matches OpenAI's published numbers for English-clean conditions and is what we report on our marketing pages.

Dragon's accuracy claims

Dragon historically advertised 99% accuracy. The fine print: this number was achieved on trained users with high-quality headset microphones in quiet environments, after extensive voice training. A first-day untrained Dragon user typically sees more like 92-94% accuracy.

For trained Dragon Professional Individual users with proper setup, real-world accuracy is typically in the 95-98% range on general dictation. For specialized domains (medical, legal), trained Dragon can edge into 98-99% on the vocabulary it has been customized for.

The honest head-to-head

For an untrained user, with a generic USB headset, dictating conversational English: Whisper-based engines (CaringDictate) usually win. Whisper does not require voice training and works well out of the box on the broad range of voices it was trained on (680,000+ hours of multilingual audio).

For a trained user, with a top-quality microphone (Plantronics, Andrea USB), dictating specialized vocabulary: trained Dragon Professional can still edge ahead on the specific vocabulary it has been trained for. The advantage narrows considerably for general writing.

For non-native English speakers, speakers with accents, or speakers with mild dysarthria from medical conditions: Whisper has a clear advantage. The training corpus was much more diverse than Dragon's, and the recognition handles speech variation noticeably better.

For speakers with significant dysarthria (severe Parkinson's, ALS affecting speech, CP affecting speech): neither Whisper nor Dragon is the right tool. Voiceitt is specifically built for this case and is significantly better than either, though much more expensive.

Speed

Whisper Turbo's real advantage over Dragon is speed. Whisper Large v3 Turbo on modern hardware (Intel i5 from the last 5 years, or Apple Silicon, or any modern AMD CPU) produces transcription faster than realtime — meaning you can dictate a sentence and the words appear before you finish speaking. Dragon also processes in realtime but has more noticeable latency on the first word of each utterance.

For long-form dictation, this speed difference is mostly psychological. For short-message dictation (Facebook comments, quick emails), the responsiveness of Whisper Turbo feels notably better.

The training-data question

Dragon's traditional accuracy advantage came from supervised voice training: you'd read a long passage to Dragon, and it would build a personal model that recognized your voice better than its default model. This worked but required significant upfront time.

Whisper does not work this way. It was trained on a vast multilingual corpus once, by OpenAI, and the resulting model is what every Whisper user runs. You do not train it on your voice. It either works on your voice or it doesn't — and the broad training data means it works on most voices reasonably well.

The implication for users: with Dragon, you can spend training time to improve accuracy. With Whisper, what you get on day one is what you get on day 365. For most users this is a feature, not a bug — they didn't want to spend an hour training Dragon anyway. For power users with technical vocabulary, the inability to voice-train Whisper is a real limitation.

The maintenance question

Dragon Professional Individual has been frozen at version 16 since 2023. The recognition engine is not being updated. Microsoft has pivoted Dragon to enterprise medical and is letting consumer Dragon age out.

Whisper, by contrast, is actively developed by OpenAI and the open source community. Whisper v3 was released in 2023, Whisper v3 Turbo in 2024, and successor models are in active development. CaringDictate ships with v3 Turbo today and will update to newer models as a free update when they prove better.

Over a five-year horizon, this is the most important difference. A frozen recognition engine will continue to work, but it will not benefit from the substantial improvements happening elsewhere in the field.

For specific user groups

Elderly users with general English vocabulary: Whisper (via CaringDictate) is the better choice. No training, no setup complexity, works well out of the box.

Medical professionals dictating clinical notes: Trained Dragon Medical One is still industry standard. Whisper's medical vocabulary is good but not as specialized.

Lawyers dictating legal documents: Trained Dragon Professional with custom vocabulary is hard to beat. Whisper handles most legal English well but lacks the customizability.

Adults with medical conditions affecting hands but not speech: Whisper via CaringDictate is what we'd recommend. No training overhead, better accuracy on speech variation, and much cheaper.

Non-native English speakers: Whisper has a clear advantage. Multilingual training corpus means it handles accents far better than Dragon's English-trained models.

Bottom line

Whisper Large v3 Turbo and trained Dragon Professional are now within a few percentage points of each other on general dictation accuracy. For most users, especially the elderly and users with medical conditions, the practical advantages of Whisper (no training, better accent handling, active development, much lower cost) outweigh Dragon's residual advantage on specialized vocabulary.

CaringDictate is built on Whisper Large v3 Turbo because we believe this is the right engine for our audience. The 30-day refund means you can compare for yourself with no risk.

Where to go from here

Read the full CaringDictate vs Dragon comparison for product-level differences (price, UX, updates, refund), and our privacy guide if local vs cloud processing matters for your situation.