SnailText
EN

Korean transcription

Korean transcription, running on your own machine

Dictate in Korean into any app on Mac or Windows. Korean is one of Whisper's weaker languages, and this page says so with the numbers. SnailText runs the model locally, so nothing is uploaded.

Download for Macand start dictating in any app
cursor.txt
말하는 중…
▍
Done

The short version

Honest version first: Korean is markedly weaker than the European languages on the same model. In the Whisper paper (Radford et al., 2022, arXiv:2212.04356), the large-v2 model scores 14.3% character error rate on Korean FLEURS, against 4.0% for Italian, 5.3% for Japanese and 5.6% for Russian in that same table - roughly three times the error rate. That is character error rate rather than word error rate, because Korean word spacing is inconsistent enough to corrupt word-level scoring. Model size matters more here than anywhere: the compact models score 36.1% and 19.6% on the same benchmark, so the free tier is not enough for serious Korean dictation. SnailText runs the model locally on Mac and Windows with no Korean-specific tuning layer, and we would rather tell you that than imply otherwise.

Korean dictation: local Whisper vs cloud STT

SnailText (local)Typical cloud Korean STT
Where audio goesStays on your device, in RAMUploaded to a server for every phrase
Works offlineYes, after the model downloads onceNo, needs a connection every time
Accuracy vs European languagesRoughly 3x the error rate, and we say soUsually presented as uniform across languages
Compact models good enough?No - 36.1% and 19.6% on the same benchmarkModel size usually not disclosed
Korean-specific tuningNone - same open model as any local Whisper setupSome vendors tune per language, most run stock
Account / costNo account to startAccount + per-minute or per-seat billing

How accurate is Korean speech to text with local Whisper?

In the Whisper paper itself (arXiv:2212.04356), the large-v2 model scores 14.3% character error rate on Korean FLEURS. For context from the same table and the same model: Italian 4.0%, Japanese 5.3%, Polish 5.4%, Russian 5.6%. Korean is roughly three times the error rate of those languages, and any page that puts it beside them without saying so is misleading you.

That figure is a character error rate, not a word error rate. The paper uses character-level scoring for Chinese, Japanese and Korean, and for Korean there is an additional reason: word spacing is applied inconsistently even by human transcribers, to the point that researchers introduced a space-normalised error rate specifically to stop spacing disagreements from corrupting Korean evaluations. So part of any published Korean number is spacing convention rather than misrecognition.

Model size matters more for Korean than for any other language on this site. On the same benchmark the compact model scores 36.1% and the small model 19.6%, against 14.3% for the largest. That is not a small gradient - it is the difference between unusable and workable. The free-tier compact models are genuinely not enough for serious Korean dictation, and we would rather say that than sell you a disappointment.

What makes Korean genuinely hard for speech models

Word spacing is inconsistent, even among people. Korean spacing rules exist but are widely varied in practice, and transcription corpora contain sequences of inconsistent spacing that produce invalid evaluations. This is why researchers proposed a space-normalised word error rate with flexible spacing rules for Korean. The practical consequence for you is that some of what looks like an error is a spacing convention the model chose differently from you.

One wrong syllable can change the meaning. Many Sino-Korean morphemes are single syllables, and phonologically similar syllables can map to different meanings or syntactic roles, so a single-character substitution can change what a sentence is asking. Research on Korean spoken question answering identifies exactly this: meaning flips hide under a low character error rate. A small number therefore understates the damage more in Korean than in a language where errors look like obvious typos.

Politeness is grammar, and it lives at the end of every sentence. Korean marks formality and deference in the verb ending. Getting that ending wrong is not a cosmetic error - it changes who the speaker is being to whom. We make no claim about handling speech levels or honorific forms, and you should check sentence endings in anything that goes to another person.

What we actually offer for Korean, and what we do not

For Spanish, German, French, Portuguese and Dutch, SnailText ships a language-specific prompt. For Korean we do not. There is no Korean tuning layer, no Korean cleanup pass and no Korean fine-tune. We do not normalise spacing, we do not handle Hanja, and we make no claim about honorific or speech-level correctness.

What SnailText gives you is the best open Korean speech model running entirely on your own hardware, with a global hotkey that pastes text at your cursor in any app, and audio that never leaves your machine. Given the accuracy above, the honest positioning is that this is a privacy-first way to run the same model everyone runs - not a claim that Korean recognition is solved.

If your Korean dictation needs to be production-accurate, use the largest model you can run and expect to proofread. If you are evaluating tools, the useful comparison is not our number against a competitor's marketing - it is whether a tool tells you the number at all.

One practical tip: set the dictation language to Korean explicitly rather than leaving it on automatic. Whisper decides the language from roughly the first thirty seconds and does not revisit that decision.

Talk instead of typing

Download for Mac

and start dictating in any app

Frequently asked questions

How accurate is Korean speech to text?

+

In the Whisper paper (arXiv:2212.04356), the large-v2 model scores 14.3% character error rate on Korean FLEURS. In the same table and on the same model, Italian scores 4.0%, Japanese 5.3% and Russian 5.6% - so Korean runs roughly three times the error rate of those languages. It is character error rate rather than word error rate because Korean word spacing is inconsistent enough to corrupt word-level scoring. Real dictation runs above the benchmark figure.

Is the free tier good enough for Korean?

+

Honestly, no, not for serious use. On the same FLEURS benchmark the compact model scores 36.1% and the small model 19.6%, against 14.3% for the largest. Korean is the language where model size matters most on this site, and the compact free-tier models are the difference between unusable and workable. Use the largest model you can run.

Why is Korean measured in character error rate?

+

Two reasons. The Whisper paper uses character-level scoring for Chinese, Japanese and Korean because word segmentation is not well defined in those scripts. For Korean specifically there is a second reason: word spacing is applied inconsistently even by human transcribers, to the point that researchers introduced a space-normalised error rate to stop spacing disagreements from invalidating Korean evaluations.

Do you tune SnailText specifically for Korean?

+

No. There is no Korean tuning layer, no Korean cleanup pass and no Korean fine-tune. We do not normalise word spacing, we do not support Hanja, and we make no claim about honorific or speech-level handling. For Korean you get the same open Whisper model everyone runs, locally and privately.

Does it get honorifics and speech levels right?

+

We make no claim about it. Korean marks formality and deference grammatically, in the verb ending at the end of each sentence, so an error there changes the social register rather than looking like a typo. Check sentence endings in anything going to another person.

Does Korean dictation work offline?

+

Yes. SnailText runs the Whisper speech model on your own Mac or Windows machine, so Korean dictation works with no internet connection once the model has downloaded. The audio is processed in RAM and is never uploaded to any server.

·

Korean speech to text, on your own machine.

Free to start on Mac and Windows. Press Option+Space (Mac) / Ctrl+Space (Windows), speak Korean, and the text lands at your cursor in any app. No account, nothing uploaded, works offline.

Download for Macand start dictating in any app