SnailText
EN

Romanian speech to text

Romanian speech to text, running on your own machine

Dictate in Romanian into any app on Mac or Windows. Romanian is still a low-resource language for speech models, and this page is honest about what that costs. SnailText runs Whisper locally, so nothing is uploaded.

Download for Macand start dictating in any app
cursor.txt
Vorbesc…
▍
Done

The short version

In the Whisper paper (Radford et al., 2022, arXiv:2212.04356), the large-v2 model scores 14.4% word error rate on Romanian FLEURS - workable, but clearly behind the major Western European languages, and research literature still describes Romanian as a low-resource language for speech recognition. That figure is clean read speech, so real dictation runs higher. Two caveats specific to Romanian are worth knowing before you start: the diacritics are frequently dropped, and the two correct Romanian letters with commas below are routinely confused with visually similar Turkish letters by software downstream. There is one more that concerns our own pipeline directly, covered below. SnailText runs the model locally with no Romanian-specific tuning layer.

Romanian dictation: local Whisper vs cloud STT

SnailText (local)Typical cloud Romanian STT
Where audio goesStays on your device, in RAMUploaded to a server for every phrase
Works offlineYes, after the model downloads onceNo, needs a connection every time
Romanian diacriticsFrequently dropped, and we do not restore themVaries; some vendors post-process
Sentence endings after hesitationCan be clipped by silence detection - see belowSame class of issue in most pipelines
Romanian-specific tuningNone - same open model as any local Whisper setupSome vendors tune per language, most run stock
Account / costNo account to startAccount + per-minute or per-seat billing

How accurate is Romanian speech to text with local Whisper?

In the Whisper paper itself (arXiv:2212.04356), the large-v2 model scores 14.4% word error rate on Romanian FLEURS. That is usable but clearly behind the big Western European languages, and it reflects the fact that research literature still describes Romanian as a low-resource language in the context of speech recognition.

That number is clean read speech and acts as a floor. Romanian ASR research also notes that dialect makes a large difference, with measurable divergence between standard Romanian and Moldovan speech, so the figure describes standard Romanian rather than the whole language area.

Accuracy tracks model size far more than it tracks local versus cloud - it is the same open-source Whisper either way. Use the largest model you can run.

What makes Romanian genuinely hard for speech models

Diacritics get dropped, and that is a real error rather than a cosmetic one. Romanian has five diacritic letters, and dropping them changes words. We do not add any Romanian post-processing to restore them, so whatever the model produces is what you get - worth a proofread, particularly on the two letters that carry commas below.

Two of those letters are routinely mis-encoded. The correct Romanian characters use a comma below, but visually similar Turkish letters with a cedilla have historically been substituted by software and fonts, and that substitution was the default for years in some environments. The result looks almost right and is wrong, which means it survives casual proofreading. This is a text-encoding hazard in whatever handles the text after us, and we make no claim to normalise it.

Silence detection can clip the end of a sentence, and this one is about our own pipeline. Romanian speech research documents that voice activity detection often truncates the final phonemes of a sentence when a speaker hesitates or a conversation turns over quickly, giving a clipped word instead of the full form. SnailText uses voice activity detection as part of its transcription pipeline, so this applies to us. We would rather name it than let you discover it: if you trail off or pause mid-sentence, check the last word.

What we actually offer for Romanian, and what we do not

For Spanish, German, French, Portuguese and Dutch, SnailText ships a language-specific prompt. For Romanian we do not. There is no Romanian tuning layer, no Romanian cleanup pass and no Romanian fine-tune. We do not restore dropped diacritics, we do not normalise the comma-versus-cedilla encoding issue, and we do not perform the morphological completion that research proposes for clipped sentence endings.

What SnailText does give you is the best open Romanian speech model running entirely on your own hardware, wired into a global hotkey that pastes text at your cursor in any app. The audio is processed in RAM and never uploaded, and it works with no connection at all.

Practical advice that follows from the section above: speak through to the end of your sentences rather than trailing off, and proofread diacritics. Both are usage habits rather than product features, and both will do more for your results than any setting.

Set the dictation language to Romanian explicitly rather than leaving it on automatic. Whisper decides the language from roughly the first thirty seconds and does not revisit that decision.

Talk instead of typing

Download for Mac

and start dictating in any app

Frequently asked questions

How accurate is Romanian speech to text?

+

In the Whisper paper (arXiv:2212.04356), the large-v2 model scores 14.4% word error rate on Romanian FLEURS. That is usable but behind the major Western European languages, and research literature still describes Romanian as low-resource for speech recognition. It is clean read speech, so everyday dictation runs higher. Dialect matters too - published work notes measurable divergence between standard Romanian and Moldovan speech.

Will it get Romanian diacritics right?

+

Not reliably, and we do not restore them. Romanian diacritics are meaning-bearing, so dropping one is a genuine error rather than a cosmetic issue. There is a second hazard: the two correct Romanian letters that carry commas below are routinely substituted by visually similar Turkish letters with cedillas in software and fonts. That looks almost right and is wrong. We do not normalise it.

Why does the last word of my sentence sometimes get cut off?

+

This is a known behaviour of voice activity detection and it applies to our pipeline too. Romanian speech research documents that silence detection often truncates the final phonemes of a sentence when the speaker hesitates or turns over quickly, producing a clipped word instead of the complete form. Speaking through to the end of your sentence rather than trailing off avoids most of it.

Do you tune SnailText specifically for Romanian?

+

No. There is no Romanian tuning layer, no Romanian cleanup pass and no Romanian fine-tune. We do not restore diacritics, normalise the comma-versus-cedilla encoding problem, or repair clipped word endings. For Romanian you get the same open Whisper model everyone runs, locally and privately.

Does it work for Moldovan Romanian?

+

Less predictably. Romanian ASR research finds that dialect makes a large difference to results and notes measurable divergence between standard Romanian and Moldovan speech. The benchmark figure on this page describes standard Romanian.

Does Romanian dictation work offline?

+

Yes. SnailText runs the Whisper speech model on your own Mac or Windows machine, so Romanian dictation works with no internet connection once the model has downloaded. The audio is processed in RAM and is never uploaded to any server.

·

Romanian speech to text, on your own machine.

Free to start on Mac and Windows. Press Option+Space (Mac) / Ctrl+Space (Windows), speak Romanian, and the text lands at your cursor in any app. No account, nothing uploaded, works offline.

Download for Macand start dictating in any app