Detect the language of a text
Paste a fragment and find out what language it is in.
Identifying a language from text has been a solved problem for decades and the solution is surprisingly simple: every language leaves a statistical fingerprint in the three-letter sequences it uses. Spanish is full of que, ent, ada; English of the, ing, and; German of sch, ich, der. Comparing how often those sequences appear in your text against each language gets it right with very little material.
That comparison is light enough to run as you type. One full sentence is usually enough; below twenty characters the result stops being reliable, and with isolated proper nouns or place names no method works.
Languages with their own script (Greek, Russian, Japanese, Korean, Arabic, Hebrew, Thai) are identified from the character range rather than the sequences, and there accuracy is essentially perfect.
How to use it
- 01 Paste a fragment of text, the longer the better.
- 02 The most likely language appears straight away, with a confidence measure.
- 03 If confidence is low, look at the other candidates: related languages get confused with short texts.
Frequently asked questions
How much text does it need?
With a sentence of twenty or thirty words accuracy is high. Below twenty characters the result is unreliable and the tool warns you.
Why does it confuse Portuguese with Spanish, or Norwegian with Danish?
Because they share much of their letter-sequence profile. With short texts the difference is tiny; with a full paragraph they separate cleanly.
Does it detect several languages in one text?
It returns the dominant one. For mixed text, analyse the fragments separately.
Related tools
This is what Eco does live.
These utilities work on files you already have. Eco transcribes and translates while the audio is still playing: meetings, classes, videos in another language.
Download Eco free10 free minutes per month · No card required