Language Detector

Identify the likely language of any pasted text, instantly in your browser.

—
Detected Language

What is the Language Detector?

This tool identifies the most likely language of a piece of text using two signals, entirely offline in your browser with no model to download: which Unicode writing system (script) the characters belong to, and — for languages that share the Latin alphabet — which language's common function words (like "the," "de," or "und") appear most often.

How it actually works

Scripts like Bengali, Devanagari, Arabic, Cyrillic, Greek, Thai, Hangul and the combination of Hiragana/Katakana are each encoded in their own distinct Unicode ranges, so text written in one of these is identified immediately and reliably just from which characters appear. English, Spanish, French, German, Portuguese, Italian and Dutch all share the Latin alphabet, so for these the tool instead counts how often each language's handful of extremely common, distinctive short words (articles, conjunctions, prepositions) appear in the text, and picks whichever language's word-list matches best.

Honest limitations

This is a lightweight heuristic, not a trained machine-learning model — it works well on a full sentence or paragraph in one of the ~16 covered languages, but can misfire on very short snippets (a few words), text that mixes multiple languages, or a language outside this list. Treat the result as a solid first guess, not a certified classification.

How to use it

  1. Paste or type any text.
  2. The detected language updates automatically as you type.

Nothing you type is uploaded anywhere — the entire analysis runs in plain JavaScript on your own device.

Want the "why" behind this tool? Read How a Language Detector Works Without Downloading an AI Model.