Hungarian text difficulty checker
Paste any Hungarian text here and Beeblio will tell you, word by word, how difficult it is. Each word is matched against frequency bands — 300, 800, 1,500, 3,000, 5,000, 8,000, and 12,000 most common Hungarian words — so you can see at a glance how much of the text sits within your current vocabulary and how much lies beyond it.
This matters because of how comprehensible input works: if you already know roughly 80% of the words in a passage, you have a real chance of picking up the rest from context. The checker helps you find texts that hit that sweet spot — challenging enough to stretch you, familiar enough to make sense. Vocabulary ranks come from the wordfreq corpus, so they reflect real-world written and spoken Hungarian rather than a textbook word list.
The results highlight new words (ones just outside your known range) and rare words (very low-frequency items unlikely to repay memorisation yet). Use that breakdown to decide whether a news article, short story, or YouTube transcript is worth tackling today — or better saved for later.
How it works
Paste a text in Hungarian (up to 6,000 characters). Every word is looked up in a frequency list; for each band — the top 300, 800, 1,500, 3,000, 5,000, 8,000 and 12,000 words — you see the share of the text a learner at that band already knows. Comprehensible reading sits around 95–98% known: we call ≥97% easy, 92–97% just right, 85–92% hard, and below that too hard.
You also get the new words worth learning (unknown but common, in frequency order) and the rare words — names, jargon, typos — that no learner should stop for.
Questions
What do the frequency bands actually mean?
Each band is a threshold: the top 300 words are the most common Hungarian words of all, the top 800 include those plus the next most frequent ones, and so on up to 12,000. If a word falls inside your chosen band, Beeblio treats it as known; if it falls outside, it's flagged as new or rare. The bands come from the wordfreq corpus, which ranks words by how often they appear across real Hungarian text.
Why does the checker sometimes flag a word I definitely know?
Hungarian is a highly agglutinative language, meaning a single root word can appear in dozens of inflected forms. Beeblio analyses the surface form it sees in the text, and an unusual suffix or compound may push a familiar root into a lower frequency band. If you see a word you already know flagged as rare, that's likely why.
How do I use this to choose the right reading material?
A good rule of thumb is to look for texts where the words within your current band cover roughly 80% of the total — that level of coverage gives you enough context to work out unfamiliar words. If the checker shows far more than 20% unknown words, the text is probably too difficult for comfortable reading practice right now, and you might save it for when your vocabulary has grown.