The 5,000 most common Macedonian words
Ranked by how often they actually turn up in spoken dialogue, not by how useful somebody guessed they would be.
Every number below is counted from a corpus, and the corpus is named at the bottom.
What each step buys you
Learn the first 100 words of Macedonian and you follow 52.8% of everyday speech. The first 1000 take you to 75.2%. After that every thousand words buys less than the thousand before it, which is exactly why the order matters.
| Words known | Speech followed | |
|---|---|---|
| 10 | 23.4% | |
| 50 | 44.4% | |
| 100 | 52.8% | |
| 250 | 62.4% | |
| 500 | 69.1% | |
| 1,000 | 75.2% | |
| 2,000 | 80.9% | |
| 3,000 | 84.1% | |
| 5,000 | 87.9% | |
| 8,000 | 91.2% | |
| 10,000 | 92.6% | |
| 15,000 | 95.0% | |
| 20,000 | 96.4% |
Comfortable reading starts around 98%. Everything before that is still worth doing, but the first hundred words buy you a foothold, not comprehension.
The first 100 words of Macedonian
In order. The percentage is the share of ordinary speech you cover once you know that word and every word above it.
- 1да4.51%
- 2не7.44%
- 3се10.22%
- 4е12.85%
- 5на14.81%
- 6го16.7%
- 7и18.52%
- 8што20.22%
- 9ќе21.85%
- 10во23.4%
- 11за24.76%
- 12тоа26.03%
- 13ти27.13%
- 14дека28.19%
- 15од29.17%
- 16со30.13%
- 17ми31.04%
- 18си31.89%
- 19ја32.68%
- 20како33.38%
- 21јас34.01%
- 22ги34.63%
- 23ова35.17%
- 24сум35.71%
- 25добро36.18%
- 26ме36.64%
- 27те37.09%
- 28само37.54%
- 29ли37.96%
- 30нема38.36%
- 31но38.75%
- 32ако39.14%
- 33а39.52%
- 34тој39.86%
- 35беше40.19%
- 36треба40.52%
- 37кога40.83%
- 38сега41.14%
- 39многу41.44%
- 40знам41.74%
- 41кој42.03%
- 42му42.32%
- 43има42.6%
- 44нешто42.88%
- 45може43.15%
- 46мене43.41%
- 47би43.68%
- 48каде43.94%
- 49сакам44.2%
- 50тебе44.45%
- 51тука44.7%
- 52дали44.94%
- 53така45.19%
- 54зошто45.43%
- 55сите45.66%
- 56ајде45.89%
- 57ред46.11%
- 58ни46.32%
- 59па46.53%
- 60знаеш46.74%
- 61овде46.94%
- 62таа47.14%
- 63можам47.34%
- 64биде47.53%
- 65по47.72%
- 66мислам47.91%
- 67сакаш48.09%
- 68до48.27%
- 69еден48.44%
- 70ви48.62%
- 71уште48.79%
- 72или48.96%
- 73некој49.11%
- 74ама49.27%
- 75време49.43%
- 76ништо49.58%
- 77тие49.73%
- 78таму49.88%
- 79сме50.03%
- 80мора50.18%
- 81сте50.32%
- 82пред50.47%
- 83можеби50.61%
- 84имам50.75%
- 85ве50.89%
- 86малку51.02%
- 87молам51.16%
- 88ние51.29%
- 89еј51.43%
- 90повеќе51.56%
- 91одиме51.69%
- 92навистина51.82%
- 93можеш51.94%
- 94колку52.07%
- 95него52.19%
- 96мојот52.32%
- 97работа52.44%
- 98тогаш52.56%
- 99оди52.67%
- 100здраво52.79%
Open all 5,000 words in the codex →
In the codex you tap the ones you already know and the number moves. It stays in your browser, no account.
Where Macedonian comes from
Proto-Indo-European › Balto-Slavic › Slavic › South Slavic › Macedonian
Cut off from the rest by Hungarian and Romanian arriving in between.
Where the numbers come from
Frequencies counted over the OpenSubtitles 2018 dialogue corpus, 20,688,310 tokens of Macedonian in total. Lists compiled by Hermit Dave, released under CC BY-SA 3.0. Subtitles skew towards conversation, which is the point: this is the language people speak, not the language people publish.