🇹🇼 Chinese (traditional) 25 million speakers

The 5,000 most common Chinese (traditional) words

Ranked by how often they actually turn up in spoken dialogue, not by how useful somebody guessed they would be.
Every number below is counted from a corpus, and the corpus is named at the bottom.

What each step buys you

Learn the first 100 words of Chinese (traditional) and you follow 45.2% of everyday speech. The first 1000 take you to 71.0%. After that every thousand words buys less than the thousand before it, which is exactly why the order matters.

Words knownSpeech followed
1021.1%
5037.2%
10045.2%
25056.1%
50063.7%
1,00071.0%
2,00078.0%
3,00081.9%
5,00086.4%
8,00090.1%
10,00091.7%
15,00094.4%
20,00096.0%

Comfortable reading starts around 98%. Everything before that is still worth doing, but the first hundred words buy you a foothold, not comprehension.

The first 100 words of Chinese (traditional)

In order. The percentage is the share of ordinary speech you cover once you know that word and every word above it.

  1. 14.39%
  2. 28.31%
  3. 311.93%
  4. 414.44%
  5. 516.25%
  6. 617.43%
  7. 7我們18.45%
  8. 819.44%
  9. 920.32%
  10. 1021.1%
  11. 11и21.83%
  12. 1222.52%
  13. 1323.2%
  14. 1423.87%
  15. 1524.46%
  16. 1625.04%
  17. 1725.6%
  18. 1826.16%
  19. 19知道26.64%
  20. 2027.13%
  21. 21什麼27.6%
  22. 2228.05%
  23. 2328.49%
  24. 2428.9%
  25. 2529.31%
  26. 2629.71%
  27. 2730.12%
  28. 28他們30.52%
  29. 2930.9%
  30. 3031.28%
  31. 3131.65%
  32. 3232.01%
  33. 3332.37%
  34. 3432.71%
  35. 3533.05%
  36. 3633.39%
  37. 3733.71%
  38. 3834.02%
  39. 3934.32%
  40. 40不是34.62%
  41. 4134.91%
  42. 4235.19%
  43. 43一個35.46%
  44. 4435.73%
  45. 45什么35.99%
  46. 4636.25%
  47. 47l36.49%
  48. 4836.73%
  49. 49可以36.96%
  50. 5037.2%
  51. 5137.42%
  52. 5237.64%
  53. 5337.86%
  54. 54這個38.08%
  55. 5538.29%
  56. 5638.51%
  57. 57就是38.72%
  58. 58τ38.92%
  59. 5939.13%
  60. 6039.33%
  61. 61這是39.53%
  62. 62這樣39.72%
  63. 63怎麼39.9%
  64. 64如果40.09%
  65. 6540.27%
  66. 6640.45%
  67. 67你們40.62%
  68. 6840.79%
  69. 6940.95%
  70. 7041.11%
  71. 7141.26%
  72. 72自己41.42%
  73. 73現在41.57%
  74. 7441.72%
  75. 75不會41.86%
  76. 76真的42.01%
  77. 7742.15%
  78. 78因為42.3%
  79. 79沒有42.44%
  80. 8042.58%
  81. 8142.72%
  82. 8242.86%
  83. 8343.0%
  84. 8443.14%
  85. 85不能43.28%
  86. 86謝謝43.41%
  87. 8743.54%
  88. 88先生43.67%
  89. 89告訴43.8%
  90. 90you43.93%
  91. 91不要44.06%
  92. 9244.18%
  93. 9344.31%
  94. 9444.43%
  95. 95the44.55%
  96. 96只是44.68%
  97. 97需要44.8%
  98. 9844.92%
  99. 99那個45.04%
  100. 100m45.16%

Chinese (traditional) does not separate words with spaces, so what is counted here are the segments the corpus splits on, not whole words. The order is still right and the coverage is still real, but read the list as building blocks rather than dictionary entries.

Open all 5,000 words in the codex →

In the codex you tap the ones you already know and the number moves. It stays in your browser, no account.

Where Chinese (traditional) comes from

Sino-Tibetan Chinese

Tone does the work that endings do in Indo-European.
Same job, opposite tool.

See the whole family tree →

Words Chinese (traditional) has that English never built

  • 缘分 yuen-FENThe pull that brings two people together, or does not

All the untranslatable words →

Where the numbers come from

Frequencies counted over the OpenSubtitles 2018 dialogue corpus, 31,871,997 tokens of Chinese (traditional) in total. Lists compiled by Hermit Dave, released under CC BY-SA 3.0. Subtitles skew towards conversation, which is the point: this is the language people speak, not the language people publish.