Coverage, Measured
A dictionary is only as useful as the languages it actually reaches. This page reports what is in it, counted from the live database rather than estimated, including the places where it is thin enough to be nearly empty.
How These Numbers Were Made
Every figure below was read from the dictionary itself on 21 September 2026, by counting rows, not by repeating an earlier summary. Where an existing claim already stood, it was recomputed rather than trusted: the collision total came back at 8,538, matching what the dictionary page already said. Two other claims did not survive the same test and were corrected the same day. A number that has not been re-measured is a number that has started to drift.
What It Holds
- 827,591 words, every one of them carrying a braille form. There are no unrendered entries in the word dictionary.
- 50,498,948 corpus occurrences stand behind the frequency figure attached to each word, so "common" is a count rather than an impression.
- 2,786,522 word-source rows and 6,686,024 word-character rows, the evidence trail from a word back to where it was found.
- 312 source files, including 150,000 Kaikki entries, 188,223 Kaikki word forms, 108,897 cuneiform rows from CDLI, and 31,102 verses from each of several Bible translations.
- 1,597 entries in the language inventory, 1,596 universal seed words, and 216 curated name aliases so one language answers to its several spellings.
- 1,634 distinct language tags appear on actual words, slightly more than the 1,597 the inventory lists. The gap is reported rather than smoothed.
Where The Mass Actually Sits
Six language tags hold 719,693 of the 827,591 words, which is 87.0 per cent of the dictionary. The remaining 1,628 tags share 107,898 words between them.
Latin Script, 378,397 words, 45.7 per cent.
Hebrew, 81,212 words, 9.8 per cent.
Latin, 79,243 words, 9.6 per cent.
Greek, 68,342 words, 8.3 per cent.
English, 66,063 words, 8.0 per cent.
Arabic, 46,436 words, 5.6 per cent.
Latin Script is a script bucket rather than a single language, which is why it leads. The honest reading of this table is that the dictionary is a Scripture-language instrument first: Hebrew and Greek together carry 149,554 words, more than English and Arabic combined.
Where It Is Thin
The breadth figure is real but it is not depth, and the difference is large enough that stating one without the other would mislead.
The median language tag holds 3 words.
1,503 of the 1,634 tags, 92.0 per cent, hold fewer than 100 words.
1,161 tags hold fewer than 10.
So "1,594 languages" means the concept spine has been seeded that widely, not that any of those languages has been served to a useful depth. For most of them the dictionary can confirm that a concept exists and very little else. That is a starting point for translation work, not a substitute for it.
What It Cannot Do Yet
Go Deeper
The Braille Universal Dictionary, what it is and how to ask it a question.
What We Believe, the doctrine this work serves.
Data Handling, how the estate treats the material it holds.