🇮🇳 हिन्दी reference data
Fully supported — read, tap-to-define, mine vocabulary, and practice frequency-banded cloze in Hindi. Devanagari writes a space between words, so a tap on किताब reaches the entry with no extra engine, and combining marks stay inside the word. NFC folds the two nukta spellings (क़ and क + ◌़) onto one key, and the danda । ends a sentence. One honest caveat: the sentence bank is smaller than the European packs, because Tatoeba holds fewer Hindi–English pairs. It covers about 72% of the two thousand most useful practice words. This is the Devanagari register. Urdu writes the same spoken language in a different script and needs its own pack.
What's inside
Provenance & licensing
Built entirely from open data — nothing here is derived from copyrighted dictionaries. See the methodology for how frequency drives vocabulary learning in Lector.
- Dictionary and inflections: Hindi Wiktionary entries extracted through Kaikki.org. Keys are the printed spelling — Hindi has no letter case and needs no fold beyond NFC, which collapses the two nukta spellings onto one key. CC BY-SA 4.0
- Cloze sentences: Hindi–English sentence pairs from Tatoeba, frequency-ranked with wordfreq and filtered against the on-device dictionary. CC BY 2.0 FR