The human languages NADA covers.

NADA is a shared standard for showing code in a human language. This page lists every language we map and how far its coverage has come. 130+ languages are usable today across 156 mapped languages, built from 4,182,470 translated terms.

Coverage grows continuously. This page is a snapshot, not the full dataset.


Three states.

Every language is in one of three states. We publish the state, not a percentage. A single percentage is easy to misread across 156 languages. Exact per-language coverage is available on request [TO CONFIRM: per-language coverage figures / where they are published].

Complete Covers the words used in real code. Usable end to end. Refined as vocabulary grows.
In progress Past the early threshold, still filling gaps. Many already cover more than 70 percent.
Planned Mapped and queued. Waiting for a contributor or council review.

Coverage, language by language.

A selection from the 156 mapped languages, grouped by script family. Names appear in English and in each language's own script. The full machine-readable register is part of the open language data [TO CONFIRM: link to the published register / dataset].

European

Spanish Español Complete
French Français Complete
German Deutsch Complete
Portuguese Português Complete
Italian Italiano Complete
Polish Polski Complete
Ukrainian Українська Complete
Greek Ελληνικά In progress

South & Southeast Asian

Hindi हिन्दी Complete
Bengali বাংলা Complete
Indonesian Bahasa Indonesia Complete
Tamil தமிழ் In progress
Telugu తెలుగు In progress
Thai ไทย In progress
Vietnamese Tiếng Việt Complete
Marathi मराठी Planned

East Asian

Chinese (Simplified) 简体中文 Complete
Japanese 日本語 Complete
Korean 한국어 Complete
Chinese (Traditional) 繁體中文 In progress

Middle Eastern & African

Arabic العربية Complete
Persian فارسی In progress
Turkish Türkçe Complete
Hebrew עברית In progress
Swahili Kiswahili In progress
Hausa Harshen Hausa Planned
Amharic አማርኛ Planned
Yoruba Èdè Yorùbá Planned

This is a partial list. 130+ languages are usable today; 79 already pass 70 percent coverage. Proprietary or restricted languages are intentionally not listed.


From a single word to a complete language.

Every language starts from English and is built term by term by people who read it natively. Each contribution is reviewed before it enters the standard, which keeps the vocabulary consistent. No single person or company decides what a word should be.

01 A term is proposed. A native reader maps a software word to the right word in their language, in the context where it is used in code.
02 The AI council reviews it. An independent panel of models checks the proposal for meaning, consistency, and register, and flags anything that reads wrong to a fluent speaker. A human steward then signs off.
03 It enters the language data. Accepted terms join the open CC-BY dataset and ship in the next build, counting toward that language's coverage.
04 Coverage moves up a state. As the words used in real code fill in, a language moves from Planned to In progress to Complete.
A language belongs to the people who speak it, and so does its place in the register.

Review thresholds and council composition are documented in the governance memo [TO CONFIRM: link to governance / council documentation].


If your language is missing or incomplete, you can help.

You do not need to be a programmer to contribute. You need to read your language well. Completing one language often advances several related languages at the same time. The data stays free and open for everyone.

Language data CC-BY · 135 languages mapped · the register is public.