A script born from necessity
When Vietnam came under the cultural orbit of the Tang Dynasty (7th–10th century CE), classical Chinese — chữ Hán — became the prestige writing system for administration, law, and literature. But Vietnamese is not Chinese. It has its own tones, its own grammar, its own vocabulary rooted in Mon-Khmer. Writing the spoken language in borrowed Chinese characters was as awkward as writing English in Japanese kanji.
Sometime around the 10th century, Vietnamese scholars invented a solution: Chữ Nôm (字喃), literally "Southern script" or "vernacular script." Instead of inventing an entirely new alphabet, they took Chinese characters and bent them to Vietnamese ends using two main strategies:
- Phonetic borrowing: A Chinese character was chosen for its sound, even if its meaning was irrelevant. The character 南 (nán, "south") could be used just for its sound value.
- Semantic + phonetic compounds: Two characters were fused — one providing meaning, one providing sound — to create an entirely new glyph not found in any Chinese dictionary. This produced characters so complex that even trained scholars sometimes couldn't read them.
The scale: a repertoire of 20,000 characters
The Chinese standard repertoire (GB 18030) encodes around 87,000 characters; Chữ Nôm drew on a working repertoire of roughly 20,000 characters, of which several thousand are Vietnamese inventions — glyphs that appear in no Chinese text, created purely to represent Vietnamese words. Vietnam's national standard TCVN 6909:2001 fixes 9,299 Nôm glyphs, about half of them Vietnam-specific; the full range may never be precisely counted because many manuscripts are still undigitized.
The script was never standardized. Different scribes created different characters for the same word, so reading Chữ Nôm required knowing the individual hand of the author. This was both a weakness (hard to learn) and a curious strength (it resisted mechanical reproduction and therefore colonial print-culture takeover — briefly).
The masterpiece: Truyện Kiều
The single greatest monument of Chữ Nôm literature is Truyện Kiều (The Tale of Kiều), a 3,254-line epic poem by Nguyễn Du (1765–1820). The story — adapted from a Chinese novel — follows a young woman named Kiều through fifteen years of tragedy, prostitution, war, and eventual reunion. Every educated Vietnamese person can recite lines from it. It is, by every measure, Vietnam's national poem.
The irony is exquisite: Vietnam's greatest work of literature was written in a script that almost nobody alive today can read in its original form. When you read Truyện Kiều in a modern edition, you are reading a transliteration into the Latin-based quốc ngữ, not the original glyphs Nguyễn Du wrote.
The erasure
French colonization arrived in stages from 1858. Catholic missionaries had already been promoting the Latin-based quốc ngữ since the 17th century (see the companion article on Alexandre de Rhodes). The French colonial administration saw a practical advantage: a population that could only read the new script was cut off from pre-colonial texts, laws, and cultural memory. In 1919, the Confucian examinations — which tested classical Chinese, never Chữ Nôm — were held for the last time and abolished. Reading Nôm had always rested on that same classical training, and now the training itself was gone. Within a generation, functional literacy in the script collapsed.
The revival that might not arrive in time
The Hán Nôm Institute (Viện Nghiên cứu Hán Nôm), founded in Hanoi in 1970, has been racing against time to digitize manuscripts before they decay. Unicode's CJK Unified Ideographs Extension blocks now include thousands of Chữ Nôm characters, allowing the script to be typed on a computer for the first time. A small but passionate community of scholars and hobbyists is relearning it. But with fewer than a few hundred fluent readers alive, the window is closing fast.