Simplified-Traditional Chinese Converter
Free online converter between simplified and traditional Chinese using a built-in common-character table. Convert large amounts of text in real time for editing, translation review, or language learning.
What Simplified-Traditional Conversion Really Is: Character Mapping, Not Translation
Simplified and traditional Chinese share the same language and the same meanings; they differ only in how a subset of characters is written. The difference comes from the Chinese script simplification that began in the 1950s. A Simplified-Traditional converter is therefore not a translator (meaning is unchanged) — it is a per-character lookup table that swaps each simplified glyph for its traditional counterpart, or vice versa. This tool ships a built-in common-character table of 181 high-frequency entries and maps each character individually. Because it reads characters, not words, the ambiguities and regional differences discussed below are unavoidable.
One-to-Many Mappings: One Simplified, Several Traditional Forms
When the simplified script was created, several traditional characters were often merged into a single simplified one. Reversing that — simplified to traditional — means one simplified character can map to two or more traditional forms. Note: such characters inevitably lose their context during a literal conversion; the tool can only pick the most frequent form.
| Simplified | Traditional candidates | Meaning / usage |
|---|---|---|
| 发 | 發 / 髮 | 發 = to emit, develop; 髮 = hair |
| 后 | 後 / 后 | 後 = before/after; 后 = monarch, queen |
| 干 | 乾 / 幹 / 干 | 乾 = dry; 幹 = trunk, to do; 干 = clean, to interfere |
| 里 | 裡 / 里 | 裡 = inside; 里 = mile, neighborhood |
| 面 | 麵 / 面 | 麵 = noodle; 面 = face, surface |
| 复 | 復 / 複 | 復 = recover; 複 = complex |
| 历 | 歷 / 曆 | 歷 = experience; 曆 = calendar |
| 台 | 臺 / 檯 / 颱 / 台 | 臺 = Taiwan; 檯 = counter; 颱 = typhoon; 台 = platform |
| 只 | 隻 / 只 | 隻 = a measure word; 只 = only |
| 制 | 製 / 制 | 製 = to manufacture; 制 = system, to control |
| 采 | 採 / 彩 / 采 | 採 = to pick; 彩 = color; 采 = demeanor |
| 折 | 摺 / 折 | 摺 = to fold; 折 = to break, to lose |
| 松 | 鬆 / 松 | 鬆 = loose; 松 = pine |
| 卷 | 捲 / 卷 | 捲 = to roll; 卷 = exam paper |
| 别 | 別 / 彆 | 別 = to separate; 彆 = awkward |
| 脏 | 髒 / 臟 | 髒 = dirty; 臟 = internal organ |
These are only the high-frequency representatives. We catalogued 48 representative one-to-many groups in total; the statistics follow.
High-Ambiguity Statistics (based on this tool's built-in table)
Placing the one-to-many mappings inside the tool's 181 built-in common characters shows how common ambiguity really is:
| Metric | Value |
|---|---|
| One-to-many groups catalogued | 48 |
| 2-way ambiguity (2 traditional forms) | 43 |
| 3-way ambiguity (3 traditional forms) | 3 |
| 4-way ambiguity (4 traditional forms) | 2 |
| Maximum traditional forms for one character | 4 (台) |
| Share of built-in common characters | about 26.5% |
Note: roughly 1 in every 4 built-in common characters can be ambiguous when converting simplified to traditional. Personal names, place names, technical terms, and idioms are the four riskiest categories — always proofread those after a whole-sentence conversion.
Character Sets and Encodings: GB2312 / GB18030-2022 / Big5 / Unicode
Behind the simplified-traditional debate sit several different Chinese encoding standards. The current mandatory national standard in mainland China is GB18030-2022 (Information Technology — Chinese Coded Character Set), which has the broadest coverage; traditional Chinese in the Taiwan region historically used Big5; and international exchange relies on Unicode. The comparison:
| Standard | Published / year | Han character scale | Encoding | Code point range |
|---|---|---|---|---|
| GB2312 | 1980 | 6,763 chars | double-byte | 0xA1A1–0xFEFE |
| GBK | 1995 | 21,003 chars | double-byte | 0x8140–0xFEFE |
| GB18030-2022 | 2022 (mandatory) | 87,887 Han chars (CJK Ext A–G) | single/double/quad-byte | 0x00–0x7F + double + quad |
| Big5 | 1984 (Taiwan region) | 13,053 chars | double-byte | 0xA140–0xF9FE |
| Unicode CJK Unified | — | 20,992 code points | UTF-8/16/32 | U+4E00–U+9FFF |
| Unicode CJK Ext A | — | 6,592 code points | UTF-8/16/32 | U+3400–U+4DBF |
| Unicode CJK Ext B | — | 42,720 code points | UTF-8/16/32 | U+20000–U+2A6DF |
This tool runs entirely in your browser and emits UTF-8, so whether your source text came from a GB18030 or a Big5 file, pasting it here converts correctly in both directions.
Variant Forms and Regional Glyph Differences
Even within "traditional", different dictionaries and regions disagree on a few characters; these are called variant forms. Common groups:
| Group | Note |
|---|---|
| 為 / 为 | 为 is simplified; the traditional standard is 為, though some older glyphs write 爲 |
| 裏 / 裡 | both traditional; 裏 is the classical form, 裡 is more common in the Taiwan region |
| 够 / 夠 | 够 is simplified; 夠 is traditional, common in the Taiwan region |
| 群 / 羣 | 群 is the mainland standard; 羣 is a variant / traditional form |
| 回 / 迴 | 迴 is a variant of 回, used in 迴旋 (revolve), 迴廊 (corridor) |
| 线 / 線 | 线 simplified; 線 traditional |
| 窗 / 窓 / 牕 | 窓 and 牕 are variant forms of 窗 (window) |
| 强 / 強 / 彊 | 強 is traditional; 彊 is a variant |
Regional Vocabulary: Chinese Mainland / Taiwan Region / Hong Kong SAR
Beyond individual characters, the three regions differ far more at the vocabulary level. The same concept can be written completely differently. Note: this tool's default "Simplified to Traditional (standard)" converts characters only and does not automatically switch to regional vocabulary; the table below helps you decide whether a second pass is needed.
| Concept | Mainland China | Taiwan region | Hong Kong SAR |
|---|---|---|---|
| software | 软件 | 軟體 | 軟件 |
| information | 信息 | 資訊 | 資訊 |
| video | 视频 | 影片 | 影片 |
| taxi | 出租车 | 計程車 | 的士 |
| printer | 打印机 | 印表機 | 列印機 |
| mouse | 鼠标 | 滑鼠 | 滑鼠 |
| internet | 互联网 | 網際網路 | 互聯網 |
| folder | 文件夹 | 資料夾 | 檔案夾 |
| memory | 内存 | 記憶體 | 記憶體 |
| CD / disc | 光盘 | 光碟 | 光碟 |
OpenCC Pipelines: s2t / s2tw / s2hk
The open-source library OpenCC offers several configurations. This tool's "Simplified to Traditional (standard)" direction is equivalent to OpenCC's s2t; if you target a specific region, here is how the other two differ:
| Config | Target | Vocabulary bias | Character layer |
|---|---|---|---|
| s2t | standard traditional | neutral, no regional words | identical |
| s2tw | Taiwan traditional | applies Taiwan-region words (軟體 / 影片 / 資訊) | identical |
| s2hk | Hong Kong traditional | applies Hong Kong SAR words (軟件 / 影片 / 的士) | identical |
Note: the three differ mainly at the vocabulary layer; the traditional character forms are essentially the same. This tool defaults to s2t, so if you write for readers in the Taiwan region or the Hong Kong SAR, after conversion you may do a second pass such as 軟件→軟體, 視頻→影片, 信息→資訊, taxi→計程車/的士, to match local usage.
Three Real Examples (default "Simplified to Traditional" direction)
The examples below use the page's default direction "Simplified to Traditional"; the input is exactly what you would type into the default text box, and the output is produced by this tool's per-character mapping.
Example 1: How one-to-many fails in practice
Input: 小明后天要出差,出发前请理发。
Output: 小明後天要出差,出發前請理發。
Analysis: both occurrences of 发 are rendered with the most common traditional form 發 — "出发→出發" is correct, but "理发" should be "理髮" (haircut, a noun). The tool gave "理發". This is the classic failure of "one character, many traditional forms" when semantics are invisible; a human should change "理發" back to "理髮".
Example 2: Default s2t output and regional vocabulary
Input: 这个软件可以处理网络文件,信息很全。
Output: 這個軟件可以處理網絡文件,信息很全。
Analysis: the character layer converts fully (软件→軟件, 网络→網絡). But 軟件 happens to match Hong Kong SAR usage; for the Taiwan region you would change it to 軟體. Meanwhile 信息 is not in the tool's built-in table and stays as-is — which also shows the default s2t does no regional vocabulary substitution, so 信息→資訊 (a s2tw-style change) must be handled separately.
Example 3: A clean one-to-one conversion
Input: 我们认真学习汉字。
Output: 我們認真學習漢字。
Analysis: every character here is a one-to-one pair with no ambiguity, so the correct traditional text comes out directly. Pure one-to-one text is the safest scenario for simplified-traditional conversion.
Further Reading
For more on text processing, see the Complete Case Converter Guide, and the tools Case Converter, Word Counter, and URL Encode.
Frequently Asked Questions
Does it cover every character? This tool ships a built-in common-character table (181 high-frequency entries) covering daily high-frequency characters; it does not guarantee coverage of every rare character or CJK extension-region character, and uncovered characters are kept unchanged.
Why do some conversions look ambiguous? Because of one-to-many mappings, where one simplified character maps to several traditional forms (e.g. 发→發/髮, 后→後/后). The tool uses the most common form; please double-check special contexts.
How does the tool handle one-to-many characters like 发, 后, 干? It picks the most frequent traditional form by character frequency (发→發, 后→後, 干→幹) and cannot sense your real meaning; confirm manually for names, places, and terms.
How different are traditional forms across the Chinese mainland, the Taiwan region, and the Hong Kong SAR? The traditional character forms are mostly the same; the bigger difference is regional vocabulary (e.g. 软件/軟體/軟件, 出租车/計程車/的士). This tool defaults to s2t (characters only), so regional vocabulary needs a second pass.
What is the difference between OpenCC s2t, s2tw, and s2hk? s2t emits standard traditional; s2tw applies Taiwan-region vocabulary; s2hk applies Hong Kong SAR vocabulary; the three are identical at the character layer and differ only at the vocabulary layer.
Why is manual review still recommended after whole-sentence conversion? Personal names, place names, technical terms, and idioms often contain high-ambiguity one-to-many characters, and mechanical conversion may pick the wrong one (e.g. 理發 should be 理髮). In formal contexts, always confirm key terms by hand.