English
This page is available in English
The whole page — instructions, buttons and results.
Switch to English在簡體、繁體(通用/台灣/香港)與詞彙變體之間轉換中文文本。簡繁轉換。全部在你的瀏覽器中運行。
Chinese Converter Tool
如何使用中文轉換器
選擇一種轉換模式
用輸入框上方的下拉菜單選擇方向。默認是簡體 → 台灣繁體(含詞彙)(软件 → 軟體)——對在兩岸之間發佈內容的東盟讀者最為實用。各模式按地區分組:通用、台灣和香港。
粘貼或輸入你的文本
在左側窗格輸入或粘貼最多 10,000 箇中文字。輸入時即時轉換,去抖延時 200 毫秒,保持流暢。右側的輸出窗格會自動更新。
複製或下載結果
點擊複製把轉換後的文本送到剪貼板,或點擊下載另存為純文本文件。差異高亮會標出哪些字發生了變化——若要乾淨地複製粘貼,可將其關閉。
一鍵互換方向
用 ⇄ 互換按鈕翻轉輸入 ↔ 輸出。模式會自動反轉(例如 cn|twp 變為 twp|cn)。校對一段譯文的兩個方向時很順手。
只轉字形的模式(cn|t、cn|tw、cn|hk)往返完全一致;含詞彙的 cn|twp ⇄ twp|cn 並不一致——詞彙映射是多對一的,例如 程序设计 → 程式設計 → 编程。
中文為何有多種書寫形式——以及它在東盟為何重要
中文並非單一的書寫語言,而是幾種關係密切的書寫體系。1956 年,中國大陸推行了《漢字簡化方案》——這項改革精簡了約 2,238 個常用字的筆畫。台灣、香港、澳門,以及新加坡和馬來西亞以外的大多數海外華人社羣,都保留了傳統字形。七十年過去,兩套體系並存:約有 14 億讀者使用簡體,約 3,500 萬人使用繁體。這種分野看似只是地域之別,但對許多東盟家庭和企業而言,這是每天都要付出的切換成本。
常見的框定——簡體對繁體——只是其中一條軸線。真實圖景至少有四條。字形有別(龍 ↔ 龙)。即便同為繁體,各地詞彙也有差異(台灣的 軟體 對香港的 軟件)。標點習慣有別。而文化語域——什麼讀起來正式、輕鬆或專業——的差異之大,足以讓一段乾淨的逐字轉換在母語讀者看來仍「不對味」。粗糙的工具只搞定了第一條軸線,其餘三條全數錯失。
"新加坡約有 290 萬華裔——約佔其公民人口的 76%——但日常媒體消費卻融合了大陸簡體、台灣繁體與香港粵語等多種來源。" —— 新加坡統計局,《2023 年人口趨勢》
字形轉換無法解決的詞彙難題
以「软件」一詞為例。只看字符的轉換器見到「软件」(簡體)便輸出「軟件」——字面上的繁體寫法。這對香港是對的,因為香港通行「軟件」一詞。但對台灣就錯了,那裏的讀者期待的是「軟體」。再舉三例:信息(大陸)↔ 資訊(台灣,「information」);程序 ↔ 程式(「program」);鼠标 ↔ 滑鼠(「mouse」)。詞組級詞典——OpenCC 所稱的「p」模式——會識別源詞並替換為該地區的目標詞,而非套用字面字符映射。
再有就是歧義。「后」字在「皇后」(empress)中意為 queen,但在「后来」(later)中意為 behind / after。一對一的盲映射無從選擇;詞組感知的詞典則至少讀取一個相鄰字來消歧。同樣的問題也出現在 干(do / dry)、乾(只表 dry)、幹(只表 to do)上——這三個繁體字都依語境映射回單一的簡體「干」。只有當詞典識別出原詞組時,往返轉換(簡體 → 繁體 → 簡體)才會保持無損。
這對東盟的企業和讀者意味着什麼
對發佈給台灣或香港客户的新加坡企業來説,軟體 與 軟件 之差,就是「聽起來地道」與「聽起來像翻譯」之差。對馬來西亞的華文報業——《星洲日報》《南洋商報》《中國報》——每天關於採用哪個地區變體的決定,都會在語氣與可信度上波及不同讀者羣。字幕和配音團隊、跨兩岸銷售的電商商品頁、面向赴巴厘島或普吉島的香港遊客的旅遊內容營銷——全都受益於詞彙模式,而不只是字符模式。
隱私這一面同樣重要。對法律草擬、內部備忘錄、財務披露,或任何受 PDPA / GDPR 數據駐留規則約束的內容而言,「在你的瀏覽器中運行」不是一項功能——而是一條合規底線。本工具在首次輸入時把 OpenCC 引擎作為自託管的 JavaScript 數據包加載;你的文字絕不會被傳輸、記錄或在我們的服務器上處理。你可以自己驗證:打開 DevTools,在輸入時盯着網絡標籤,你會看到只有一個數據包加載,之後沒有任何請求。
關於中文字形轉換,你應該知道的 10 件事
相隔三千年. 甲骨文(商朝)與現代簡體中文之間,字形保持不變的僅約 30%。
1956 年改革. 中國大陸的《漢字簡化方案》於 1956 年 1 月 28 日正式公佈——共涉及 2,238 個漢字。
台灣與香港守住了傳統. 兩地均未採用簡體——繁體中文已連續保留了七十多年。
新加坡選擇了簡體. 新加坡於 1976 年正式採用簡體中文——是唯一這樣做的主要海外華人社羣。
同字異義. 「后」在「皇后」中意為 queen,但在「后来」中意為 behind——靠上下文區分,這正是簡單轉換會失敗的原因。
詞彙分歧. 软件(大陸)、軟件(香港)、軟體(台灣)——「software」因地區不同而有三種説法。
OpenCC 始於 2010 年. 這款開源轉換器經過 15 年以上的詞典打磨,由世界各地的譯者共同貢獻。
新加坡 76% 為華裔. 據統計局《2023 年人口趨勢》——這讓得當的中文工具成為日常所需。
馬來西亞 22% 為華裔. 約 670 萬華裔,在學校多用簡體,在文化媒體中則多用繁體。
在瀏覽器中即可完成. 現代 JavaScript 可在客户端完整運行 OpenCC 的詞典引擎(約 1.1 MB)——你的文字永遠不會離開你的設備。
常見問題
-
簡體中文減少了常用字的筆畫(例如 龍 → 龙、後 → 後),由中國大陸於 1956 年推行。繁體中文保留了較舊的字形,至今仍通行於台灣、香港、澳門,以及新加坡和馬來西亞以外的大多數海外華人社羣。除了筆畫精簡,兩套體系在用詞(詞彙)、標點習慣,以及偶爾的語序上也有差異。官方簡化的漢字約有 2,200 個,而繁體則數百年來大體保持不變。
-
兩地都使用繁體字,但在哪些傳統字形算「標準」以及詞彙上有所分歧。台灣遵循教育部的標準字形(常稱為「國字標準字體」)。香港遵循香港增補字符集,其中包含粵語專用字以及略有不同的字形標準。對大多數日常文本而言差異細微,但在出版、字幕和法律文件上就很重要——這正是 OpenCC 為各地分別提供詞典的原因。
-
「含詞彙」模式(twp/cnp)會進行詞組級替換,而不僅僅是逐字轉換。不啓用它時,「软件」只會變成「軟件」(僅字符轉換)。啓用後,「软件」會變成「軟體」——這才是台灣讀者真正使用的詞。程序 → 程式(program)、信息 → 資訊(information)、鼠标 → 滑鼠(mouse)亦同。對營銷文案、技術文檔,或任何講求地道地區中文的場合,你幾乎總是會想用詞彙模式。
-
不會。所有轉換都在你的瀏覽器中以 OpenCC 引擎本地完成(首次輸入時加載約 1.1 MB 的 JavaScript 數據包)。你的輸入文本、輸出文本以及轉換過程都不會接觸我們的服務器。這對法律草擬、財務材料、內部備忘錄,或任何受 PDPA / GDPR 數據駐留要求約束的內容都很重要。你可以自行驗證:打開 DevTools → 網絡,看看你輸入時發生了什麼——只有數據包會加載,你的文字不會被上傳。
-
OpenCC 是基於詞典的轉換器,而不是翻譯器——這既是它的長處,也是它的邊界:它不會像大型語言模型那樣憑空造字,但也讀不懂詞典以外的語境。對現代標準中文的連貫文句,它是可靠的:詞典收錄了多字詞條,所以「头发 → 頭髮」「干杯 → 乾杯」都正確,而且能原樣轉換回來。人名才是它出錯的地方,而且是無聲地出錯。簡化時把若干個不同的繁體字併成了一個,反向轉換時詞典只揀最常見的那一個,並不會先問這是不是名字:「我姓余」變成「我姓餘」,「干先生」變成「幹先生」,兩個都錯。可是「范冰冰」和「钟先生」又是對的,因為這兩個名字收在詞典裏。頁面上沒有任何提示告訴你落在哪一種情況,所以人名請逐一人手校對。至於文言文、古代名稱,或詞典未收錄的專門術語,OpenCC 會保留原字不變,而不是去猜。
-
新加坡官方使用簡體中文,所以大多數輸入都是簡體。如果你是寫給新加坡讀者:無需轉換。如果你要發佈給台灣讀者,選「簡體 → 台灣繁體(含詞彙)」(本工具的默認模式——cn|twp)。如果你要發佈給香港讀者,選「簡體 → 香港繁體」(cn|hk)——注意 OpenCC 沒有為香港單設詞彙模式,因為粵語專用詞最好手動處理。
-
可以,但有須注意之處。字形會正確轉換——學而時習之(繁體)↔ 学而时习之(簡體)沒問題。不過文言文使用的詞彙和語法與現代中文不同,所以詞彙模式(twp/cnp)派不上用場。對文言文,請用純字形模式(cn|t 或 t|cn)。文言成語通常在各地都有穩定字形,轉換起來很乾淨。
-
有三個原因。第一,有些字在兩套體系中完全相同——人、中、文、大 無需轉換。第二,人名、地名和品牌名即便用另一套體系書寫,也常按慣例保留原形。第三,少數罕見或專門的字可能未收錄在 OpenCC 的詞典中。如果你發現明顯的遺漏,很可能是詞典缺口——OpenCC 在 GitHub 上開源,並接受貢獻。
-
有——每次轉換上限 10,000 字。這是安全限制,而非技術限制(OpenCC 可處理數百萬字)。轉換以 200 毫秒去抖,而過長的文本會讓瀏覽器標籤頁在詞典引擎遍歷輸入時短暫卡頓。1 萬字的上限涵蓋了大多數用例(一篇典型新聞報道約 500–1500 字;一篇長篇文章約 3000–5000 字)。要處理更大批量,請把文本分段。
-
可以。OpenCC 採用 MIT 許可證,而轉換後的文本是你自己的工作成果——使用本工具不帶來任何授權限制。我們不對你的輸入或輸出主張任何權利。本工具本身免費,無需註冊;我們通過頁面上的展示廣告盈利,而非收費或付費分級。至於專業出版,我們仍建議由人工複核一遍語域、語氣,以及詞典可能遺漏的地區慣用語。
Related News
You may be interested in these recent stories from our newsroom.
-
Saudi Arabia's humain-m3 Is MiniMax's M3, Under MiniMax's Licence
Same 428 billion parameters, same 23 billion active, same mixture-of-experts design as the model Shanghai open-weighted in June. The Arabic...
-
OpenAI's GPT-6 Astra Costs 2.5 Times More and Matches GPT-5.6 on the Independent Index
OpenAI called the launch the start of the AGI era. Artificial Analysis, which OpenAI neither ran nor selected, scores Astra 61 — exactly wha...
-
Six AI Releases in Three Days, and Two Were New Models
The 72-hour launch count folds together two new models, a safeguard variant, two point releases, a pricing tier and a feature that is not a...
方法與來源
計算方式
Conversion is a dictionary lookup, not a translation. The OpenCC engine walks the input against a trie of single- and multi-character entries, longest match first, rewrites what it finds and passes through what it does not. Eight directed modes cover general Simplified/Traditional plus Taiwan and Hong Kong standard forms; the two Taiwan phrase modes additionally substitute regional vocabulary, so 软件 becomes 軟體 rather than only 軟件. Everything runs in the browser from a self-hosted bundle.
本工具採用的標準
- Longest-match-first over multi-character entries, not character-by-character substitution. This is why 头发 → 頭髮 and 干杯 → 乾杯 are correct and survive a round trip back to the input, where a per-character table would produce 头髮 and 干杯 from the wrong 干.
- Character form and regional vocabulary are separate choices, kept separate. cn|tw rewrites the glyphs only; cn|twp also swaps the terms. Picking the wrong one yields text that is orthographically Traditional and lexically mainland — 臺灣資訊軟體 comes back as 台湾资讯软体 under tw|cn but 台湾信息软件 under twp|cn.
- Taiwan and Hong Kong are not one "Traditional" and the tool does not treat them as one. Verified against the served bundle: 里面 → 裡面 in Taiwan mode but 裏面 in Hong Kong mode; 着急 → 著急 in Taiwan mode but 着急 in Hong Kong mode.
- ⚠️ NAMES ARE WHERE THIS BREAKS, AND IT BREAKS SILENTLY. Simplification merged distinct Traditional characters; converting back, the dictionary resolves each merge by frequency rather than by whether the string is a name. Verified 2026-08-31: 余华 → 餘華 and 我姓余 → 我姓餘 are both wrong (the surname is 余; 餘 means "surplus"), and 干先生 → 幹先生 is wrong. Yet 范冰冰 and 钟先生 come out right because those names are in the dictionary. Nothing on the page tells you which case you are in, so proofread every name by hand in either direction.
- It is a converter, not a translator, and it never invents. A character with no dictionary entry is passed through unchanged rather than guessed at — which is the right behaviour for 文言文, ancient names and specialist terms, and is the one respect in which it is genuinely safer than a language model.
- The text never leaves the browser. The engine is a self-hosted bundle loaded on first input, and no input, output or intermediate state is transmitted or logged.
資料來源
- opencc-js v1.3.1 (MIT), the conversion engine: https://github.com/nk2028/opencc-js — the JavaScript port of Open Chinese Convert. The bundle served at /tools/chinese-converter/opencc-bundle.js was confirmed byte-identical to that package's dist/umd/full.js on 2026-08-31 (SHA-256 01c0921d3685f410…, 1,124,229 bytes), so the audited package and the shipped copy are the same file.
- Open Chinese Convert (OpenCC), Apache-2.0: https://github.com/BYVoid/OpenCC — the upstream project the conversion tables originate from.
- Dictionary licensing: per opencc-js's own THIRD_PARTY_LICENSES.md the bundled tables are generated at build time from opencc-data (https://github.com/nk2028/opencc-data) and redistributed under the Apache License 2.0 — a different licence from the MIT-licensed engine code, and one this tool previously did not name.
- Taiwan standard character forms: 常用國字標準字體表 (4,808 characters, 1982) and 次常用國字標準字體表 (6,341 characters, 1982), Ministry of Education, Republic of China: https://language.moe.gov.tw/ — the orthography the tool's Taiwan modes target.
- Hong Kong standard character forms: 常用字字形表 (List of Graphemes of Commonly-used Chinese Characters), Education Bureau, first issued 1986 and reissued in typeset form in 2007 with 4,762 characters: https://www.edbchinese.hk/lexlist/ — the orthography the tool's Hong Kong modes target, and the reason 裡/裏 and 著/着 differ from Taiwan.
可能使本頁過時的因素
- The dictionaries are compiled into the vendored bundle and nothing refetches at runtime. Updating them means bumping opencc-js and re-copying the bundle; a bump that is not re-copied would leave the audited version and the served bytes disagreeing, which is why the byte comparison above is recorded rather than a version banner.
- ⚠️ Regional vocabulary dates faster than character forms do. The phrase tables encode what a term was called when they were compiled — new technical vocabulary diverges between markets before any dictionary records it, so a recent coinage will convert its characters and keep its mainland wording.
- Hong Kong has no vocabulary mode, only character forms. Cantonese-specific lexis is not substituted by any mode and has to be handled by a human.
Pick up where you left off
Stored only in this browser — never sent to our servers.