English
This page is available in English
The whole page — instructions, buttons and results.
Switch to English在简体、繁体(通用/台湾/香港)与词汇变体之间转换中文文本。简繁转换。全部在你的浏览器中运行。
Chinese Converter Tool
如何使用中文转换器
选择一种转换模式
用输入框上方的下拉菜单选择方向。默认是简体 → 台湾繁体(含词汇)(软件 → 軟體)——对在两岸之间发布内容的东盟读者最为实用。各模式按地区分组:通用、台湾和香港。
粘贴或输入你的文本
在左侧窗格输入或粘贴最多 10,000 个中文字。输入时即时转换,去抖延时 200 毫秒,保持流畅。右侧的输出窗格会自动更新。
复制或下载结果
点击复制把转换后的文本送到剪贴板,或点击下载另存为纯文本文件。差异高亮会标出哪些字发生了变化——若要干净地复制粘贴,可将其关闭。
一键互换方向
用 ⇄ 互换按钮翻转输入 ↔ 输出。模式会自动反转(例如 cn|twp 变为 twp|cn)。校对一段译文的两个方向时很顺手。
只转字形的模式(cn|t、cn|tw、cn|hk)往返完全一致;含词汇的 cn|twp ⇄ twp|cn 不一致——词汇映射是多对一的,例如 程序设计 → 程式設計 → 编程。
中文为何有多种书写形式——以及它在东盟为何重要
中文并非单一的书写语言,而是几种关系密切的书写体系。1956 年,中国大陆推行了《汉字简化方案》——这项改革精简了约 2,238 个常用字的笔画。台湾、香港、澳门,以及新加坡和马来西亚以外的大多数海外华人社群,都保留了传统字形。七十年过去,两套体系并存:约有 14 亿读者使用简体,约 3,500 万人使用繁体。这种分野看似只是地域之别,但对许多东盟家庭和企业而言,这是每天都要付出的切换成本。
常见的框定——简体对繁体——只是其中一条轴线。真实图景至少有四条。字形有别(龍 ↔ 龙)。即便同为繁体,各地词汇也有差异(台湾的 軟體 对香港的 軟件)。标点习惯有别。而文化语域——什么读起来正式、轻松或专业——的差异之大,足以让一段干净的逐字转换在母语读者看来仍「不对味」。粗糙的工具只搞定了第一条轴线,其余三条全数错失。
"新加坡约有 290 万华裔——约占其公民人口的 76%——但日常媒体消费却融合了大陆简体、台湾繁体与香港粤语等多种来源。" —— 新加坡统计局,《2023 年人口趋势》
字形转换无法解决的词汇难题
以「软件」一词为例。只看字符的转换器见到「软件」(简体)便输出「軟件」——字面上的繁体写法。这对香港是对的,因为香港通行「軟件」一词。但对台湾就错了,那里的读者期待的是「軟體」。再举三例:信息(大陆)↔ 資訊(台湾,「information」);程序 ↔ 程式(「program」);鼠标 ↔ 滑鼠(「mouse」)。词组级词典——OpenCC 所称的「p」模式——会识别源词并替换为该地区的目标词,而非套用字面字符映射。
再有就是歧义。「后」字在「皇后」(empress)中意为 queen,但在「后来」(later)中意为 behind / after。一对一的盲映射无从选择;词组感知的词典则至少读取一个相邻字来消歧。同样的问题也出现在 干(do / dry)、乾(只表 dry)、幹(只表 to do)上——这三个繁体字都依语境映射回单一的简体「干」。只有当词典识别出原词组时,往返转换(简体 → 繁体 → 简体)才会保持无损。
这对东盟的企业和读者意味着什么
对发布给台湾或香港客户的新加坡企业来说,軟體 与 軟件 之差,就是「听起来地道」与「听起来像翻译」之差。对马来西亚的华文报业——《星洲日报》《南洋商报》《中国报》——每天关于采用哪个地区变体的决定,都会在语气与可信度上波及不同读者群。字幕和配音团队、跨两岸销售的电商商品页、面向赴巴厘岛或普吉岛的香港游客的旅游内容营销——全都受益于词汇模式,而不只是字符模式。
隐私这一面同样重要。对法律草拟、内部备忘录、财务披露,或任何受 PDPA / GDPR 数据驻留规则约束的内容而言,「在你的浏览器中运行」不是一项功能——而是一条合规底线。本工具在首次输入时把 OpenCC 引擎作为自托管的 JavaScript 数据包加载;你的文字绝不会被传输、记录或在我们的服务器上处理。你可以自己验证:打开 DevTools,在输入时盯着网络标签,你会看到只有一个数据包加载,之后没有任何请求。
关于中文字形转换,你应该知道的 10 件事
相隔三千年. 甲骨文(商朝)与现代简体中文之间,字形保持不变的仅约 30%。
1956 年改革. 中国大陆的《汉字简化方案》于 1956 年 1 月 28 日正式公布——共涉及 2,238 个汉字。
台湾与香港守住了传统. 两地均未采用简体——繁体中文已连续保留了七十多年。
新加坡选择了简体. 新加坡于 1976 年正式采用简体中文——是唯一这样做的主要海外华人社群。
同字异义. 「后」在「皇后」中意为 queen,但在「后来」中意为 behind——靠上下文区分,这正是简单转换会失败的原因。
词汇分歧. 软件(大陆)、軟件(香港)、軟體(台湾)——「software」因地区不同而有三种说法。
OpenCC 始于 2010 年. 这款开源转换器经过 15 年以上的词典打磨,由世界各地的译者共同贡献。
新加坡 76% 为华裔. 据统计局《2023 年人口趋势》——这让得当的中文工具成为日常所需。
马来西亚 22% 为华裔. 约 670 万华裔,在学校多用简体,在文化媒体中则多用繁体。
在浏览器中即可完成. 现代 JavaScript 可在客户端完整运行 OpenCC 的词典引擎(约 1.1 MB)——你的文字永远不会离开你的设备。
常见问题
-
简体中文减少了常用字的笔画(例如 龍 → 龙、後 → 后),由中国大陆于 1956 年推行。繁体中文保留了较旧的字形,至今仍通行于台湾、香港、澳门,以及新加坡和马来西亚以外的大多数海外华人社群。除了笔画精简,两套体系在用词(词汇)、标点习惯,以及偶尔的语序上也有差异。官方简化的汉字约有 2,200 个,而繁体则数百年来大体保持不变。
-
两地都使用繁体字,但在哪些传统字形算「标准」以及词汇上有所分歧。台湾遵循教育部的标准字形(常称为「國字標準字體」)。香港遵循香港增补字符集,其中包含粤语专用字以及略有不同的字形标准。对大多数日常文本而言差异细微,但在出版、字幕和法律文件上就很重要——这正是 OpenCC 为各地分别提供词典的原因。
-
「含词汇」模式(twp/cnp)会进行词组级替换,而不仅仅是逐字转换。不启用它时,「软件」只会变成「軟件」(仅字符转换)。启用后,「软件」会变成「軟體」——这才是台湾读者真正使用的词。程序 → 程式(program)、信息 → 資訊(information)、鼠标 → 滑鼠(mouse)亦同。对营销文案、技术文档,或任何讲求地道地区中文的场合,你几乎总是会想用词汇模式。
-
不会。所有转换都在你的浏览器中以 OpenCC 引擎本地完成(首次输入时加载约 1.1 MB 的 JavaScript 数据包)。你的输入文本、输出文本以及转换过程都不会接触我们的服务器。这对法律草拟、财务材料、内部备忘录,或任何受 PDPA / GDPR 数据驻留要求约束的内容都很重要。你可以自行验证:打开 DevTools → 网络,看看你输入时发生了什么——只有数据包会加载,你的文字不会被上传。
-
OpenCC 是基于词典的转换器,而不是翻译器——这既是它的长处,也是它的边界:它不会像大语言模型那样凭空造字,但也读不懂词典之外的语境。对现代标准中文的连贯文句,它是可靠的:词典收录了多字词条,所以「头发 → 頭髮」「干杯 → 乾杯」都正确,而且能原样转换回来。人名才是它出错的地方,而且是无声地出错。简化时把若干个不同的繁体字并成了一个,反向转换时词典只挑最常见的那一个,并不会先问这是不是名字:「我姓余」变成「我姓餘」,「干先生」变成「幹先生」,两个都错。可是「范冰冰」和「钟先生」又是对的,因为这两个名字收在词典里。页面上没有任何提示告诉你落在哪一种情况,所以人名请逐一人工校对。至于文言文、古代名称,或词典未收录的专门术语,OpenCC 会保留原字不变,而不是去猜。
-
新加坡官方使用简体中文,所以大多数输入都是简体。如果你是写给新加坡读者:无需转换。如果你要发布给台湾读者,选「简体 → 台湾繁体(含词汇)」(本工具的默认模式——cn|twp)。如果你要发布给香港读者,选「简体 → 香港繁体」(cn|hk)——注意 OpenCC 没有为香港单设词汇模式,因为粤语专用词最好手动处理。
-
可以,但有须注意之处。字形会正确转换——學而時習之(繁体)↔ 学而时习之(简体)没问题。不过文言文使用的词汇和语法与现代中文不同,所以词汇模式(twp/cnp)派不上用场。对文言文,请用纯字形模式(cn|t 或 t|cn)。文言成语通常在各地都有稳定字形,转换起来很干净。
-
有三个原因。第一,有些字在两套体系中完全相同——人、中、文、大 无需转换。第二,人名、地名和品牌名即便用另一套体系书写,也常按惯例保留原形。第三,少数罕见或专门的字可能未收录在 OpenCC 的词典中。如果你发现明显的遗漏,很可能是词典缺口——OpenCC 在 GitHub 上开源,并接受贡献。
-
有——每次转换上限 10,000 字。这是安全限制,而非技术限制(OpenCC 可处理数百万字)。转换以 200 毫秒去抖,而过长的文本会让浏览器标签页在词典引擎遍历输入时短暂卡顿。1 万字的上限涵盖了大多数用例(一篇典型新闻报道约 500–1500 字;一篇长篇文章约 3000–5000 字)。要处理更大批量,请把文本分段。
-
可以。OpenCC 采用 MIT 许可证,而转换后的文本是你自己的工作成果——使用本工具不带来任何授权限制。我们不对你的输入或输出主张任何权利。本工具本身免费,无需注册;我们通过页面上的展示广告盈利,而非收费或付费分级。至于专业出版,我们仍建议由人工复核一遍语域、语气,以及词典可能遗漏的地区惯用语。
Related News
You may be interested in these recent stories from our newsroom.
-
Saudi Arabia's humain-m3 Is MiniMax's M3, Under MiniMax's Licence
Same 428 billion parameters, same 23 billion active, same mixture-of-experts design as the model Shanghai open-weighted in June. The Arabic...
-
OpenAI's GPT-6 Astra Costs 2.5 Times More and Matches GPT-5.6 on the Independent Index
OpenAI called the launch the start of the AGI era. Artificial Analysis, which OpenAI neither ran nor selected, scores Astra 61 — exactly wha...
-
Six AI Releases in Three Days, and Two Were New Models
The 72-hour launch count folds together two new models, a safeguard variant, two point releases, a pricing tier and a feature that is not a...
方法与来源
计算方式
Conversion is a dictionary lookup, not a translation. The OpenCC engine walks the input against a trie of single- and multi-character entries, longest match first, rewrites what it finds and passes through what it does not. Eight directed modes cover general Simplified/Traditional plus Taiwan and Hong Kong standard forms; the two Taiwan phrase modes additionally substitute regional vocabulary, so 软件 becomes 軟體 rather than only 軟件. Everything runs in the browser from a self-hosted bundle.
本工具依据的标准
- Longest-match-first over multi-character entries, not character-by-character substitution. This is why 头发 → 頭髮 and 干杯 → 乾杯 are correct and survive a round trip back to the input, where a per-character table would produce 头髮 and 干杯 from the wrong 干.
- Character form and regional vocabulary are separate choices, kept separate. cn|tw rewrites the glyphs only; cn|twp also swaps the terms. Picking the wrong one yields text that is orthographically Traditional and lexically mainland — 臺灣資訊軟體 comes back as 台湾资讯软体 under tw|cn but 台湾信息软件 under twp|cn.
- Taiwan and Hong Kong are not one "Traditional" and the tool does not treat them as one. Verified against the served bundle: 里面 → 裡面 in Taiwan mode but 裏面 in Hong Kong mode; 着急 → 著急 in Taiwan mode but 着急 in Hong Kong mode.
- ⚠️ NAMES ARE WHERE THIS BREAKS, AND IT BREAKS SILENTLY. Simplification merged distinct Traditional characters; converting back, the dictionary resolves each merge by frequency rather than by whether the string is a name. Verified 2026-08-31: 余华 → 餘華 and 我姓余 → 我姓餘 are both wrong (the surname is 余; 餘 means "surplus"), and 干先生 → 幹先生 is wrong. Yet 范冰冰 and 钟先生 come out right because those names are in the dictionary. Nothing on the page tells you which case you are in, so proofread every name by hand in either direction.
- It is a converter, not a translator, and it never invents. A character with no dictionary entry is passed through unchanged rather than guessed at — which is the right behaviour for 文言文, ancient names and specialist terms, and is the one respect in which it is genuinely safer than a language model.
- The text never leaves the browser. The engine is a self-hosted bundle loaded on first input, and no input, output or intermediate state is transmitted or logged.
资料来源
- opencc-js v1.3.1 (MIT), the conversion engine: https://github.com/nk2028/opencc-js — the JavaScript port of Open Chinese Convert. The bundle served at /tools/chinese-converter/opencc-bundle.js was confirmed byte-identical to that package's dist/umd/full.js on 2026-08-31 (SHA-256 01c0921d3685f410…, 1,124,229 bytes), so the audited package and the shipped copy are the same file.
- Open Chinese Convert (OpenCC), Apache-2.0: https://github.com/BYVoid/OpenCC — the upstream project the conversion tables originate from.
- Dictionary licensing: per opencc-js's own THIRD_PARTY_LICENSES.md the bundled tables are generated at build time from opencc-data (https://github.com/nk2028/opencc-data) and redistributed under the Apache License 2.0 — a different licence from the MIT-licensed engine code, and one this tool previously did not name.
- Taiwan standard character forms: 常用國字標準字體表 (4,808 characters, 1982) and 次常用國字標準字體表 (6,341 characters, 1982), Ministry of Education, Republic of China: https://language.moe.gov.tw/ — the orthography the tool's Taiwan modes target.
- Hong Kong standard character forms: 常用字字形表 (List of Graphemes of Commonly-used Chinese Characters), Education Bureau, first issued 1986 and reissued in typeset form in 2007 with 4,762 characters: https://www.edbchinese.hk/lexlist/ — the orthography the tool's Hong Kong modes target, and the reason 裡/裏 and 著/着 differ from Taiwan.
可能导致本页过时的因素
- The dictionaries are compiled into the vendored bundle and nothing refetches at runtime. Updating them means bumping opencc-js and re-copying the bundle; a bump that is not re-copied would leave the audited version and the served bytes disagreeing, which is why the byte comparison above is recorded rather than a version banner.
- ⚠️ Regional vocabulary dates faster than character forms do. The phrase tables encode what a term was called when they were compiled — new technical vocabulary diverges between markets before any dictionary records it, so a recent coinage will convert its characters and keep its mainland wording.
- Hong Kong has no vocabulary mode, only character forms. Cantonese-specific lexis is not substituted by any mode and has to be handled by a human.
Pick up where you left off
Stored only in this browser — never sent to our servers.