Chinese Charset Converter (GBK/Big5/UTF-8 编码)

Convert Chinese text between UTF-8 and the legacy GB2312 / GBK / Big5 byte encodings with a live hex preview. Text to bytes turns UTF-8 text into the target charset's hexadecimal bytes; bytes to text turns those hex bytes back into readable text. Decoding uses the browser's native TextDecoder; encoding is table-driven and runs entirely locally. Characters missing from the target charset are flagged inline — never silently mis-encoded.
For whole files (upload + convert + download), use the File Encoding Converter. All processing happens locally in your browser — your text is never uploaded to any server.
Operation: Charset:
Hex case: Separator:
Open File reads a UTF-8 text file; Ctrl+Enter converts.
Enter text or hex bytes, pick the charset and operation, then click Convert Now.
How to read the output: 中 in GBK is D6 D0, in Big5 A4 A4 and in UTF-8 E4 B8 AD. The hex preview groups bytes with spaces by default; uncheck that to get the compact form used by many tools. When a character does not exist in the chosen charset it is shown as 3F and listed below so nothing is silently mis-encoded. Decoding tolerates spaces, commas, line breaks, 0x prefixes and mixed case in the hex input.
What are GB2312, GBK, GB18030 and Big5?

Before UTF-8, Chinese text was stored in regional byte standards: GB2312 (1980, the 6,763 common hanzi), extended by GBK (1995), extended again by GB18030 — each a superset of the last — while Big5 is the traditional-character standard from Taiwan and Hong Kong. One hanzi is nearly always two bytes in these sets, which is why the hex preview reads in byte pairs.

The sets overlap but are not equal: a character that exists in one may be missing from another, so converting between charsets can lose characters — this page lists every character it cannot map (3F) instead of silently substituting one.