We use cookies.This website uses essential cookies to operate core features. With your consent, we also use analytics cookies to understand traffic and improve the service. For more details, see our .
CJK to Unicode Converter
If this tool helped you, you can buy us a coffee ☕
A bidirectional converter for CJK characters and Unicode code points. Supports multiple input formats to solve encoding issues in programming and text processing.
Supports single character, U+XXXX, 0xXXXX, decimal, and backslash-u hexadecimal forms
Enter a character or Unicode code point
Online JSON to TOML Converter
Quickly convert JSON data to TOML format. A free online tool perfect for migrating configuration files.

Unicode Encoder & Decoder
A tool for bidirectional conversion between text and Unicode escape sequences (such as \uXXXX format).
JSON to Protobuf Converter
Free online JSON to Protobuf converter. Instantly generate Protocol Buffers message definitions from JSON data.
When processing multilingual text, encoding issues with Chinese, Japanese, and Korean (CJK) characters often lead to garbled text or display anomalies. This tool provides precise bidirectional conversion between CJK characters and code points using the Unicode standard. A single code point (e.g., U+4F60) corresponds to a basic unit of an ideograph, supporting the conversion of character sets like Hanzi, Kana, and Hangul.
Q: What are the Unicode code points for "你好"?
U+4F60 U+597D. These are the standard code points for the two Chinese characters "你" and "好" in the Unicode Basic Multilingual Plane (BMP).
Q: Does the tool support rare Chinese characters in the extension blocks?
It supports all CJK Unified Ideographs within the BMP, including Extension A characters. However, the validity of code points for some rarely used characters in Extensions B-F may need to be verified.
We recommend processing no more than 500 characters at a time. Non-Unicode encodings (such as GB2312) must be converted to UTF-8 first. The conversion results do not include character property metadata.
When handling CJK text in development, it is highly recommended to use the standard U+XXXX format. For example, the Japanese word "日本語" can be converted to U+65E5 U+672C U+8A9E. This format is universal and highly readable across most programming languages. Note that UTF-8 encoding and Unicode code points are different concepts—the former is a byte sequence, while the latter is an abstract character number.