Text Encoding Detector
Check text or hexadecimal bytes for recognized BOMs and valid UTF-8, while reporting ambiguity honestly.
Text encoding detector
A QUICK WALKTHROUGH
How to use this tool
- Choose text or hex bytes.
- Enter the value and inspect it locally.
- Read the evidence-based result and its limitations.
Evidence-based detection
Recognized UTF-8, UTF-16LE, and UTF-16BE BOMs are reported. Strict UTF-8 validity is reported when no BOM settles the question.
Ambiguity is explicit
A no-BOM byte sequence can be valid in several encodings. This tool does not claim certainty or guess legacy encodings from text patterns.
Bounded local inspection
Input is limited to 100,000 characters and is processed in the browser without upload.
GOOD TO KNOW
Common questions
Does it identify every encoding?
No. It reports only limited BOM and strict UTF-8 evidence.
What does no BOM mean?
It means the bytes may be ambiguous; valid UTF-8 is evidence, not proof of the original encoding.
Does it guess Windows or regional encodings?
No.