Code tool

Text Encoding Detector

Check text or hexadecimal bytes for recognized BOMs and valid UTF-8, while reporting ambiguity honestly.

In-browser processingNo account requiredPrivacy details ↗

Text encoding detector

This modest detector only checks BOMs and strict UTF-8. It does not guess legacy encodings.

Enter text or bytes to inspect.

A QUICK WALKTHROUGH

How to use this tool

  1. Choose text or hex bytes.
  2. Enter the value and inspect it locally.
  3. Read the evidence-based result and its limitations.

Evidence-based detection

Recognized UTF-8, UTF-16LE, and UTF-16BE BOMs are reported. Strict UTF-8 validity is reported when no BOM settles the question.

Ambiguity is explicit

A no-BOM byte sequence can be valid in several encodings. This tool does not claim certainty or guess legacy encodings from text patterns.

Bounded local inspection

Input is limited to 100,000 characters and is processed in the browser without upload.

GOOD TO KNOW

Common questions

Does it identify every encoding?

No. It reports only limited BOM and strict UTF-8 evidence.

What does no BOM mean?

It means the bytes may be ambiguous; valid UTF-8 is evidence, not proof of the original encoding.

Does it guess Windows or regional encodings?

No.