Text tool

Duplicate Content Checker

Find shared text fragments in two pasted texts locally. Each text: at most 64 KiB UTF-8 and 40,000 normalized Unicode characters. Show at most 100 fragments, 240 characters each.

In-browser processingNo account requiredPrivacy details ↗

Find shared text fragments in two pasted texts locally. Each text: at most 64 KiB UTF-8 and 40,000 normalized Unicode characters. Show at most 100 fragments, 240 characters each.

Normalize with Unicode NFKC, lowercase, whitespace collapse and trimming. Compare exact sliding windows of 12–120 characters (default 24); merge adjacent matches only when their positions align in both texts. Repeated windows use the first position in the second text. Positions are 1-based normalized Unicode character offsets, not source offsets. This heuristic can miss paraphrases or repeated occurrences and match common phrases. It does not search websites or determine plagiarism, authorship or AI use.

Processed only in your browser; input is not uploaded or saved.

A QUICK WALKTHROUGH

How to use this tool

  1. Paste input, choose options, then analyze and review the report.

Scope and limits

Normalize with Unicode NFKC, lowercase, whitespace collapse and trimming. Compare exact sliding windows of 12–120 characters (default 24); merge adjacent matches only when their positions align in both texts. Repeated windows use the first position in the second text. Positions are 1-based normalized Unicode character offsets, not source offsets. This heuristic can miss paraphrases or repeated occurrences and match common phrases. It does not search websites or determine plagiarism, authorship or AI use.

Processed only in your browser; input is not uploaded or saved.

Find shared text fragments in two pasted texts locally. Each text: at most 64 KiB UTF-8 and 40,000 normalized Unicode characters. Show at most 100 fragments, 240 characters each.

GOOD TO KNOW

Common questions

Scope and limits

Normalize with Unicode NFKC, lowercase, whitespace collapse and trimming. Compare exact sliding windows of 12–120 characters (default 24); merge adjacent matches only when their positions align in both texts. Repeated windows use the first position in the second text. Positions are 1-based normalized Unicode character offsets, not source offsets. This heuristic can miss paraphrases or repeated occurrences and match common phrases. It does not search websites or determine plagiarism, authorship or AI use.