Extract Long Words
Paste text, choose a minimum length, and list longer words locally.
A QUICK WALKTHROUGH
How to use this tool
- Paste the text to scan.
- Choose a minimum length from 1 to 1,000 Unicode code points per word.
- Extract unique matches, then copy or download the result.
Unicode word and length rules
A word is a word-like segment from Intl.Segmenter when available, containing a Unicode letter or number. Length counts Unicode code points: a supplementary character counts as one and a combining mark counts separately. A conservative Unicode letter/number fallback is used when segmentation is unavailable.
Unique, stable results
Matches are deduplicated by exact text, including case and accents. The first spelling and occurrence are retained. Results are sorted longest first; equal-length words keep their first-occurrence order. Processing stays in the browser and accepts up to 100,000 UTF-16 code units.
GOOD TO KNOW
Common questions
How is length counted?
Length uses Unicode code points, not UTF-16 code units or visible grapheme clusters. A supplementary character counts as one; a base character and combining mark count as two.
Are duplicates removed?
Yes. Exact text matches appear once in first-occurrence order. Case and accents remain significant.
What order are results in?
Longer words come first. Equal-length words retain the order of their first matching occurrence.