What this tool does
Compare shared local word sequences in two supplied texts. The tool lowercases text, removes much punctuation and matches overlapping word n-grams between the two inputs. Its similarity percentage derives from matching n-gram counts and is capped at 100; highlights show locally matched positions.
How to use Text Similarity Checker
- Prepare the input. Paste both texts you are authorized to compare.
- Run or configure the tool. Select Compare texts.
- Check and use the output. Review the highlighted phrases and context instead of treating the percentage as a verdict.
How this tool works
The tool lowercases text, removes much punctuation and matches overlapping word n-grams between the two inputs. Its similarity percentage derives from matching n-gram counts and is capped at 100; highlights show locally matched positions.
Worked example
Input: Text A and Text B: The blue bird sings softly.
Output: Similarity: 100% for these identical supplied texts.
Limits, assumptions and interpretation
No web search, journal database, institutional repository or plagiarism corpus is queried.
- Overlapping n-grams can overcount matches; the percentage is a local similarity heuristic, not a measured percentage of copied writing.
- Short texts, non-ASCII languages, paraphrases and common phrases can produce misleading results.
- It cannot decide plagiarism, permission, attribution or originality; investigate similarities in context.
Supported inputs and limits
- Text A and Text B: The blue bird sings softly.
- No web search, journal database, institutional repository or plagiarism corpus is queried.
- Overlapping n-grams can overcount matches; the percentage is a local similarity heuristic, not a measured percentage of copied writing.
- Short texts, non-ASCII languages, paraphrases and common phrases can produce misleading results.
- It cannot decide plagiarism, permission, attribution or originality; investigate similarities in context.
Frequently asked questions
What does this tool actually do?
The tool lowercases text, removes much punctuation and matches overlapping word n-grams between the two inputs. Its similarity percentage derives from matching n-gram counts and is capped at 100; highlights show locally matched positions.
What should I check before using the result?
No web search, journal database, institutional repository or plagiarism corpus is queried. Overlapping n-grams can overcount matches; the percentage is a local similarity heuristic, not a measured percentage of copied writing. Short texts, non-ASCII languages, paraphrases and common phrases can produce misleading results. It cannot decide plagiarism, permission, attribution or originality; investigate similarities in context.
Is information sent to a server?
Tool inputs are processed in this browser. This product does not use analytics or send your input to an external API. Clicking an external website link still visits that website.