Meelu

Free similarity checker

You have two pieces of writing and you need to know how much of one is in the other. Paste a text into each box — or a web address, and the page is fetched for you — and get a percentage, the sentences that match, and the passages the two share word for word. Nothing is uploaded. No sign-up.

0 / 5,000
0 / 5,000

Free, no sign-up. Pasted text is compared inside your browser and is never uploaded — only a URL is sent to our server, and only so the page can be fetched. Up to 5,000 words a side.

  • No account
  • Nothing uploaded
  • Both directions, not an average
  • Matched sentences listed
  • CSV export

What this document similarity checker reports

Both directions, not one average

How much of A appears in B is a different number from how much of B appears in A, and the gap is the whole story when one text is longer. Both are reported in plain words.

Shows you the sentences that match

Every sentence is paired with its closest match on the other side, so you can read the overlap for yourself instead of arguing with a percentage.

Longest shared passages

The longest runs of identical words the two texts share, with a word count on each. A forty-word identical run is more convincing evidence than any percentage.

What is only in yours

The sentences with nothing close to them on the other side — the part that is genuinely original, listed so you can see what survived.

Pasted text never leaves the browser

The comparison is JavaScript running on your machine. Nothing is uploaded, nothing is stored and nothing is added to anybody's archive of submitted work.

CSV export, no sign-up

Every matched pair with its score, ready for Excel, Sheets or Numbers — so an editor can work down the list without opening the tool.

How to check similarity between two texts

  1. Step 1

    Put a text in each box

    Paste the writing itself, or paste a URL and the page is fetched and reduced to its readable text. One side can be text and the other a URL.

  2. Step 2

    Press compare

    The whole comparison happens on your own machine, in under a second. Nothing is uploaded and nothing is kept.

  3. Step 3

    Read the matches, not just the number

    Shaded sentences in the side-by-side view jump to their partner when clicked. Work down the matched pairs, then export the list to CSV.

What each number means

Most free similarity checkers give you one percentage and no way to argue with it. These are the five figures this one reports and what each is actually measuring.

Overall similarity
The higher of the two directions below. Under 30% the two texts are original with respect to each other; 30–60% is partly duplicated and worth reading through; over 60% means one document is largely the other one.
How much of A appears in B
Containment, one way round. If this is high and the other is low, a short text has been copied into a long one — which is the case a single averaged score hides completely.
Word-for-word overlap
How much of this text appears in the other completely untouched. This is the number to quote when the question is whether something was copied rather than written.
Matched-sentence coverage
How much of the text sits in a sentence that has a close partner on the other side. Higher than the word-for-word figure when copying has been lightly tidied up.
Longest shared passage
The longest unbroken run of identical words the two share. Short runs are ignored, because a handful of words in common is just how English works.

Duplicate content and plagiarism are not the same problem

The same percentage answers two different questions depending on who is asking, and the searches that land on a similarity checker come from both.

Duplicate content is an SEO problem

Two of your own pages saying the same thing. Nobody has done anything wrong; the pages are just competing. Google picks one and filters the other out of the results, so the duplicate quietly stops earning traffic. The fix is a canonical tag, a redirect or a merge — not a rewrite. Most of it is created by templates and URL parameters rather than by writers.

Plagiarism is an authorship problem

Somebody else’s words presented as yours. Here the percentage matters far less than the passages: one uncited paragraph is serious at 4% overall, and a page of properly quoted and referenced material is fine at 40%. Read the matched sentences and the longest shared runs, and ignore the headline.

This tool measures the overlap. Which of the two problems you have is your call, and it changes what you do about it.

If you came here looking for Turnitin

A lot of people searching for a Turnitin similarity checker — or a Turnitin similarity checker free of charge — want the Turnitin similarity report without paying for Turnitin. That is not something any free tool can give you, and the sites promising it are either selling something else or keeping your paper. Here is the honest version.

  • Turnitin’s number comes from its archive. Your paper is matched against tens of millions of student submissions, licensed publisher content and a web crawl. The archive is the product. Nobody else has it, so nobody else can reproduce the number.
  • This compares two documents you supply. If you already know the source you worked from — lecture notes, a shared draft, an earlier version of your own assignment, a classmate’s essay you both built from the same outline — this will show you exactly how much of it survived into your draft, sentence by sentence.
  • When that is enough, and when it is not. Enough: checking your own rewrite actually reads as a rewrite, or settling an argument about two documents in front of you. Not enough: predicting a submission score, or finding out whether a passage exists somewhere on the internet. For the second one you need a service with an index, and you should expect to pay for it.
  • Your institution already has the report. Most universities let you see the similarity report on a draft submission before the final one. That report is the only one that counts, and it is free to you. If you searched for a similarity checker Turnitin would agree with, that draft submission is it — no third-party number, this one included, predicts the score.

If you came here looking for an AI similarity checker

“AI similarity checker” gets searched for two different things, and the free AI similarity checker results that come back rarely say which one they are. This is an AI similarity checker free of any model at all, so it is worth being specific.

If you mean a checker that uses AI: this one does not, on purpose. For measuring how much two documents share, the old methods win on the things that matter here — the same texts always give the same number, it runs instantly on your own machine, and there is no company to send your draft to. What you give up is spotting a passage rewritten in different words. That is the trade, said plainly rather than hidden behind an AI-powered badge.
If you mean a checker for AI-written text: that is a different question — not how similar two texts are, but whether one of them came out of a model. Use the free AI content detector instead. It is the wrong tool for duplicate content and this is the wrong tool for detecting a model, and running the pair of them answers both.

What it does not do

  • No database, no web search. It compares the two documents you give it. It cannot tell you whether a sentence appears anywhere else, because there is no index behind it to look in.
  • It compares wording, not meaning. A passage genuinely rewritten in different words scores near zero, so a careful rewrite can still slip past. No checker that reads wording can honestly claim otherwise.
  • Not a code similarity checker. It finds identical runs in source files, but a copy with everything renamed slips past. Use a purpose-built one for that.
  • Not a face similarity checker. This compares writing. If you are after a face similarity checker online free, you want an image tool — there is no upload here and nothing that looks at photographs.
  • No file upload. Word documents, PDFs and Google Docs need their text copied in. Nothing is parsed, and nothing is kept.
  • Built for English. Other languages written in the same alphabet mostly work. Languages that do not put spaces between words will give wrong numbers.
  • 5,000 words a side. Longer texts are cut and the report tells you it happened. Compare them in sections rather than trusting a truncated score.

Frequently asked questions

What is text similarity?

Text similarity is a measure of how much two pieces of writing have in common, usually given as a percentage. Lexical similarity counts shared words and phrases, so it catches copying. Semantic similarity compares meaning, so it catches the same idea written differently. Most checkers report the lexical figure, because it is the one you can point at and verify.

What is duplicate content?

Duplicate content is the same or near-identical text appearing at more than one address, either within a site or across sites. It happens most often by accident: printer-friendly pages, URL parameters, filtered category listings, product descriptions copied from a supplier, or the same article republished under two paths. It is a housekeeping problem rather than a moral one, but it costs traffic all the same.

How much similarity is acceptable?

There is no universal figure, and any tool that gives you one is guessing. Academic institutions typically accept somewhere between 10% and 25% including quotations and references, and set their own bar. For two pages on the same website, anything above about 60% means one of them should be merged or canonicalised. Read the matched passages rather than the headline number: 40% made of properly cited quotes is fine, 10% made of one lifted paragraph is not.

What is semantic similarity?

Semantic similarity measures whether two texts mean the same thing, regardless of the words used. "Cut costs" and "reduced expenditure" share no wording but carry the same claim, and a semantic comparison scores that as a match. It is normally computed by turning each text into a vector with a language model and measuring the angle between them. Lexical methods cannot do this, which is why heavy paraphrasing slips past them.

How is text similarity calculated?

The common method breaks both texts into overlapping runs of words, called shingles, and counts how many runs they share. A second pass pairs each sentence with its closest match on the other side and scores the pair on shared vocabulary, usually with TF-IDF weighting and cosine similarity so that common words count for less. The two numbers answer different questions: verbatim overlap and reworded overlap.

Can paraphrasing be detected?

Light paraphrasing, yes. Swapping a few words, reordering a clause or changing the tense leaves enough shared vocabulary for a sentence to still register as a match. A genuine rewrite in different words will not be caught by any method that compares wording alone, because there is nothing left to compare. Detecting that needs a semantic model, and even then the result is a judgement rather than proof.

What is the difference between plagiarism and duplicate content?

Plagiarism is using someone else's work without credit, which is an academic and ethical question. Duplicate content is the same text existing at two addresses, which is a technical and SEO question. Self-duplication is common and is not plagiarism. A checker can measure overlap; only a person can decide whether that overlap was dishonest.

Does duplicate content hurt SEO?

Not as a penalty. Search engines filter rather than punish: when two pages say the same thing, one is chosen for the results and the other is dropped, so the duplicate quietly stops earning traffic. The real costs are crawl budget spent on pages that will never rank and link signal split between two URLs that should have been one. The fix is a canonical tag, a redirect or a merge.

Duplication is rarely the only thing wrong

If two of your own pages came back heavily similar, the next questions are which one Google has picked, whether the canonical tags agree with you and how many other pages are in the same state. The free SEO audit tool crawls the site and answers all three, including duplicate titles and descriptions across every page it reads. The broken link checker covers the other half of a content clean-up: the links that stopped working while nobody was looking.

All three are pieces of Meelu, a desktop app where an AI marketing agent runs your marketing on your own machine — no word cap, no rate limit, and nothing leaves the laptop. Join the waitlist.