Long runs, from one source
Several long passages return the same page or closely related pages. That pattern deserves careful review, especially when the wording is not quoted or attributed.
Paste one document and search its passages against indexed public web pages. See the matching sources and the exact text worth reviewing.
No word limit · files up to 20 MB are read here · extracted passages are sent securely for public-web search
Searching public web pages for matching passages…
This can take a little while because each passage is checked separately.A match is not proof of plagiarism, and no match is not proof of originality. Check quotations, citations, dates and source context before reaching a conclusion.
No percentage was guessed from an incomplete search.
There is one document input. A file is converted to text in your browser, then query-sized passages go to the search endpoint and are checked against indexed public pages.
Paste a document of any length, or choose a PDF, Word, RTF, Markdown or text file. The browser extracts readable text before the online check begins.
The text is divided into meaningful exact-phrase queries and a server-side search provider looks each one up. A long document is sent in parts, because one request can only issue so many searches.
The result shows a weighted similarity percentage, every passage checked, source titles and links, and the original document with passages that returned sources highlighted.
A similarity figure on its own says very little. Correct quotations, standard wording and copied paragraphs can return the same source links. The result therefore keeps the passages and their context beside the percentage.
Several long passages return the same page or closely related pages. That pattern deserves careful review, especially when the wording is not quoted or attributed.
Quotations, references, legal language, definitions and standard technical phrases often return sources. They are real overlaps but are not misconduct by themselves.
None of the exact passages returned a public page from the configured search index. Private repositories, paywalled archives and unindexed pages remain outside that result.
Almost always a document and a suspicion about where part of it came from. These are the cases that arrive, and every one of them is a comparison against sources somebody already has.
An essay and the three articles on the reading list it was supposedly written from.
Two students, one paragraph, and a conversation nobody wants to have.
Work submitted again a year later, to a marker who kept the original.
A filed piece and the published article it reads uncomfortably close to.
A bid that reuses a competitor clause for clause.
A submission and the file a freelancer sent the week before.
Copy lifted wholesale from a competitor, down to the typos.
A methods section that appeared first in somebody else preprint.
A dissertation that says the same thing twice, three chapters apart.
An application that reuses a funded proposal without saying so.
Two reports from the same office, one clearly built from the other.
A draft checked against its own sources before anybody else sees it.
The figure is the share of checked words that sit inside a passage for which the web search returned at least one indexed public source. A passage returned by three pages still counts once, so the overall figure cannot exceed one hundred per cent.
It is a search-coverage measurement, not a judgement about authorship. The marker at thirty per cent is a reading aid rather than a decision line. A correctly cited paper can cross it, while rewritten copying can stay below it.
| Share | Band | What it means | What to do next |
|---|---|---|---|
| 0 – 5 | Nothing matched | No checked passage returned a source from the configured public-web index. | Do not call it cleared. The source may be private, unindexed, paywalled, translated or paraphrased. |
| 5 – 15 | Incidental overlap | A small number of passages returned indexed pages, often because of quotations or shared phrasing. | Open the sources and confirm that quoted or borrowed wording is attributed correctly. |
| 15 – 30 | Worth a look | Several passages have online matches, so their location and context start to matter. | Check whether each match is quoted, cited, common language or presented as original work. |
| 30 – 60 | Substantial overlap | A large share of the checked words belongs to passages that returned public sources. | Review the leading source links, dates and highlighted blocks. Do not decide from the number alone. |
| 60 – 100 | Largely the same text | Most checked passages returned one or more indexed sources. | Establish publication order and attribution before treating the overlap as misuse. |
The bands are fixed and published, so repeated checks can be read consistently even when the underlying index changes.
A high figure is not an accusation and a low one is not a clearance. The result covers only the public pages returned for the exact passages searched at that moment.
The document is normalised and divided into passages that usually contain about twenty-six words. The split prefers sentence endings and never creates a query below eight words. Every word remains covered, so a short tail at the end cannot silently disappear from the score.
Each passage is sent as an exact quoted query through the Original or AI server endpoint. The endpoint uses a server-held search credential, accepts only same-origin JSON requests and returns no cached response. The credential is never placed in page code or local storage. A long document is sent in parts, because one request can only issue so many searches before a serverless function runs out of room.
If the provider returns public pages, that passage is marked and the page titles, addresses and snippets are shown. The overall percentage is weighted by words rather than by sentence count, so a four-word fragment cannot influence the score as much as a full paragraph.
This does use the network. Files are parsed locally, but their extracted text is divided into passages and sent through our Cloudflare-hosted endpoint to the configured search provider. We do not add that text to a submissions archive. The provider and ordinary hosting logs remain subject to their own policies.
Distinctive verbatim passages on indexed public pages are the strongest case for this method. Exact quoted search removes much of the topic-level noise that would appear if ordinary keyword results were treated as copying.
Search coverage is never complete. Indexes omit private repositories, some paywalled pages, recently published text and pages blocked from crawling. Exact phrases also miss real paraphrase. A zero result means the provider returned no page for these queries; it does not mean the writing has been certified as original.
A flagged passage may be a quotation, reference entry, shared prompt, legal formula or standard language in a technical field. Those are genuine search matches but are not automatically plagiarism, which is why the output keeps every source link available for inspection.
A missed passage may come from a private or unindexed source, may have been translated, or may be paraphrased too heavily for an exact query. Search outages and exhausted quota are handled as errors; they are never converted into an invented unique result.
Read the highlighted passages, not the percentage. Where the matches fall, how long they run, and whether they are attributed will tell you everything the number cannot.
This checker is useful for finding likely public sources and opening them immediately. A school or publisher may also have licensed journals and a private archive of earlier submissions. Those collections can find material that a public search index cannot, so the two results can differ without either one being fabricated.
The useful part of an online plagiarism checker result is the evidence behind it. Original or AI shows the searched passages, returned pages, per-source coverage and highlighted document together, with provider failures kept separate from originality findings.
The document view marks each query-sized passage that returned at least one public page, so the percentage can be traced back to visible text.
Returned page titles, addresses and search snippets are listed by coverage. Open the original page to decide whether the wording is quoted, licensed, common or misused.
Paste text or choose PDF, Word .docx, RTF, Markdown and plain text. File extraction happens in the browser, and a document of any length is split into parts for searching.
A document past the size one request can carry is split, searched in parts and merged back into a single result, with every source counted once.
Original or AI does not add checked documents to a private corpus or compare future visitors against them. Extracted passages do travel to the search service for this requested check.
Only a compact result summary is kept in your browser history. The document text and source snippets are not written into that saved result.
The privacy boundary is different from the media detectors. A chosen file is opened and converted to text in your browser, so its binary contents are not uploaded. When you press the check button, the extracted text is divided into passages and sent to our Cloudflare Pages Function, which forwards exact queries to the configured search provider. Original or AI does not add the document to a submissions archive, but this is network processing and should not be described as fully local.
A source search is useful, but copied work also has changes a reader notices before any tool runs. Those changes can guide the source links and highlighted passages you inspect most carefully.
Every one of these has an innocent explanation, most often a document written over several weeks. Use them to guide review, not to reach a conclusion about the writer.
A public-web checker and an institutional system search different collections. The first can show open source links immediately; the second may also cover licensed journals and a private submission archive.
| Capability | Original or AI | Institutional checker |
|---|---|---|
| File parsing happens in the browser | Yes Extracted text is searched online | Partly Depends on the service |
| No Original or AI submissions archive | Yes Text is processed for the requested search | Partly Institutional archives may retain submissions |
| Clear online word limit | Yes No word limit | No Usually a few hundred words |
| Matched passages shown in place | Yes Highlighted in the text | Partly A percentage and links |
| Coverage attributed per source | Yes | Partly |
| Counts a shared passage once | Yes Cannot exceed 100% | No Shares often sum past 100% |
| Provider failures shown as errors | Yes Never converted into a unique score | Partly Behaviour varies |
| Reads PDF and Word before transmission | Yes | Partly Often parsed on the server |
| Works without an account | Yes | Partly Often after a sign-up wall |
| Searches indexed public web pages | Yes Exact quoted passage queries | Partly Coverage depends on provider |
| Searches journal and archive databases | No Institutional services only | Yes Available to licensed institutional services |
| Catches paraphrase | No Word matching cannot | No Sometimes implied |
The institutional column describes common licensed systems, not every product. Database coverage, retention and result logic vary, so consult the policy of the specific service used by your school, publisher or employer.
Paste one document and review exact passages against indexed public web pages. Source links, passage results and the highlighted document stay visible together.
Check for plagiarismNo sign-up. No word limit. No Original or AI submissions archive.