Practice

Checking PDF and DOCX without copying the text

You can upload .docx, .pdf, .txt and .md up to ten megabytes — the text is extracted for you. With a scanned PDF that contains only an image of the page this is not possible and you need a different version of the document.

Why not to copy by hand

Copying out of a PDF breaks the line wrapping and loses the paragraphs. Detection also works with the construction of the text, so broken wrapping shifts the result.

Across thirty papers it is also half an hour of extra mechanical work.

Scanned PDFs

A scan is an image. Technically there is no text in it, so nothing can be extracted and the check does not run.

The fix is to ask for the original document. For work produced on a computer, one always exists.

What to watch for before uploading

Take out quotations and appendices. Someone else's text only adds noise to the result.

For long work, upload chapter by chapter rather than the whole thing at once.

FAQ

Which formats are supported?

Text files .txt and .md, documents .docx and .pdf. Other extensions are rejected.

How large a file goes through?

Up to ten megabytes. Larger work is better split anyway.

Does the uploaded file stay stored?

The extracted text is saved to your history, not the file itself.

Try it on your own text

Paste a text or upload a file and look at the score and at the specific sentences that came out suspicious.

Run a detection

More articles