Extract all text from a PDF as a plain text file
Drag & drop your file here
click to browse
Max file size: 20MB
pdfty's extract tool pulls every character of text out of a PDF and delivers it as a clean plain-text file. It beats select-all copy-paste, which routinely scrambles multi-column layouts, drops text boxes, and chokes on long documents โ extraction processes all 50 pages in one pass with reading order preserved.
Extraction runs on PyMuPDF, which reads the PDF's internal text structure directly and reconstructs natural reading order โ including two-column layouts, tables, headers, and footnotes. Full Unicode comes through intact: Cyrillic, CJK, Arabic, accented characters. The result is ideal for feeding documents into translation tools, search indexes, or AI assistants, for quoting from papers, and for word counts.
One important note: this works on PDFs that contain a text layer. Scanned documents are just pictures of text โ run them through our OCR tool (powered by Tesseract) first, then extract. Free for files up to 20 MB and 50 pages, no watermark, no signup. Files are permanently deleted within 1 hour.
Drag and drop or click to select. Up to 20 MB and 50 pages free.
PyMuPDF reads the text layer and reconstructs reading order. A second or two for most files.
Check the first pages of extracted text on screen.
One plain-text file, with page breaks marked.
Your PDF and the text file are deleted from our servers within 1 hour.