Add your files
Drop one PDF or two hundred. Batch conversion is not a paid feature here.
The words in a PDF need to be fed into a script, a search index, a spreadsheet or an LLM without any of the formatting.
Only characters come out. Fonts, positions, images, tables and page furniture are all discarded, leaving a plain text stream in the reading order the extractor inferred.
Reading order is inferred, not stored, so two-column academic papers commonly come out with lines from both columns interleaved. Headers, footers and page numbers land inline in the middle of the text too. A scan with no text layer is rejected with a message saying exactly that, rather than returning an empty file.
There is no text layer to extract, so OCR has to run. Set the OCR mode to auto and it falls back automatically.
PDF stores glyph positions, not spaces. Extraction has to infer word boundaries from gaps, and tight kerning or unusual fonts fool it.
Yes, give a page range.
Drop one PDF or two hundred. Batch conversion is not a paid feature here.
Plain Text is lossless, so nothing is thrown away in the conversion.
Instant, because the file never went anywhere. Nothing expires, since nothing was stored.