PDFGPT

Extract, convert, merge, split and compress PDFs.

PDFGPT handles the routine PDF operations: pulling text out, converting to plain text, merging several files into one, splitting a long document, and compressing scans. It distinguishes text-layer PDFs from scanned images and treats each correctly.

Workspace

The PDFGPT interface runs here. While it is being connected, the guidance below covers how to use it, what it handles well and where it falls short — the same information you would want before running a first job.

How to use PDFGPT

  1. 1Upload your PDF and check whether the tool reports a text layer or a scanned document.
  2. 2Choose the operation: extract text, convert, merge, split or compress.
  3. 3For merges, set the page order; for splits, set the page ranges.
  4. 4Run the operation and review the preview before downloading.
  5. 5Compare a page of the result against the original, especially after compression.

Practical examples

Extracting from a two-column research paper

Input

A 14-page digitally generated PDF with a text layer, two columns and footnotes.

Result

Exact characters with no recognition error, running headers flagged for removal, and a warning that column reading order should be checked at each section break.

Compressing a scanned archive

Input

A 180-page colour scan at 400 DPI, 240 MB.

Result

Converted to greyscale at 300 DPI, reduced to roughly 34 MB, with text still sharp at full zoom and the original retained as the master copy.

Supported formats

  • PDF (text-layer)
  • PDF (scanned images)
  • Plain text output
  • Encrypted PDFs once unlocked

Features

  • Direct text-layer extraction with no recognition error
  • Recognition for scanned pages
  • Merge with page-order control
  • Split by page range
  • Compression with colour-depth and resolution options
  • Page rotation correction

Limitations

  • Encrypted files must be unlocked with their password before processing.
  • Multi-column reading order can be wrong; check paragraph continuity.
  • Compression is not reversible, so keep an uncompressed master.
  • Table structure and precise layout are not preserved in plain-text output.
  • Very large files are more reliable processed in batches.

Privacy and data handling

Uploaded PDFs are processed to produce your result and are not kept as a permanent archive. Redact sensitive fields you do not need processed, and remember that a black rectangle drawn over selectable text does not remove it.

The full statement is in our privacy policy.

PDFGPT questions

Guides for PDFGPT

Related tools