PDF text to safe HTML
Escapes extracted PDF text into a page-grouped HTML document. It does not recover design, images, link structure, or OCR.
β οΈ Scanned PDFs require OCR. The output is extracted text, not a reconstruction of the source document.
