Convert PDF To and From Other Formats
Convert tools turn a PDF into another format, such as Word, Excel, PowerPoint, images, plain text, HTML, Markdown, CSV or XML, and turn other files into a PDF: Office documents, images, HTML, Markdown, web pages and eBooks.
A conversion rebuilds the content in the new format rather than copying it, so layout, fonts and tables can differ from the original. Each tool page says what it keeps and what it drops; check the result before you rely on it.
- PDF to WordPDF to Word converts a PDF that contains real text into an editable DOCX, DOC or ODT file with LibreOffice. It suits memos, reports and forms; a scan that is only pictures has no text to import.
- Images to PDFImages to PDF puts each uploaded PNG, JPEG, GIF or BMP image on its own page and returns one PDF, in upload order. Page size follows the image size.
- PDF to ImagesPDF to Images renders every page of a PDF as a picture with Ghostscript and delivers the files in a ZIP. The form defaults to PNG at 300 DPI; text becomes pixels and links do not survive.
- HTML to PDFHTML to PDF renders an uploaded .html or .htm file, or a ZIP that contains HTML, into a PDF with WeasyPrint. JavaScript does not run, so pages built by scripts come out incomplete.
- PDF to PowerPointPDF to Presentation imports a PDF into LibreOffice Impress and returns a PPTX, PPT or ODP file. Treat it as a draft: fonts may be substituted and objects approximated, so proof it before presenting.
- PDF to TextPDF to Text extracts the text that already exists in a PDF and returns it as a plain .txt file, or as RTF. It does not run OCR, so image-only scans come out empty.
- PDF to PDF/APDF to PDF/A rewrites a PDF with Ghostscript in PDF/A-1 or PDF/A-2 mode. It is a conversion, not a certificate: check the result in a PDF/A validator if archival compliance is required.
- PDF to MarkdownPDF to Markdown turns the existing text of a PDF into a .md file, with headings, paragraphs and lists where it can detect them. It does not OCR, so scanned pages come out empty.
- PDF to HTMLPDF to HTML converts a PDF with Poppler's pdftohtml and delivers a ZIP holding the HTML pages and their images. It is a layout dump rather than clean web markup, and it does not OCR.
- PDF to CSVPDF to CSV finds aligned table rows in a PDF's text and writes each detected table as a CSV file; several tables come back in a ZIP. It needs a text layer and returns nothing for prose or scans.
- PDF to ExcelPDF to Excel finds aligned table rows in a PDF's text and writes each table to its own sheet of an .xlsx workbook. It needs a text layer, and returns nothing for prose or scanned pages.
- Markdown to PDFMarkdown to PDF renders a CommonMark .md file to a PDF with WeasyPrint, using a font stack that covers Chinese, Japanese and Korean text. The file must be UTF-8.
- eBook to PDFeBook to PDF converts an EPUB, MOBI, AZW3 or FB2 file to PDF with Calibre's ebook-convert. There are no paper-size or font options, and DRM-protected files fail.
- Office to PDFFile to PDF converts Word, Excel, PowerPoint, OpenDocument, RTF and plain-text files to PDF with LibreOffice. Layout follows LibreOffice's import, so it can differ from Microsoft Office.
- PDF to XMLPDF to XML imports a PDF through LibreOffice and returns an Office-flavored XML file. It is a best-effort text and layout export, not a copy of the PDF's structure or tags.
- URL to PDFURL to PDF fetches a public web address with WeasyPrint and saves the page as a PDF. JavaScript does not run, so pages that need scripts or a login come out empty or incomplete.