Extract a PDF's Structure as XML, Online for Free
Get metadata, bookmarks, page dimensions, and text as structured XML — no upload, no signup, up to 100 MB.
Your file never leaves your deviceDrag & drop your PDF here
or Max 100 MB · Free · Processed entirely in your browserHow to Extract a PDF's Structure as XML
-
Upload your PDF Drag & drop or browse your PDF.
-
Confirm No settings needed — we extract the structure automatically.
-
Review and download Check the preview, then save your XML instantly.
Your PDF, as structured data
- Private by designYour PDF is processed on your device and never uploaded.
- InstantNo upload queue — your browser does the work locally.
- Free, no signupNo account, no credit card, no trial that expires.
- More than just textMetadata, bookmarks, and each page's size and rotation — not only the words.
- Preview before you downloadCheck the extracted XML right in your browser first.
Frequently Asked Questions
Is it free?
Yes. No signup required.
Is my file safe?
Yes — your PDF is processed entirely in your browser and never uploaded.
What's included in the XML?
Document metadata (title, author, subject, keywords, creator, producer, dates — each element included only when the PDF has it), the top-level bookmark list, and one page element per page with its width, height, rotation, and extracted text.
Is the output valid XML?
Yes — well-formed XML 1.0 with UTF-8 encoding. Special characters in titles, metadata, and text are properly escaped, so any XML parser can read the file.
Does it include images or formatting?
No. This extracts structure and text, not visual content — images, fonts, and layout aren't included. For a visual export, try one of our PDF to image or PDF to HTML tools instead.
Does it work on scanned PDFs?
Page metadata (size, rotation) will still be extracted, but scanned pages have no real text layer, so each page's text will be empty. Run an OCR tool first if you need text from a scan.
What if my PDF has no bookmarks?
The outline element is simply empty — most PDFs don't have bookmarks.
Need JSON instead?
Our PDF to JSON tool exports the exact same structure as JSON — same fields, same one-entry-per-page shape.
Max file size?
100 MB.