PDF to Markdown Converter — Clean Markdown, Even from a Scanned PDF
How do you convert PDF to Markdown, including scanned PDFs?
To convert PDF to Markdown, a tool has to work out the document's structure (which line is a heading, which lines form a list, where a table's rows and columns sit) and write that structure in Markdown syntax. Most PDF to Markdown converters get there by reading the text layer stored inside the file. That works for a digital PDF, but a scanned PDF is only a picture of each page, so a text-layer tool returns an empty file or a few stray characters. Linnk reads the page image itself. AI models (Linnk uses ChatGPT, Claude and Gemini models) look at each page, rebuild it as structured text and write the Markdown file from that, which is why scans, phone photos and digital PDFs all convert the same way. Headings, lists and tables come through, the tables as Markdown tables. Numbers, IDs, dates and amounts are copied digit for digit; a digit that can't be read is marked with a question mark and an unreadable passage is marked [illegible], rather than guessed. The document stays in its own language by default. The same run also gives you plain text and a searchable PDF, and a Word file can be built on request. A clean, straight, well-lit page usually converts best; a faint or curled scan may need a quick check.
Last updated: October 2026
How to convert PDF to Markdown
From a PDF, scanned or digital, to a .md file in three steps.
Upload the PDF or photo
Drop a PDF, scanned or digital, or a JPG, PNG, WEBP or HEIC photo of a page. A photo is read as a one-page scan. Markdown is already selected, and the output language is set to Keep original.
Check the preview
Conversion usually takes under a minute per page, often about half a minute on a typical scan, and the page shows an estimate while it works. Copying the recognised text is free. Look out for a question mark or [illegible], which flag the spots the page didn't show clearly.
Download the Markdown file
Sign in with a free account to download the .md file. Plain text and a searchable PDF come from the same run, and you can ask for a Word file, which takes a few extra minutes to build.
What this PDF to Markdown converter does
Built for text you plan to reuse: in prompts, notes, docs and sites.
Scans convert, not just digital PDFs
Text-layer converters need text stored inside the PDF. A scanned contract, a photocopied chapter or a photo of a printed handout has none, so those tools come back empty. Linnk reads the page image, so scans and photos go through the same way digital files do.
Headings, lists and tables kept as structure
Section titles become Markdown headings, bullet and numbered points become lists, and tables become Markdown tables you can paste into a README or a note. The structure is rebuilt from what the page shows, not from the order the PDF happened to store its text in.
Numbers copied, not guessed
Version numbers, dates, amounts and IDs are copied digit for digit. A digit too faint to read is marked with a question mark and an unreadable passage is marked [illegible], so neither a reader nor a model downstream gets fed a confident guess.
Markdown that suits LLMs and RAG
Markdown is compact, and its headings give a chunker natural places to split, which is why it's a common input for LLM prompts and retrieval pipelines. Do the PDF to Markdown step once, then paste the result into a chat, split it by heading or index it.
Keep the language, or translate it
Keep the original language, or translate in the same run. A Japanese manual stays in Japanese by default, and Arabic, Hebrew, Chinese and Korean documents are read too, so the Markdown matches the source you'll quote from.
More than one format per upload
The same conversion gives you plain text and a searchable PDF rebuilt as real text, and a Word file can be built on request. If the document has tables, a spreadsheet and CSV files are available from the same run as well.
PDF to Markdown conversion facts
What to expect before you upload a PDF, a scan or a photo.
Input
- Scanned PDFs
- Read from the page image
- Digital PDFs
- Read the same way
- Photos
- JPG, PNG, WEBP, HEIC, HEIF
- Each photo
- Read as a one-page scan
Output
- Markdown (.md)
- Headings, lists and text
- Tables
- Written as Markdown tables
- Plain text, searchable PDF
- From the same run
- Word (DOCX)
- Built on request
Accuracy
- Numbers and dates
- Copied digit for digit
- Unreadable digits
- Marked with a question mark
- Unreadable passages
- Marked [illegible]
- Tables across pages
- Not joined; one table per page
Access
- Preview
- Free, no account
- Copy text
- Always free
- Download
- Free account
- Longer documents
- Paid plans, see linnk.ai/pricing
PDF to Markdown options compared
How Linnk compares with the usual ways of getting Markdown out of a PDF.
| Feature | Linnk PDF to Markdown | Open-source PDF-to-Markdown libraries | Copy and paste | Online PDF converters |
|---|---|---|---|---|
| Setup | Upload in the browser, nothing to install | Install and run locally; easy to script for batches | None | Upload in the browser |
| Scanned PDFs and photos | Read from the page image by AI models | Some offer an OCR mode; results depend on setup | Nothing to copy from a scan | Varies; many read only the text layer |
| Headings and lists | Rebuilt as Markdown headings and lists | Varies by library | Pasted as plain lines | Varies |
| Tables | Markdown tables | Varies by library | Usually pasted as loose text | Varies |
| Unreadable digits and passages | Marked with a question mark or [illegible] | Varies | Not applicable; text is copied as stored | Varies |
| Where your file goes | Uploaded; converted files deleted after 72 hours | Stays on your machine | Stays on your machine | Uploaded or processed in the browser, depending on the site |
| Other outputs from the same upload | Plain text, searchable PDF; Word on request | Depends on the library | None | Sometimes other formats |
Comparison as of September 2026. Features of other tools vary and change often; check each one for its current options.
Who converts PDF to Markdown
People who want a document's text in a form they can edit, search and pass to other tools.
Developers building LLM and RAG apps
Turn scanned manuals and vendor PDFs into Markdown before chunking and embedding them, so a spec table arrives as rows and columns instead of a jumble of words.
Researchers and students
Convert papers, book chapters and photocopied readings into Markdown notes for Obsidian or Logseq, with the section headings in place for linking and outlining.
Technical writers
Move a legacy PDF manual into a Git-based docs repo or a static site generator such as MkDocs or Hugo, then edit and review it like any other page.
Writers and editors
Get an old article, a manuscript or a scanned typescript back as editable Markdown for a writing app, without retyping it page by page.
Teams keeping a wiki
Import a scanned policy or a supplier's PDF handbook into Notion or a team wiki as Markdown, so the headings and tables stay editable instead of sitting in an attachment.
Analysts working across languages
Convert a Chinese, Arabic or German report to Markdown in its own language, so quotes and figures match the source, then search it or compare versions with a diff.
PDF to Markdown FAQ
How do I convert PDF to Markdown?
Can I convert a scanned PDF to Markdown?
Why use Linnk for PDF to Markdown instead of an open-source library?
What does the Markdown file contain?
Is the Markdown suitable for ChatGPT, Claude or a RAG pipeline?
Can I convert an image to Markdown?
Which files can I upload?
Which languages does it read?
How long does it take?
Is the PDF to Markdown converter free?
What happens to my files after the conversion?
Convert your PDF to Markdown
Drop a scan, a photo or a digital PDF and check the preview before you download.