Extract Links from PDF
List every link in a PDF with the words you see and the address it actually opens — and a warning whenever those two disagree. Web and email addresses typed as plain text are found too, and no link is ever opened.
Extract PDF Links
Add your files. Recommended settings are ready, so you can process in one click.
PDF, Word, images, audio, video, archives and more
Up to 100 files · 250 MB · No account- 1Add files
- 2Choose an action
- 3Download or keep going
What happens to your file.
The PDF is opened in your browser and every page is read for link annotations. For each one, the address stored in the document is taken as-is, and the text items whose centres fall inside the link’s rectangle are collected to give you the words a reader actually sees. Those two are then compared: if the visible text names a domain that is not where the link goes, it is flagged. Every other judgement — the scheme, look-alike characters, a bare IP address, a username hidden in the address, a known shortener, a redirect carrying another URL — is made by reading the address itself. Process file does the same on our server and also searches each page’s text for web and email addresses that were typed without a link, which PDFs made without hyperlinks are full of; those are listed as not clickable, and an address that already has a link on the page is listed once.
Where it stops: No link is ever opened, so this cannot tell you whether a link still works or where a shortener finally lands; it reports what the document says, not what a server would answer. Typed addresses are read from the page text, so one split across two lines may be cut short, and a scanned PDF has no text to read at all: run OCR PDF on it first. A flagged link is not proof of anything bad, and an unflagged one is not a guarantee: this reads addresses, it does not check reputation.
What the warnings mean
| Warning | Why it matters |
|---|---|
| The text names a different site | The page says one thing and the link does another, which is how a phishing document works |
| Runs a script or opens a local file | A javascript:, data: or file: target is not a web page at all |
| Non-Latin characters in the address | Letters that look identical to Latin ones can imitate a familiar name |
| Points at a bare IP address | Legitimate organisations almost always use a domain name |
| A username in the address | Everything before the @ is ignored by the browser, so the real site is hidden |
| Passes through a redirect | The address carries another URL, so the real destination is elsewhere |
| A shortened link | The destination cannot be known without opening it |
Why the links are not clickable here
The list shows every target as plain text. That is deliberate: this tool exists for documents you do not yet trust, and a page that invited you to click the very links you came to check would defeat the point. Copy an address if you want it, once you have read where it goes.
What people use it for.
Everything you need, without a paywall.
Common questions.
How do I see where a link in a PDF really goes?
Add the PDF and every link is listed with its visible text next to the address actually stored in the document. Where those disagree, the row is flagged. Nothing is uploaded and no link is opened.
Does it find URLs that are not clickable?
Yes, with Process file. Web and email addresses typed into the text without a link attached are listed too, marked as not clickable, so a PDF made without hyperlinks still gives you its full list. The visual editor lists clickable links only.
Can I extract email addresses from a PDF?
Yes. Clickable mailto: links and email addresses written in the text are both listed, each with the page it is on. An address that appears more than once on a page is listed once.
Are internal links included?
Yes. A table-of-contents entry pointing to another page is listed with the page it goes to, alongside the external links.
Can it tell me whether a link is dangerous?
It tells you what the address says and where it disagrees with the text you would see. It cannot check a site’s reputation, and it never opens a link to find out. Treat it as a first read, not a verdict.
Can I check whether the links still work?
Not here — that would mean opening every link, which this tool deliberately does not do. Export the list as CSV and feed it to a link checker.
Is my PDF uploaded?
Not with the visual editor, which reads it in your browser. Process file sends it to our server, where it is written to a temporary folder for that request and deleted as soon as the job finishes. Either way, no address in it is ever requested.
You might also need.
Extract clean text from PDF files online. Copy your text instantly or download it as a TXT file for free.
Inspect PDF title, author, subject, keywords, producer, creator, creation dates, page count and document properties.
Remove invisible characters, zero-width marks and hidden PDF data before sharing: metadata, comments, bookmarks and embedded attachments. Your text stays real, selectable text and the page is never turned into an image.
Extract PDF bookmarks and table-of-contents entries with hierarchy and destination pages into JSON.