PDF Link Results
Cleaned and deduplicated per pageWebsite icons are loaded lazily from domain.glass. Only the extracted domain name is requested; the PDF and link text remain in your browser. Internal and non-web targets are never opened automatically.
PDF Link Extractor with Page Numbers and Link Text
This browser-based PDF link extractor reads link annotations and visible text from every page. It cleans control characters and excess whitespace, removes harmless trailing punctuation from URL-shaped text, and deduplicates the same target on the same page.
CSV exports contain only Link, Link Page, and Link Text columns. Favicons are visual aids in the table and are intentionally excluded from downloads.
Remove Links Instead?
Remove clickable links, scripts, forms, attachments, and other active PDF features.
PDF Link Extractor Questions
Can it find URLs that are not clickable?
Yes. Keep Include visible URL text selected to scan page text for HTTP, HTTPS, www, and domain-shaped URLs.
Can it extract internal PDF and file links?
Yes. Select Include internal and other links to show internal destinations and non-HTTP schemes such as file and mailto. These targets are shown as text rather than opened automatically.
What is included in the CSV?
Each CSV row contains the cleaned link, its PDF page number, and sanitized one-line link text when that text is available.
Is the PDF uploaded?
No. PDF.js reads the file locally in this browser. The tool has no PDF upload or remote URL loader.
