Website URL Extractor
List every address a page refers to, not just its links. The extractor collects links, images and srcset candidates, scripts, stylesheets, icons and other link tags such as canonical, alternate, manifest and preconnect, frames and embeds, video and audio sources, form actions, URLs in meta tags and url() references in inline CSS – each with the place it was found and its host. Filter by type and export as text or CSV.
- Encrypted connection
- No sign-up
- Free to use
How to use Website URL Extractor
- Enter a page address (or paste HTML).
- Click “Extract URLs”.
- Filter by type.
- Download as TXT or CSV.
Website URL Extractor features
13 kinds of URL
Links to CSS url().
Where found
Element and attribute.
Hosts
Third parties at a glance.
Export
Download results as TXT or CSV.
Paste mode
Analyse HTML you paste, e.g. from a staging site.
Safe fetching
Public addresses only, with size and time limits.
When to use Website URL Extractor
- Migrations: finding every asset to move.
- Privacy reviews of third-party hosts.
- Content Security Policy planning.
- Debugging missing resources.
Website URL Extractor FAQ
How is it different from the Link Extractor?
The Link Extractor lists clickable links; this tool lists every resource the page loads or refers to.
Does it read external CSS files?
No, only inline CSS; use the CSS Extractor for stylesheets.
Are duplicates removed?
Yes, per type.
Can it help with a CSP?
Yes – the host list shows which domains a Content-Security-Policy must allow.
Every dependency of a page
A page depends on many addresses besides its links: scripts, fonts, images and frames from several hosts. Seeing them all is the first step for migrations, privacy reviews and security policies.
The CSV can be sorted by host to group third-party dependencies.
How it works: our server downloads the page once through a guarded fetcher that only connects to public addresses, follows a limited number of redirects and stops after a size and time limit. The HTML is then analysed in your browser as inert text – scripts on the page never run and nothing is stored.
What it cannot see: content and resources that a page adds with JavaScript after it loads, pages behind a login, and servers that block automated requests. For those, open the page in your browser, use its developer tools, or paste the page source where the tool offers a paste option.
Use the results as a starting point: fix the items marked red first, review the yellow warnings in context, and run the check again after a change. Requests are rate-limited to keep the service fair; if you check many pages in a row, wait a few minutes.
Related checks on this site cover the rest of a technical review – speed and Core Web Vitals, security headers, structured data, accessibility and SEO signals – so you can work through a whole site audit one topic at a time.
Who it is for: site owners checking their own pages, developers debugging a release, SEO and marketing teams auditing clients or competitors, and students learning how the web works. No account or installation is needed, and the results are plain text and tables you can copy into a report or ticket.