Domain Extractor
Strip URLs down to domains or root domains, with counts and deduplication.
Pull every link out of a block of text or HTML.
Used to mark links internal or external.
| URL | Anchor text | Type | rel |
|---|
URL Extractor finds every web address in whatever you paste. Give it plain text and it picks out anything that looks like a URL. Give it HTML and it reads the actual anchor tags, which means you also get the anchor text and the rel attribute of each link rather than just the address.
That second mode is the useful one for auditing a page. Paste the source of an article and you get a table of every link it contains, with its anchor text, whether it is internal or external relative to a domain you nominate, and whether it carries nofollow or sponsored. Scanning that table takes seconds; scanning the raw HTML does not.
Filters let you keep only external links, only internal ones, or only those matching a pattern, and the result can be copied as a plain list or downloaded as CSV for a spreadsheet.
The url extractor is used by writers, developers, students, marketers and anyone else who needs the job done once without installing software. Common cases include:
Text mode matches anything that looks like a URL anywhere in the input, including inside plain prose. HTML mode parses the markup and reads real anchor elements, so it also gives you the anchor text, the rel value and the target of each link.
In HTML mode, a relative address such as /about is resolved against the domain you supply, so it appears as a full URL and is correctly marked as internal. Without a domain it is listed as it was written.
No. A page running in your browser cannot request another site's HTML, because browsers block cross-origin requests. Use your browser's view-source or developer tools to copy the HTML, then paste it here.
Text mode also matches addresses beginning with www, which covers most of what appears in plain prose. Addresses written without either, such as a bare example.com, are too easily confused with ordinary words to match reliably.
Yes. Use the pattern filter with the domain name, and only matching URLs are kept.
No. Parsing happens in your browser, which matters when the page you are auditing is not published yet.
If the url extractor is not quite what you need, these other free tools solve closely related problems.
Strip URLs down to domains or root domains, with counts and deduplication.
Remove repeated URLs, treating trailing slashes and tracking parameters as the same.
Remove utm and click-tracking parameters, fragments and stray characters from URLs.
Convert a whole list of URLs and anchor texts into HTML or Markdown links.
Split a URL into every component, including a table of query parameters.