Extract URLs from Text

–links found
InputCaught?Why
https://tooldune.com/x?a=1yesfull scheme form
www.example.org/pageyeswww prefix counts
example.comnobare domain - could be prose
(https://tooldune.com).yes, cleanedbrackets and periods trimmed
The recognized shape: http or https addresses and www-prefixed links, with queries and fragments intact and prose punctuation trimmed off the ends. Bare domains stay in your text on purpose - zero false positives beats greedy matching when you will act on the list. The URI grammar this implements: RFC 3986, Uniform Resource Identifier (text).
Pair tools: the email extractor for addresses instead of links, find and replace for the cleanup pass, and text compare to diff the before and after.

Paste a document, a chat export, a scraped page or a decade-old bookmarks dump, and this tool lines up every web link it finds - https and http addresses, plus bare www. forms - listed one per line with a running count and an optional dedupe that folds repeated links into one entry. Trailing punctuation that glued itself to the link (the period at the end of a sentence, the bracket of a forum post) comes off automatically.

Everything runs in this browser tab: no upload, no logging, no account - the text and the extracted list stay on your device, with the source kept locally so a refresh does not lose it. The honest boundary: a bare domain like example.com without a scheme or www is left alone, because words such as that are ordinary English; the tool extracts links that are unambiguous about being links.

How to use

  1. Paste your text - the link list builds live as you type, one address per line in order of appearance.
  2. Tick dedupe to fold repeats; untick for the raw every-occurrence list.
  3. Copy the block straight into your spreadsheet, link checker or migration sheet.

Frequently asked questions

How do I extract URLs from text?

Paste the text above. Every http:// or https:// address and every bare www. link is pulled out in order of appearance, one per line, with trailing punctuation stripped. Dedupe folds repeats, and the count up top tells you exactly how many links you caught before you paste them anywhere. Like its sibling tool for email addresses, there is no button - extraction is live.

Why isn't a bare domain like example.com caught?

Deliberate design: without a scheme or www, a string such as example.com is indistinguishable from ordinary prose - sentence-ending words, file names, product codes. Requiring https://, http:// or a www. prefix keeps the false positives at zero, which matters more than completeness when you are building a link list you will act on. Paste the domain with its scheme and it is caught like any other link.

Does it strip punctuation stuck to the link?

Yes - the period, comma, closing parenthesis, bracket and quote that prose glues onto the end of a link are trimmed automatically, so a sentence ending in https://tooldune.com/ yields the clean address. What stays: query strings, hash fragments and paths, because they are part of the link. The line between them is punctuation, not structure.

Is my text uploaded anywhere?

No. Extraction runs entirely in this browser tab - no server, no logging, no account. Your text stays in the page's local storage on your own device (up to 20,000 characters) so a refresh does not lose it, and nothing leaves your machine unless you copy it out yourself.

What counts as a link here, formally?

The tool recognizes web addresses in the scheme-plus-path shape defined by the URI standard - the grammar lives in RFC 3986, the internet's official Uniform Resource Identifier specification. This extractor implements the practical subset: http, https and www-prefixed web links with their queries and fragments. Schemes like ftp or mailto are left to their own tools - email addresses have a dedicated extractor one link away.

Related tools