EXU

URL Extractor

Find HTTP(S) URL candidates in pasted text

Extraction
πŸ”’ 100% client-side β€” your data never leaves this page
Maintained by Evanβ€’Updated: September 29, 2026
Options
Input Text

Paste logs, markdown, or JSON and run Extract URLs first to get a deduped list; domain distribution and scenario guidance are available in Advanced mode.

Extracted URLs
Extracted URLs will appear here
πŸ”’ Processed locally in your browser
Page reading mode

The full guide also includes pitfalls, worked examples, snippets, FAQs, and related tools for checking results or troubleshooting.

About this tool

Scan pasted logs, chat text, or source text for HTTP/HTTPS-prefixed candidates. Selected trailing sentence punctuation is removed, then exact strings are deduplicated and sorted. Query variants and case differences remain distinct. This is a text scan, not HTML, Markdown, JSON, or URL validation; escaped strings and URLs containing closing delimiters may need manual correction. Input is saved as a draft in this browser until cleared.

Production Snippets

Exact duplicates collapse; query variants remain separate

text

INPUT
GET https://b.example.org/p?id=1
Retry https://b.example.org/p?id=1.
See https://a.example.org/help and https://b.example.org/p?id=2
Bare example.org and ftp://files.example.org are skipped.

OUTPUT
https://a.example.org/help
https://b.example.org/p?id=1
https://b.example.org/p?id=2

Frequently Asked Questions

Which strings are found?

Candidates beginning with http:// or https://, case-insensitively. Bare hostnames, relative paths, and ftp:// URLs are skipped. JSON escape sequences and HTML entities are not decoded.

What changes during cleanup and deduplication?

Closing brackets or parentheses terminate a match, and trailing periods, commas, semicolons, exclamation marks, or question marks are stripped. Remaining exact strings are deduplicated and sorted; this is not canonical URL normalization.

What does the domain breakdown count?

It groups the unique output URLs by parseable full hostname. A malformed candidate may remain in the URL list but be absent from the breakdown. It does not group by registrable domain or retain raw occurrence counts.

Does it visit or validate the extracted URLs?

No. Candidates are not fetched, checked for safety, or confirmed reachable. Retain source context and verify a destination before taking action.

Does the tool keep my input?

Yes. It saves the draft in this browser local storage and restores it on return. Clear removes the saved draft. Processing is local and does not send the text to a server.

Keep browsing