HTML Link Extractor
Extract quoted href values from a tags
Paste HTML to collect quoted href values from a tags, trimming outer whitespace and deduplicating in first-seen order. Anchor text, rel, and base URL resolution are not included.
The full guide also includes pitfalls, worked examples, snippets, FAQs, and related tools for checking results or troubleshooting.
About this tool
Paste an HTML snippet to collect quoted href values from a tags. The output keeps the first occurrence of each exact value, including relative paths, fragments, and non-HTTP schemes. It does not include anchor text or rel attributes, decode HTML entities, resolve a base URL, or check destinations. Matching uses a lightweight pattern rather than a complete HTML parser.
Production Snippets
Relative hrefs stay relative; repeated values appear once
text
INPUT
<a href="/guide">Guide</a>
<a href="/guide" rel="nofollow">Another label</a>
<a href="mailto:[email protected]">Email</a>
<a data-href="/not-an-href">Placeholder</a>
OUTPUT
/guide
mailto:[email protected]Frequently Asked Questions
Which attributes are included?
Only quoted href values on a tags, with leading and trailing whitespace removed. data-href, image src, unquoted href, anchor text, and rel are not exported.
How does deduplication work?
Exact extracted strings are deduplicated in first-seen order. Case, fragments, entity spelling, and query parameters are not normalized.
Are relative URLs or HTML entities resolved?
No. /guide, #section, mailto:[email protected], and literal & sequences remain as written. Resolve them separately using the source page context when needed.
Does it inspect the page or check broken links?
No. It only scans the pasted source, without fetching URLs or running JavaScript. Source inside comments or scripts can resemble anchors, and malformed markup can cause missed or false matches.
Keep browsing