Bulk URL cleaner
Paste a list of links from wherever it came from. Duplicates out, tracking parameters out, broken lines separated — with a count of exactly what changed.
Runs entirely in your browser — nothing is uploaded, and it is free with no account.
Why the count matters before you check
If you bought 200 links, the delivery sheet says 200, and the list cleans down to 176, that gap is worth resolving before you check anything — not after. Twenty-four of those lines were the same URL twice, or the same URL with a campaign parameter, or a row that never was a URL. Any tool that charges per URL charges for all 200.
The http/https split is the one that catches people out. Two entries pointing at the same page over different schemes are two different URLs, and on a site that redirects one to the other only one is ever going to be in Google's index under that exact address. Seeing the split before you check is the difference between reading a result and re-running it.
Once the list is clean, the actual question is how many of them Google has — how to check if a page is indexed covers which methods answer that honestly and which quietly do not. For a list of bought or built links specifically, how to check if your backlinks are indexed covers what to check on each linking page besides the index.
Questions
- Are two URLs that differ only by a tracking parameter the same page?
- For indexing purposes, almost always yes — and that is why stripping them matters. ?utm_source=x does not change what Google indexes, but it does make two lines look different to any tool counting them, so you pay twice for one page.
- Should I strip the fragment?
- Yes, for index checking. Everything after # is never sent to the server and is not a separate URL in Google’s index — example.com/page#section and example.com/page are one page. The exception is a site using the long-obsolete #! scheme, which nothing modern does.
- Does http:// and https:// count as one URL or two?
- Two, and the tool treats them as two. They are genuinely different URLs and only one of them is normally canonical — which is worth knowing, because a list that is half http on an https site is a list where half the entries may report as not indexed correctly.
- Why does it not remove the trailing slash by default?
- Because /page and /page/ can legitimately be different pages, and on many sites one redirects to the other. Removing it is offered as a choice rather than assumed — guessing here would silently change what you are checking.
Checking whether those URLs are actually in Google?
That is the part a browser cannot do for you. indexaction runs up to 1,000 URLs in one
check and returns a verdict for each — no site: searches, no CAPTCHA. New
accounts get 50 free credits.
Other free tools
- robots.txt tester — Paste a robots.txt and a URL, see which line decides it and why.
- Sitemap URL extractor — Turn a sitemap into a plain list of URLs you can paste anywhere.
- SERP snippet preview — See where Google cuts your title and description, measured in pixels.