Keyword List Cleaner

Nothing to clean yet.
What to fix

What does a keyword list cleaner do?

It takes the messy list you have and returns the list you meant. Merge exports from Search Console, a competitor scrape and two brainstorm docs and you do not have one keyword list - you have the same keyword three times, in three casings, with trailing spaces and a scattering of empty rows.

This tool removes duplicates, normalises casing and whitespace, strips punctuation left by exports, and optionally drops stop words - all in your browser, in one pass, on lists of tens of thousands of rows.

How to remove duplicate keywords from a list

  1. Paste the whole list into the left box - from a CSV, a spreadsheet column, or several sources stacked together.
  2. Leave Remove duplicates, Lowercase and Trim and collapse spaces on. Those three together catch nearly every duplicate, because most are not exact matches - they differ only by case or a trailing space.
  3. Read the counter: it tells you how many rows went in, how many came out, and how many were duplicates versus blanks.
  4. Copy the cleaned list, or download it as a .txt.

Order matters here. Lowercasing and trimming before deduplicating is what turns Running Shoes, running shoes and running shoes  into one keyword instead of three. This tool always runs them in that order.

Why duplicate keywords cost you money

Because every duplicate is a paid lookup for an answer you already have. Any tool that charges per keyword - ours included - bills the row whether or not it has seen it before.

Deduplication is the biggest win and the easiest to skip, because duplicates are invisible in a spreadsheet of four thousand rows. Lists assembled from several sources routinely shed a fifth to two-fifths of their rows to exact duplicates and whitespace before a single genuine keyword is lost.

Should you remove stop words from keywords?

Usually not, which is why the option is off by default. Dropping the, for and of tightens a list when you are building modifier sets or grouping themes.

It actively damages long-tail research. Insurance for small business and insurance small business are different searches with different volumes and different intent, and collapsing them loses the distinction you were paying to discover. Turn it on deliberately, for a specific job - not as general tidying.

What each cleaning option does

  • Split on commas and tabs - turns a pasted CSV row into one keyword per line. Leave it on unless your keywords genuinely contain commas.
  • Trim and collapse spaces - the invisible killer. running shoes with a trailing space is a different string to every tool on earth, and looks identical to you.
  • Lowercase everything - search engines are case-insensitive, so casing only manufactures duplicates.
  • Drop blank lines - clears the empty rows exports leave behind.
  • Remove duplicates - keeps the first occurrence, drops the rest.
  • Strip punctuation - clears quotes and brackets. Hyphens and in-word apostrophes survive, because t-shirt and men's are real keywords.
  • Sort A–Z - alphabetical rather than input order, for eyeballing a long list.

Clean the list first, then look it up

The order matters more than it sounds. Clean, then send for volume and CPC - not the other way round. Cleaning afterwards means you have already paid for the duplicates.

  1. Paste everything you have collected into the box above.
  2. Take the cleaned output to KWScanner, which handles up to 3,000 keywords per run and bills per keyword rather than per seat.
  3. Group what survives into lists for the pages and campaigns you are briefing.

If you are also sizing up the competition, the bulk DR checker and bulk domain age checker handle the domain side of the same question.

Is this keyword cleaner private?

Completely. The cleaning is plain string work performed by your own browser - there is no request to a server, nothing is logged, and nothing is stored. Disconnect from the network and the tool still works, which is the only honest test of that claim.

Common questions

Why clean a keyword list before uploading it?

Because every duplicate is a wasted lookup. Keyword tools that charge per keyword - ours included - bill the row whether or not you have seen it before. A list exported from three sources typically loses 20-40% of its rows to duplicates and whitespace alone.

Does removing stop words help?

It depends what the list is for. If you are building modifier sets or grouping themes, dropping the, for, of and so on tightens the list. If you are researching long-tail phrases, do not: "insurance for small business" and "insurance small business" are different searches with different volumes. The option is off by default for that reason.

Is my list uploaded anywhere?

No. The cleaning runs entirely in your browser - nothing is sent to a server, logged, or stored. You can disconnect from the network and it still works.

How many keywords can it handle?

Tens of thousands comfortably. It is plain string work in your own browser, so the practical limit is your machine rather than any quota we impose.

What counts as a duplicate?

An exact match after the other options you have enabled have run. So with lowercasing on, "Running Shoes" and "running shoes" are one keyword; with it off, they are two. The first occurrence is the one kept.

What do I do with the cleaned list?

Take it to a keyword tool for volume, CPC and competition. KWScanner accepts up to 3,000 keywords in a single run and bills per keyword, which is exactly why trimming the list first is worth the thirty seconds.

Clean list. Now get the numbers.

A tidy list is only worth having if you find out what people search for. KWScanner returns volume, CPC and competition for up to 3,000 keywords in a run - priced per keyword, credits never expire.