Keyword Density Analyzer
Paste your content, optionally enter a target keyword, and see frequency + density for every word.
| # | Word / Phrase | Count | Density | Status |
|---|
Green = 0–2% (healthy) | Amber = 2–4% (borderline) | Red = >4% (likely over-optimized)
Why Your Keyword Density Numbers Are Probably Lying to You
There is a persistent myth in SEO that you can pick a magic density percentage — say, 1.5% or 2% — stuff your target phrase in exactly that many times, and watch your rankings climb. A decade of testing across millions of pages has largely dismantled this idea. And yet keyword density remains a genuinely useful diagnostic — not as a ranking lever to pull, but as a warning system that catches bad habits before they cost you.
The distinction matters. Density analysis is a tool for editors and writers, not a ranking formula. Understanding what the numbers actually mean — and what they don't — changes how you use them entirely.
What Keyword Density Actually Measures
Density is simple arithmetic: the number of times a word or phrase appears divided by the total word count, expressed as a percentage. If your 1,000-word article uses the word "mortgage" 15 times, your density for that term is 1.5%.
The calculation itself is unambiguous. What gets messy is deciding which words count. Should you include stop words like "the," "and," "for" in the denominator? Should you count multi-word phrases as one unit or multiple single words? Should you stem words so that "optimize," "optimizing," and "optimization" all count together?
Different tools answer these questions differently, which is why you'll sometimes see wildly different density figures for the same content depending on which analyzer you use. When you run your own analysis, knowing how the tool tokenizes text — whether it strips punctuation, lowercases everything, handles hyphens — lets you interpret results with appropriate skepticism.
The History of Density as an SEO Signal
In the late 1990s and early 2000s, early search engines leaned heavily on term frequency as a proxy for topic relevance. A page that mentioned "digital camera" forty times in five hundred words was, by that crude measure, very much about digital cameras. Webmasters quickly learned to exploit this. Pages were stuffed with invisible text, keyword-loaded alt attributes, and footer blocks of repeated phrases.
Google's response was gradual but systematic. The Panda algorithm update in 2011 hit thin and over-optimized content hard. Penguin in 2012 went after manipulative linking patterns but also reinforced the idea that mechanical optimization signals were suspect. By the mid-2010s, Google had shifted substantially toward entity recognition, semantic relationships, and behavioral signals — things that keyword density simply cannot capture.
Today, a search quality rater looking at your content isn't counting how many times you used your target phrase. They're asking whether the page satisfies the search intent behind the query. Density analysis has no answer to that question.
When High Density Is a Real Problem
Despite everything above, density analysis still catches genuine issues. The clearest case is unintentional repetition. A writer drafting a 2,000-word guide on "content marketing strategy" might use the exact phrase dozens of times simply because it's the natural shorthand for the topic — without realizing the page now reads like a product description written by an algorithm.
Readers notice before Google does. Prose that repeats the same phrase every third sentence creates cognitive friction. The content feels mechanical, which reduces time-on-page and increases bounce rates — both behavioral signals that matter for rankings even if density itself doesn't.
High density analysis also flags thin content problems. A 300-word page where the same five words account for 12% of the total word count isn't just over-optimized; it's probably too short and too narrow to serve any query with real depth.
Practical Thresholds Worth Knowing
Industry convention — supported by various correlation studies, though not by any official Google guidance — suggests that single-word keyword density above 4% is where content starts to look manipulative in automated quality assessments. Between 2% and 4% is a borderline zone: not alarming, but worth examining the surrounding context. Under 2% for any given term is generally healthy for naturally written content.
These thresholds apply to individual words. For multi-word phrases, the math changes significantly. A two-word phrase appearing at 1% of total word count would mean the phrase appears roughly once every 100 words — which in a 1,500-word article is fifteen repetitions of the same exact phrase. That's almost certainly too many for natural writing.
The real signal is relative density across all terms. If your top five highest-density words are all variations of your target keyword, the content probably lacks topical breadth. A well-written, genuinely useful piece will show a diverse vocabulary distribution — the top terms should be varied, with your primary keyword appearing at a similar rate to other topically relevant terms.
Using Density Analysis Before and After Drafting
The most practical application is running an analysis immediately after finishing a draft, before any SEO review. At that stage, you're not trying to add keywords — you're looking for the opposite problem: have you been so focused on one phrase that the content reads robotically?
A useful exercise: run the analysis, note the top twenty words by frequency, and read the list as a proxy for your content's topical coverage. If the list is narrow and repetitive, the article itself is probably narrow and repetitive. If it shows a rich vocabulary around your topic — varied terms, related concepts, naturally occurring synonyms — the content is likely serving readers well.
After publishing, density analysis becomes a diagnostic for underperforming pages. If a well-linked page isn't ranking for its intended query despite having broadly good technical SEO, running a density check occasionally reveals over-optimization that predates better writing habits. A careful rewrite that redistributes and varies keyword usage — while improving overall quality — often moves these pages without any other intervention.
What Density Analysis Cannot Do
It cannot tell you whether your content matches search intent. It cannot measure whether your page answers questions better than the pages currently ranking. It cannot assess E-E-A-T signals, page experience metrics, or the quality of your internal linking structure. It gives you vocabulary statistics — and vocabulary statistics are one small input into a very large optimization picture.
The temptation to treat density as a direct ranking factor persists partly because it's quantifiable in a discipline where most important factors are frustratingly fuzzy. But over-reliance on any single quantifiable metric is exactly how pages get written for crawlers instead of humans.
Use the numbers as a sanity check. Keep your density in the healthy range not because Google is counting, but because content that reads naturally tends to be content that actually helps people — and over time, that correlation with reader satisfaction is what search engines are trying to measure anyway.