Content Term Coverage

Paste 3–10 competitor URLs — the pages that rank for the query you care about — and, optionally, your own page URL or draft text. The tool fetches each page, extracts its main content, counts one- to three-word terms in the language you pick, and lists the terms that appear in at least 60 % of the competitor pages, with the count per page and the count in yours. Terms you never use are listed separately; word counts and H1–H3 outlines sit alongside. No SERP scraping, no content score — a coverage matrix you can read and export as CSV. Nothing is stored beyond 24 hours.

0 URLs — 3 to 10 per run

Compared like a fetched page, without headings. Never stored.

Pages are fetched by our server, one per second per host, and their text is discarded when the comparison ends. Only URLs, counts and this table are kept, for 24 hours.

How it works

1. You name the competitor pages (3–10 URLs); the tool does not fetch search results. Our server requests each page once, at most one request per second per host, 2 MB per page, 15 s per request, redirects followed up to five times; your own page, if given by URL, is fetched the same way — pasted text is used as is
2. Main content = the <main> or <article> element when the page has one, otherwise the body with navigation, header, footer, aside and forms removed, scripts and styles stripped. Word counts and H1–H3 headings are taken from that content
3. Terms: words lowercased and split on non-letters; 1-, 2- and 3-word sequences counted; sequences that begin or end on a stopword of the chosen language are skipped (so 'cost per click' counts, 'of click' does not). Japanese is split at script boundaries, not by a morphological analyser — treat its results as coarse
4. A term is listed when it appears in at least 60 % of the competitor pages that could be read (never fewer than two) and at least twice in total; up to 150 terms, ordered by number of pages, then total occurrences
5. Missing = listed terms that occur zero times in your text. There is no score: which missing terms matter is a judgement about the topic, not a count

About term coverage and what the matrix does and does not tell you

The idea behind term coverage is simple: the pages that rank for a query tend to cover the same set of subtopics, and those subtopics leave a trace in vocabulary. If nine of ten pages about email benchmarks mention 'unsubscribe rate' and yours does not, either you left something out or you called it something else. The matrix makes that visible per term and per page, so you can decide which. Commercial tools (Surfer, Clearscope and their relatives) do the same counting and then wrap it in a score and a target; this tool stops at the counting.

You supply the competitor URLs yourself. That is a limitation — a scraper that pulled the top 10 for you would be more convenient — and also the point: you choose pages that are actually comparable (informational against informational, product against product), and the tool does not need to scrape a search engine to work. Three pages are the minimum for the 60 % threshold to mean anything; ten is the cap because the count per page becomes noise beyond that.

Reading the matrix: terms that appear in every competitor and not in yours are the first thing to look at, but they are not automatically gaps. A term can be a competitor's brand, a navigation label that survived boilerplate removal, or a phrase your text expresses with different words. The per-page counts help — a term that appears 20 times in one page and once in the others is that page's obsession, not the topic's — and the H1–H3 outlines show whether a term is a section of its own or a passing mention.

Word count is shown as a median, not a target. Longer pages are not better; pages that cover more of what the query is about tend to be longer, which is not the same claim. If your page is a third of the median and misses half the shared terms, the two facts probably have the same cause. If it is shorter and misses nothing, it is shorter.

Frequently asked questions

Why do I have to paste the competitor URLs?
Because the tool does not scrape search results — that would be fragile, against the search engines' terms, and would make the tool's availability depend on them. Search your query, pick the pages that are genuinely competing with yours (same intent, same kind of page), and paste those. Three to ten.
Where is the content score?
There is none. A score would weigh terms nobody has justified weighing, and a target score would push text toward stuffing. The matrix shows which terms the competitors share and how often each page uses them; deciding what your page should cover is editorial work, and the tool does not pretend to do it.
Some of the listed terms are nonsense — navigation labels, cookie notices, a competitor's brand.
Boilerplate removal is heuristic: <main> and <article> when present, otherwise nav, header, footer and aside stripped. Sites that put their menu inside the article, or repeat a call-to-action in the body, leak such terms into the list. Ignore them; the per-page counts usually make them obvious (present everywhere at a low count, or only in one page at a high count).
Why is a competitor page marked as failed?
It returned an error status, was not HTML, took longer than 15 s, or refused our crawler (bot protection answers 403 or 429 to unknown user-agents). Failed pages are excluded from the threshold and the counts; the summary says how many were read. If a key competitor fails, paste its text as your 'own' input in a second run to at least see its terms.
Does it handle languages other than English?
Turkish, German and Russian use the same word splitting with their own stopword lists; results are as good as for English. Japanese has no word spaces, and the tool splits at script changes (kanji runs, kana runs, Latin runs) rather than with a morphological analyser — phrases come out coarse, and the counts are indicative only. Pick the language of the pages, not of the interface.
Can I use it on my draft before it is published?
Yes — choose 'Paste text' and paste the draft. It is compared exactly like a fetched page, minus headings. Up to 200,000 characters.
What happens to the pages you fetch?
The text is used inside one job on our server and discarded when the job ends; the result keeps only the URLs, word counts, headings and the term table. The job itself is deleted after 24 hours. Your pasted text is never stored beyond that.
SEO AuditEnter a URL and the crawler checks up to 200 pages the way a search engine's first pass would: status codes, noindex and canonical directives, missing or duplicate titles and descriptions, missing H1s, thin pages, images without alt, redirect chains, broken links, orphan pages, robots.txt and sitemap coverage. Every page gets the same on-page checks as the single-page checker; the report groups them by issue and by page. No score, no fabricated priorities — findings with counts. Download as CSV. Nothing is stored beyond 24 hours.On-Page SEO CheckerEnter a URL and, optionally, a target keyword. The checker fetches the page as Chrome and as Googlebot and reports what the raw HTML says: title and meta description with character and pixel length, robots meta and X-Robots-Tag, canonical, H1 count and heading outline with skipped levels, image alt coverage, internal and external links and nofollow share, Open Graph and Twitter card, hreflang validity, JSON-LD types, lang, viewport, charset, word count and keyword density and placement. No score — a list of what passed, what to look at and what is only information. Nothing stored.Google SERP SimulatorSee how your page will look in Google search results before you publish. Type a title, meta description and URL — or pull them from a live page — and get a pixel-accurate desktop and mobile preview, pixel and character counters, the exact character where Google cuts the text, and a list of findings. Free, no signup, nothing stored.Internal Link CheckerEnter a URL and the crawler maps the internal link graph of up to 200 pages: how many pages link to each URL, how many links each page sends out, how many clicks each page sits from the start, and which pages nobody links to — orphans that only the sitemap knows about. The table sorts by incoming links so the pages your own site treats as unimportant are at the bottom. Download as CSV. Nothing is stored beyond 24 hours.