Most tools sold as website audit software never read your writing.
They crawl diligently, report broken links, missing meta descriptions, redirect chains and slow templates, then hand you a spreadsheet. All of that is worth knowing. None of it answers the question a marketing lead usually arrives with, which is whether the words on the site are any good and which page is letting the rest down.
This page compares five tools people reach for, explains what each one genuinely inspects, and names the gap between them. We build one of the five, and that is stated wherever it matters. A roundup listing only its author is worth nothing to anybody.
What does a copy audit cover that an SEO audit does not?
Judgement about the prose itself, rather than the markup around it.
A technical audit answers whether search engines can reach a page and make sense of what they find. A copy audit answers whether a person who lands there finds anything worth reading. The two overlap less than the shared word "audit" suggests, and a site can pass one comprehensively while failing the other.
The distinction has grown sharper as answer engines began summarising pages rather than merely ranking them. Interchangeable copy is easy to crawl, easy to index, and offers a model nothing specific to quote.
How do the five compare?
| Screaming Frog | Ahrefs Site Audit | Clearscope | Hemingway | SiteTell | |
|---|---|---|---|---|---|
| What it crawls | A whole site, from a seed URL | A whole site, on a schedule | One page, against a target query | Nothing; you paste the text | A whole site, up to 400 pages |
| Does it read the copy itself | Spelling and grammar, plus readability | Not the prose; on-page elements only | Term coverage against a query | Yes, readability and style | Yes, register and structure |
| What you get back | Over 300 issue types, exported to a spreadsheet | A health score, and over 170 issue types by category | A content grade, with terms to add | A grade level, with passages highlighted | The rule that fired, quoted on the exact sentence |
| Free tier | 500 URLs per crawl | 5,000 pages a month, via their webmaster tools | A trial | The online editor | The whole scan, every page, no signup |
| Who it is built for | Technical SEOs and agencies | Anyone maintaining a site's search health | Content teams optimising for search | Anyone editing prose | Teams auditing a live marketing site |
Read the second row first. It is the one that separates these products, and the one most roundups get wrong in both directions.
What do the crawl-based tools inspect?
Structure, mostly, with a genuine edge into content on one of them.
Screaming Frog crawls outward from a seed URL and reports over 300 issue types, covering broken links, redirect chains, duplicate content, absent metadata and indexability. Its free version handles 500 URLs per crawl, ample for a small property and quickly exhausted by a large one. Worth correcting a widespread claim: Screaming Frog genuinely does examine content, checking spelling and grammar across many languages as well as readability, so calling it purely technical understates the product. Register lies outside its remit, meaning nobody will tell you your paragraphs echo every competitor's.
Ahrefs Site Audit crawls on a schedule and produces a health score spanning more than 170 issue categories, validating structured data against Google and Schema.org requirements along the way. Their documentation states the scope plainly, covering technical and on-page elements rather than prose quality. Such candour does them credit, and it draws exactly the line this article concerns.
Both deserve a place in any serious stack. Neither was built to warn you your careers page reads like a template.
What do content optimisation tools inspect?
Coverage against a target query, which is a narrower question than it first appears.
Clearscope takes a topic, a URL or existing text, compares it against pages currently ranking for that query, and recommends terms and structure to improve relevance. For a team producing articles aimed at specific searches, the workflow is well designed and the output is directly actionable.
Its frame is inherently page-by-page, and it measures relevance rather than distinctiveness. A page can hit every recommended term and still read exactly like the competitors it was measured against, since those competitors supplied the target. Optimising toward a consensus tends to produce consensus prose, which is a reasonable trade when ranking is the goal and a problem when differentiation is.
What do prose editors inspect?
Readability, on whatever you hand them.
Hemingway grades reading difficulty and highlights long sentences, passive voice and adverbs. It has been good at this for over a decade and remains the fastest way to find out whether a paragraph is heavy. It makes no claim about authorship, which is unusual and welcome in this market.
The constraint is scope. Hemingway reads text you paste in, so you pick the sample, and picking your own sample is how bad pages stay hidden.
Where does the gap sit?
Between "the page is technically sound" and "the page says something only you could say".
Nothing above reads a whole site for register. Crawlers cover every page but judge structure. Editors judge writing but read one page. Optimisation tools judge relevance against a query. The uncovered question, which is what a whole site sounds like when read end to end, is where we built SiteTell: it takes a domain, crawls up to 400 pages, and for each sentence that trips a rule reports which rule fired and why. No score without a reason attached, and no claim about who wrote anything.
Being explicit about the trade: we inspect no links, no metadata, no page speed, no schema. Anybody needing those should run Screaming Frog or Ahrefs, and many teams run one of those alongside us rather than instead.
Why does reading every page matter?
Because the page doing the damage is almost never the page you would have chosen to check.
When we scanned 97 domains and 10,126 pages on 14 September 2026, properties carrying 200 or more pages averaged 89.0 across them while their worst single URL averaged 52.3. The typical page on a large domain is fine. Somewhere down the tail sits something scoring in the low fifties, usually unread since launch, still collecting search traffic. Across the full dataset, 17% of sites had a worst page beneath 60, and the average URL carried 2.27 flags.
Sampling cannot locate it, and there lies the entire case for crawling. It also explains why worst-first ordering beats a site-wide average: averages describe how things generally stand, whereas the tail tells you what to repair on Monday.
What should a copy audit actually check?
Six things, and most tooling covers two of them.
Accuracy. Whether claims, prices and product descriptions still match reality. Nothing automated catches this; it needs somebody who knows the product reading the page.
Consistency. Whether the same feature carries the same name everywhere. Crawl exports help, since you can grep a column of extracted text for competing terms.
Register. Whether the writing sounds like a specific company or like the category average. This is the one nothing in the standard stack measures.
Readability. Whether sentences carry their weight. Hemingway and Screaming Frog both touch this.
Structure. Whether headings describe what follows, which matters more now that answer engines lift sections rather than pages.
Coverage. Whether pages exist for the questions buyers actually ask, which is where query-based tools earn their place.
An audit covering only the technical half will return a clean report on a site whose copy is quietly costing conversions, and the report will be accurate as far as it goes.
How should you put a stack together?
By question, since no single product covers the territory.
Ask whether search engines can crawl and understand the site, and you want Screaming Frog or Ahrefs. Ask whether a specific article can rank for a specific query, and Clearscope earns its place. Ask whether a draft reads clearly, and Hemingway answers in seconds. Ask which of your live pages reads like everyone else's, and you need something that crawls and reports patterns.
A sensible starting order for a small team: run a free technical crawl first, because broken pages make every other question moot, then read the copy site-wide, then optimise individual pages once you know which ones deserve the effort.