SiteTell
// comparison

SiteTell and Hemingway: readability, register, and which one catches what

Hemingway grades how hard your writing is to read. SiteTell reports whether it sounds like everybody else's. Where the two overlap, where they conflict, and what our scan of 97 domains says about checking a site one page at a time.

data measured September 14, 2026ruleset 2026-09-12
About these figures
They were produced under ruleset 2026-09-12, which changed on September 16, 2026. A scan you run today uses the newer one, so your score may differ from what these figures imply. See the current dataset.

Hemingway is the closest thing we have to a peer, and the comparison is worth making carefully rather than competitively.

Almost every other tool people put beside us guesses at authorship. Hemingway never has. It reads your prose, tells you where the reading gets heavy, and leaves the question of who typed it entirely alone. That shared refusal puts both tools on the same side of a line most of this market sits on the other side of.

They still measure different things, and the distinction is easy to state: Hemingway asks whether a sentence is hard to read, while SiteTell asks whether it sounds like everybody else's. A paragraph may score beautifully on the first count and dismally on the second.

What does each one actually measure?

What Hemingway and SiteTell each measure, what they accept as input, and how much writing each one reads at a time
HemingwaySiteTell
What it measuresReadability: sentence difficulty, adverbs, passive voiceRegister: the patterns that make copy read as interchangeable
What you get backA grade level, with passages highlighted by categoryThe rule that fired, quoted on the exact sentence
What it takes as inputPasted or typed textA domain
How much it reads at onceOne piece of writingUp to 400 pages per scan
Finds your worst page for youNo, you choose what to paste inYes, results are ordered worst page first
Rewrites the text for youYes, in the paid tier, with grammar and tone toolsYes, on the sentences it flagged
Works offlineYes, in the desktop versionNo, it crawls a live site
Makes a claim about who wrote itNoNo
Free tierThe online editor, plus a two-week trial of the paid versionThe whole scan, every page, no signup
Claims checked September 22, 2026

Grade level against register, in short. Those are separate properties of a text, and improving one does nothing automatic for the other.

Where do the two genuinely overlap?

At sentence rhythm, and that is roughly the extent of it.

Hemingway flags sentences running long or turning dense, nudging you toward variation. We measure something adjacent via a rule called sentence_length_variance, firing whenever the sentences on a page cluster too tightly around one length. Across our scan of 97 domains and 10,126 pages, taken on 14 September 2026 under ruleset 2026-09-12, it caught 41% of sites.

Notice the direction, though. Hemingway wants sentences shorter. We want them uneven. A page built entirely from crisp nine-word declaratives sails past Hemingway and trips our rule, since uniformity is itself a signature of machine drafting. Run both products and you may receive advice pulling opposite ways, less a contradiction than proof they are grading different exams.

What has Hemingway no category for?

The three commonest structural tells in our dataset, none touching readability at all.

redundant_closing_paragraph fired on 66% of the 97 domains, the single most widespread habit we recorded. It catches a final paragraph restating whatever the preceding ones already established. Hemingway reads such a paragraph as perfectly lucid, because it is; lucidity was never the difficulty.

em_dash_frequency fired on 63%. Dashes are legitimate punctuation, deployed beautifully by good authors for a century, and our rule measures density instead of presence. One caveat belongs beside the figure: this count straddles a rule revision dated 12 September 2026, when we shifted from one flag per page to one per sentence, leaving totals incomparable across the boundary.

rule_of_three_pattern also fired on 63%. Triples chosen for cadence instead of content rank among the most recognisable mannerisms in machine-drafted copy.

Flagged — redundant_closing_paragraph
Suggested rewrite
Illustrative. Nothing here is quoted from a scanned site.

Paste that original sentence into Hemingway and it comes back readable. Grade eight or thereabouts, no adverbs, active voice. It is also the most skippable sentence on the page, and a reader has seen it a thousand times.

Why does scope change the answer?

Because on a large property the writing costing you money is rarely the writing you would think to inspect.

Hemingway reads whatever you paste in, so you choose the sample. Nobody feeds 200 URLs through an editor individually, meaning people realistically check their homepage, their pricing page, plus whatever draft sits open. Our figures suggest the habit misses the problem entirely.

Sites carrying 200 or more pages, 21 of them here, averaged 89.0 across their pages while their worst single URL averaged 52.3. The typical page on a big domain reads fine. Somewhere down the tail sits something scoring in the low fifties, never revisited since launch. Compare single-page sites, 24 of them, averaging 94.9, where a bad page has nowhere to hide.

That spread of nearly 37 points between average and worst makes the case for crawling. Hemingway is untouched by it; the point concerns sampling, since choosing your own pages hides the tail by construction. Across the whole dataset, 17% of sites had a worst page scoring under 60.

What does a readability grade miss?

Word choice, for one, at least the kind that signals a register rather than a difficulty.

Hemingway will tell you a word has a simpler alternative. It has no opinion about whether a word has become a cliché of software marketing, because that judgement requires knowing what thousands of other sites currently say. Our vocabulary rules exist for exactly that, and they accounted for 42% of everything flagged across the scan.

The five most widespread vocabulary flags, by share of the 97 domains

Each one is a perfectly ordinary English word, and our ruleset marks all five ambiguous, since legitimate uses obviously exist. A geologist writing about terrain has earned the second. A locksmith describing what their hardware physically does has earned the first. Our flag reports frequency within a context; the judgement stays yours.

Hemingway passes all five without comment, correctly, since none of them makes a sentence harder to follow. They make it harder to distinguish, a separate failure needing separate equipment.

Can you run both?

Yes, and the sequencing tends to matter more than the choice.

Hemingway belongs at the drafting stage, where it responds instantly and you are still deciding what the sentence will be. A crawl belongs after publication, when there are enough pages that nobody remembers what all of them say. Running our scan first and Hemingway second is the awkward order, because you end up rewriting sentences that a readability pass would have restructured anyway.

One practical note: our ruleset reads published HTML, so anything behind a login, inside a PDF or sitting in a draft is invisible to us. Hemingway covers that territory comfortably.

What does Hemingway do better?

Several things, genuinely.

It works while you type, and we cannot. Our crawler reports on published URLs, so the feedback loop runs in hours instead of seconds. Mid-draft, Hemingway sits beside the cursor and answers immediately.

Its desktop edition runs offline, mattering enormously to anybody handling confidential or unpublished material. We fetch live addresses, leaving anything unshipped invisible.

It also reads absolutely anything: essays, fiction, a cover letter, an internal memo. Our ruleset was calibrated against marketing and product copy, so aiming it at a short story would yield confident nonsense. Inside its own territory Hemingway has excelled for over ten years, and it survived because the founding idea was sound.

Which one should you use?

By the job in front of you, rather than by whichever sounds grander.

Editing one piece right now, wondering whether the prose feels heavy? Reach for Hemingway. Wondering which live URLs read like a competitor's, and precisely what went wrong inside them? Only a crawler answers this, since you cannot inspect what never occurred to you to open.

Plenty of teams run both at different moments: Hemingway while drafting, a site scan afterwards, with neither asking who typed the words.

Questions people actually ask

Is SiteTell a Hemingway alternative?

Partly. Both read your writing and neither makes any claim about authorship, so they belong in the same family. They measure different properties though: Hemingway grades readability on text you paste in, while SiteTell crawls a domain and flags the patterns that make copy read as generic. Many teams use both at different stages.

Does Hemingway check a whole website?

No. Hemingway reads text you paste or type into its editor, one piece at a time, and does not crawl. Auditing every page on a domain needs a crawler; SiteTell covers up to 400 pages per scan and orders results worst page first.

Does Hemingway detect AI writing?

No, and that is deliberate on their part. Hemingway reports readability and style, and makes no statement about whether a person or a model produced the text. SiteTell takes the same position, which is what makes the two tools comparable at all.

Can good readability scores still mean generic copy?

Yes, and it is common. The most widespread pattern in our dataset, a closing paragraph restating everything above it, fired on 66% of 97 domains and reads perfectly clearly. Readability and originality are independent properties, so a page can score well on one and poorly on the other.

Do Hemingway and SiteTell ever give conflicting advice?

Occasionally, on sentence length. Hemingway encourages shorter sentences, while our sentence_length_variance rule fires when sentences cluster too tightly around a single length, which happened on 41% of sites we scanned. A page of uniformly short sentences satisfies one tool and trips the other.

Which is better for a large site?

A crawler, for the simple reason that you cannot paste 200 pages into an editor. Among sites of 200 or more pages in our sample, the average page scored 89.0 while the worst page averaged 52.3, so the damage sits in a tail that manual checking rarely reaches.

Is Hemingway free?

The online editor is free to use. There is a paid version with grammar and tone tools and a two-week trial that requires no card, plus a one-off desktop application that works offline. Consult their site for current prices, since we do not republish figures we cannot verify directly.

What does SiteTell cost?

Crawling costs nothing, with no account required, and you keep every flag we raise. Charges begin only if you want our generated replacements for the sentences we marked. You therefore see the entire diagnosis up front and decide afterwards whether buying the repairs appeals.

Want to know which of these your site does?

SiteTell crawls every page and flags the exact sentences, with the rule that fired next to each one. Free for the whole site.

Scan your site free
SiteTell

The read a careful editor would give your site, run in twenty seconds. Built by Rally Digital.

PRODUCT
COMPANY
LEGAL
© 2026 Rally Digital