Extractability Scorer A model quotes a sentence, not a page.

Score a page

One page, not the whole site. Nothing is stored, and the scoring runs in your browser.

Checks run
7
Scored in
Your browser
Stored
Nothing

It lifts one line out of your copy, strips everything around it, and drops it into a paragraph about something else. Whatever that line was leaning on is gone. Score any page for how many of its sentences survive the journey.

What it checks

Two classes, and the difference decides the count. Blocking faults mean the sentence cannot stand up alone at all. Weakening ones mean it survives, but arrives worth less.

  • Opens with an orphan reference

    blocking

    The subject lives in a previous sentence. Lifted out, the claim is about nothing.

    Name the subject inside the sentence — the company, the product, the thing.

  • Points at the page around it

    blocking

    Above, below and here do not exist once the sentence is quoted somewhere else.

    Restate what is being referred to, or drop the pointer.

  • Nothing a model can verify

    blocking

    Superlatives carry no checkable content, so there is nothing to corroborate off-site and nothing worth quoting. This one is graded: when the superlative *is* the claim — "we are world-class" — the sentence has nothing left once it is doubted. When it only decorates a claim that survives without it — "integrates seamlessly with Salesforce" — the sentence is weaker, not broken, and it is marked that way.

    Replace the adjective with the fact that earned it.

  • Hedged

    weakening

    An answer engine is picking a sentence to assert. Hedged claims get passed over for ones that commit.

    Commit, or state the condition under which it holds.

  • Unquantified quantity

    weakening

    Many and most are not numbers. They cannot be checked, compared, or repeated with confidence.

    Use the figure, or say you do not have one.

  • Figure with no source

    weakening

    A number without attribution is the first thing a careful model declines to repeat.

    Name the source in the same sentence as the number.

  • Too long to quote whole

    weakening

    A sentence that has to be cut to be used gets cut where the model chooses, not where you would.

    Split it. One claim per sentence.

Questions

What is extractability?

Whether a claim still means something once it is lifted off your page. A model composing an answer does not quote your page — it quotes a sentence from it, stripped of the paragraph above and dropped into a reply about something else. Anything that sentence was leaning on is gone. Extractability is the property of surviving that.

Why does an orphan subject matter so much?

Because it is the difference between being cited and being paraphrased. "It reduces onboarding time by half" is useless quoted alone — a model either skips it, or attributes it to whatever noun happens to be nearest in its own draft, which is how brands end up credited with a competitor's claim. "Acme reduces onboarding time by half" cannot be misattributed.

Is this the same as readability?

No, and they pull in different directions. Readability rewards short sentences with pronouns doing the connective work, because a human reads top to bottom and carries context with them. Extraction has no context to carry. A page can score well on Flesch-Kincaid and be almost entirely unquotable.

Why flag words like world-class and seamless?

Because there is nothing behind them for a model to corroborate. An answer engine is deciding which source to stand behind, and it can check a founding date, a named customer, or a figure with a source. It cannot check an adjective. Superlatives are not just weak writing here — they are structurally uncitable.

Does the tool store my page?

No. The page is fetched once and passed straight to your browser, and every score you see is computed there from the text. Nothing is written down and nothing is queued for a human to look at.