Docket / Learn

Learn

Reference pages, written to be read rather than skimmed for keywords.

Answer extractability: can an AI quote your page?

The check reads your question-form headings and nothing under them. What that measures, what nobody can measure, and why one is a proxy for the other.

Trust and authorship signals a crawler can see

About, Contact and a privacy page are the reliable half. The authorship half is looser, and this is the page that says which is which.

The email-capture check that never looks at your form

It reads body text, and body text is defined as the page minus nav, header, footer, aside and form. What that means for a footer signup.

Does your copy read as AI-written?

Docket matches a register, not authorship — it cannot tell you who or what wrote a page. What it looks for, why it needs more than one tell, and what that finding is worth.

Who blocks Googlebot, and who blocks GoogleOther

We read robots.txt across a large public sample. A small share deny Googlebot and rather more deny GoogleOther — what that can, and cannot, tell you about a site.

Does Cloudflare block GPTBot?

Cloudflare changed what a new domain does by default, and the answer differs by which OpenAI crawler is asking. How to check your own.

Does noindex stop AI crawlers?

noindex tells an indexer not to list a page; it cannot stop a crawler fetching it. What robots.txt and your server control instead, and which layer is refusing yours.

Click depth and orphan pages

How many clicks from the homepage a page sits, and which pages nothing links to — plus the crawl conditions under which neither number means anything.

URL structure: checked vs folklore

What Docket actually enforces about URLs, and which popular rules have no published source behind them.

Page weight and how heavy is too heavy

Docket grades the HTML document, not the whole page — and says 'at least' when the read was capped.

How Docket decides what to fix first

Severity, impact, reach and effort — the formula an audit ranks your site with, written out.

AI crawler directives: which to block

Training crawlers and search crawlers are different decisions with different costs.

Is your page actually indexed?

Whether a page is in Google, Bing or Brave is a fact about their index, not your HTML. What Docket asks each engine, and the places it refuses to guess.

AI crawlers and your sitemap

A Sitemap: line is a non-group record, so every crawler is offered the same one — including the crawlers a site is trying to keep out.

www or no www, and whether it still matters

Every page-one result agrees you should pick one. None of them counted. We probed a random sample to see how many hosts still serve both.

Internal links that carry UTM parameters

A site tagging its own internal links splits its analytics and can split its canonical signals. We crawled a sample and counted — and the affected hosts split into two different problems.

Googlebot's 2MB cutoff

It reads the first 2MB and indexes that as the whole page. We measured well-known homepages and found five already past it.

How to tell whether an audit tool is lying to you

Four questions to ask of any SEO finding, each of them learned here by getting it wrong first.

Log file analysis: what Googlebot actually fetched

A user-agent proves nothing, which is why Google publishes 1,641 crawler IP prefixes. Reading a log honestly, and where the dedicated tool wins.

Brand consistency: the question no crawler asks

Of 11 company sites linking social profiles, 9 declared none of them in schema. What the brand lane checks, and where design tools beat it.

Domain authority without a subscription

Ranked from the public Common Crawl link graph. The useful part is what it says when it cannot see you, which is what it said about this site.

SEO monitoring: what changed, not what is wrong

Re-audits while the app is open, regressions first, and the two comparisons Docket refuses to make because the number would look useful and be wrong.

Conversion audit: 9 checks on your landing pages

Ranking and then failing to say what to do next costs the same as not ranking. The 9 mechanical checks, and the judgement calls Docket refuses to make for you.

Marketing tag audit: is your tracking on every page?

Tags are installed on templates; sites grow pages built from other templates. The 6 tracking checks, and the four-minute version you can do by hand.

The contact address that cannot receive mail

An address on a domain with no MX record bounces to the sender and never reaches you, and no tool asks whether yours works.

Which pages an AI answer replaces

Ranking and not being visited. Measured on two live sites — this one at 5% fully substitutable, a delicatessen at 0% — and three ways we measured it wrong first.

AI search visibility

The three gates a model has to clear before it can cite you — access, rendering and entity clarity — with measured data on who is blocking what.

What an SEO audit covers

Every area, in the order they should be worked, and the three tests a report has to pass to be worth acting on.

JavaScript rendering

What a crawler that does not run JavaScript misses — measured on a page that serves 0 characters of text and renders 2,068.

sameAs and entity signals

The cheapest entity signal there is, and the share of major sites that skip it — measured, with the dataset attached.

Canonical tags

Google calls rel=canonical a hint and overrules it routinely. The seven ways it gets set wrong, and what each Search Console status is actually telling you.

Internal link equity

The ranking signal your pages pass to each other, measured on our own site — where the download page held a fifth of what an average page did.

What Docket checks

All the checks, by area, with what each one actually looks at.