Understand
How search engines see this page
Fetch a URL on this site and read the tags a crawler gets in the first response, then compare that text with the page after JavaScript.
Inspect a URL on this site
The server fetches the HTML and reads the tags in that response. The host has to be blackboxpersonalization.com, or this dev origin. The request uses a normal user agent, so the document is the one a crawler downloads before it renders.
Fetching the HTML…
How a crawler walks this site
The path is robots.txt, then sitemap.xml, then a fetch, a render, the index, and ranking. The first two boxes are this site’s live files.
01
robots.txt
The crawler reads which paths are allowed before it spends a fetch.
Reading robots.txt…
02
sitemap.xml
The sitemap lists URLs. It is a hint, not a guarantee that each URL is indexed.
Reading sitemap.xml…03
Fetch
The crawler downloads the HTML. The inspector reads that same kind of response.
04
Render
Googlebot then runs JavaScript. The comparison shows words that appear only after that.
05
Index
A noindex tag, or a disallow rule, keeps the URL out of the index.
06
Rank
Relevance, links, and page experience decide the order. Core Web Vitals are one part of that.
Core Web Vitals
Google publishes three page-experience measures: largest contentful paint, interaction to next paint, and cumulative layout shift. A visit is in the “good” range at 2.5 seconds, 200 milliseconds, and 0.1. They are a ranking input beside the words on the page and the links to it. This lab does not invent a score.
The performance lab measures LCP, INP, CLS, and TTFB for this visit.
SEO and analytics
A visit from a search results page is stored here with referrer class search. The host list is Google, Bing, DuckDuckGo, Yahoo, Baidu, and Ecosia. That class is this site’s label. It is separate from the query the visitor typed.
In GA4, unpaid search usually lands in the Organic Search channel. A Google results visit with no campaign parameters is commonly source google and medium organic. Parameters from the UTM builder replace that medium. The source lab shows the same referrer decision on this site.
Search Console is not connected. This server has no Search Console credentials, so this page has no query, click, or impression rows.
| Query | Page | Clicks | Impressions |
|---|---|---|---|
| No rows. Search Console credentials are absent. | |||
Fixes on this site
- Person JSON-LD names Neil Black and links to the about page. WebSite JSON-LD names the site, the www URL, and that person as publisher.
- Each blog post adds Article JSON-LD with the headline, description, date, author, and the share image.
- Every route gets a 1200×630 Open Graph image and a Twitter image.
- Canonical URLs use the www host. The title template appends the site name.
- robots.txt allows the public site, disallows account, admin, API, and the white-paper download and thanks paths, and points at sitemap.xml.
- Account, admin, sign-in, register, and the white-paper thanks page send noindex.
- The document language is English. hreflang is omitted because the site has one locale.
Recommended next
Why these picksReading path transitions…