Skip to content
Black BoxPersonalization

Understand

How search engines see this page

Fetch a URL on this site and read the tags a crawler gets in the first response, then compare that text with the page after JavaScript.

Inspect a URL on this site

The server fetches the HTML and reads the tags in that response. The host has to be blackboxpersonalization.com, or this dev origin. The request uses a normal user agent, so the document is the one a crawler downloads before it renders.

Fetching the HTML…

How a crawler walks this site

The path is robots.txt, then sitemap.xml, then a fetch, a render, the index, and ranking. The first two boxes are this site’s live files.

  1. 01

    robots.txt

    The crawler reads which paths are allowed before it spends a fetch.

    Reading robots.txt…
  2. 02

    sitemap.xml

    The sitemap lists URLs. It is a hint, not a guarantee that each URL is indexed.

    Reading sitemap.xml…
  3. 03

    Fetch

    The crawler downloads the HTML. The inspector reads that same kind of response.

  4. 04

    Render

    Googlebot then runs JavaScript. The comparison shows words that appear only after that.

  5. 05

    Index

    A noindex tag, or a disallow rule, keeps the URL out of the index.

  6. 06

    Rank

    Relevance, links, and page experience decide the order. Core Web Vitals are one part of that.

    Core Web Vitals for this visit

Core Web Vitals

Google publishes three page-experience measures: largest contentful paint, interaction to next paint, and cumulative layout shift. A visit is in the “good” range at 2.5 seconds, 200 milliseconds, and 0.1. They are a ranking input beside the words on the page and the links to it. This lab does not invent a score.

The performance lab measures LCP, INP, CLS, and TTFB for this visit.

SEO and analytics

A visit from a search results page is stored here with referrer class search. The host list is Google, Bing, DuckDuckGo, Yahoo, Baidu, and Ecosia. That class is this site’s label. It is separate from the query the visitor typed.

In GA4, unpaid search usually lands in the Organic Search channel. A Google results visit with no campaign parameters is commonly source google and medium organic. Parameters from the UTM builder replace that medium. The source lab shows the same referrer decision on this site.

Search Console is not connected. This server has no Search Console credentials, so this page has no query, click, or impression rows.

QueryPageClicksImpressions
No rows. Search Console credentials are absent.

Fixes on this site

  • Person JSON-LD names Neil Black and links to the about page. WebSite JSON-LD names the site, the www URL, and that person as publisher.
  • Each blog post adds Article JSON-LD with the headline, description, date, author, and the share image.
  • Every route gets a 1200×630 Open Graph image and a Twitter image.
  • Canonical URLs use the www host. The title template appends the site name.
  • robots.txt allows the public site, disallows account, admin, API, and the white-paper download and thanks paths, and points at sitemap.xml.
  • Account, admin, sign-in, register, and the white-paper thanks page send noindex.
  • The document language is English. hreflang is omitted because the site has one locale.
Why these picks

Reading path transitions…