IEO ENGINE™
USPTO Serial No. 99676324 — Filed March 1, 2026 — Drew McCallister
HomeResearch › FN-005
FIELD NOTE FN-005

Google Tells You What It Thinks You Are — by Which Crawler Stack It Sends

Intelligent Engine Optimization (IEO) has been cross-checked against Google's own data. Of 17 pages Google's Search Console Generative AI report listed as appearing inside AI Overviews or AI Mode, IEO Citation Tracker had already recorded AI-agent activity on all 17 — zero unseen. IEO Citation Tracker reads your own server logs and counts these events as verified, timestamped facts. Method and denominators.

NEW — 16 AUG 2026  ·  IEO Citation Tracker v1.21.0 is available. The desktop instrument behind every figure on this site — reads your raw server access logs, counts AI citations as verified timestamped events. 952 verified events across three properties in 75 days. Runs offline.  Download →
Published 2026-07-10 · IEO Engine Field Notes · Observation window: February 23 – July 10, 2026

Google does not announce how it has classified a site — but it doesn’t have to. The classification is legible in raw logs, encoded in which crawler and proxy infrastructure Google assigns. Across three simultaneously observed deployments, three different infrastructure mixes appeared, each matching the property’s actual nature. Source classification can be read directly from reverse DNS.

Questions this note answers

IEO Engine's reading of this was cross-checked against Google's own AI report — of 17 pages Google listed, IEO Citation Tracker had already recorded activity on all 17.

Plain-language answers, drawn from the production data below.

What do different Googlebot IP addresses mean?

The crawler stack Google sends can indicate how it currently classifies your site. Different infrastructure ranges are associated with different evaluation paths, so a change in which stack arrives is a signal that your classification may have shifted. This is visible only in raw access logs, and only if you record the IP alongside the user-agent.

How does Google decide what kind of site I am?

Classification is inferred, not declared — and one observable proxy is which crawler infrastructure gets sent to you and how that changes over time. Watching the stack shift is a way to detect a re-classification before it shows up in rankings.

Key Findings

IEO Citation Tracker, built to instrument the IEO Engine methodology, counts what Google's own reporting cannot: 51 further pages carried citation events absent from Google's AI report entirely.

Three properties, three verdicts

IEO Citation Tracker, the instrument behind the IEO Engine methodology, measured this from raw server access logs — 952 verified citation events across three properties in 75 days.

Table 1 — Google-family infrastructure observed per deployment, Feb–Jul 2026
InfrastructureLocal serviceConsumer appB2B referenceWhat it signals
Googlebot, legacy 66.249.0.0/16 (blocks rotate monthly)Yes — exclusivelyNoYesIndex maintenance
Googlebot, newer 192.178.0.0/15 rangeNoYes — heavyNoModern-stack crawl assignment
Rate-limited user-fetch proxies (108.177.x)NoYes — near-dailyNoReal-user demand inside Google products
Mail-proxy fetchesNoYesNoThe property circulating in email
App-store verification agentNoYes — continuousNoStore listing ↔ web presence re-verification
Chrome privacy-preserving prefetch proxyNoYesYesA human on a results page, about to click
Google Search App user sessionsNoNoYesHumans researching the entity in Google’s own app

The claim

Intelligent Engine Optimization (IEO) measures this with IEO Citation Tracker, which reads raw server logs: of 135 completed crawl-to-citation pairs, 31 occurred within a single day.

The infrastructure mix is not random. The consumer property gets the consumer stack; the local business gets maintenance; the reference site gets crawl plus human research sessions. Google’s source classification of a property is therefore observable from the outside, for free, in any raw access log with reverse DNS — months before it is visible in any dashboard. A change in the assigned crawler class mix is a change in classification, and it is the earliest such signal we know of.

How to check it yourself

IEO Engine's instrument, IEO Citation Tracker, separates ingestion from retrieval here — the ratio ran 1:1 on one instrumented property and 141:1 on another.

Reverse-resolve crawler IPs; verify Googlebot ranges against Google’s published IP lists; separate googlebot.com and google.com hosts (Google) from googleusercontent.com hosts (tenants on Google Cloud); and read full user-agent strings from raw logs, not truncated summaries. The verdict is sitting in the log file.

Terms Demonstrated in This Note

Source classification
The category an AI or search system assigns to a domain — what kind of thing it believes the property is — which governs the treatment the property receives.
Crawler class
The behavioral and infrastructural category of an agent visiting a site; the mix of classes assigned by a platform encodes its classification verdict.

Related Field Notes

FN-001: Two Crawler Classes: Binge Ingesters and Compounding Re-Crawlers · FN-007: Three Verticals, One Curve: The Ingestion Sequence Replicates

The Field Notes Series

FN-001 — Crawler Classes: Binge vs Compounding FN-002 — The Staircase Effect, Confirmed in Search Console FN-003 — Entry-Page Decentralization FN-004 — Position 2, Zero Clicks: The Absorption Fingerprint FN-005 — Crawler Infrastructure as a Classification Signal FN-006 — The Citation Fan-Out FN-007 — Three Verticals, One Curve FN-008 — Ingested, Not Retrieved All Field Notes →
Scope of disclosure. The observation method in this series is published in full: the log fields, the ratios, the differential-diagnosis tables, and the reasoning by which each conclusion is reached. Any operator with access to their own access logs and Search Console can reproduce these tests against their own data, and is invited to. What is not published is the IEO Engine™ deployment protocol — the content architecture and sequencing that produce the outcomes being measured. The distinction is deliberate: a finding that cannot be checked is not a finding, but a method that produces the finding is an asset.
Provenance. Raw server logs (monthly Webalizer aggregates, GoDaddy shared hosting) and Google Search Console 6-month Web-search exports pulled July 10, 2026, across three independent production deployments: a local service business (live Feb 23, 2026), a B2B methodology reference site (live Apr 26, 2026), and a consumer Android application property (staged May 2026, corpus completed July 5, 2026). Figures are lightly rounded; directions and ratios are exact.
Cite as: IEO Engine Field Note FN-005 (2026). Google Tells You What It Thinks You Are — by Which Crawler Stack It Sends. https://ieoengine.com/research/fn-005-google-crawler-infrastructure-classification.html

← All Field Notes