Everyone Is Quoting the Same Number and Getting Different Answers

Shalin Siriwardhana

Summary

Last week, I wrote about vendor content citing sources that turned out not to exist. This is the harder version of that same. The practical question is what this changes for SEO, content quality, and AI search visibility.

A close-up shot of a person's hand holding a printed report with a large, bold number circled in red, resting on a wooden table next to a pair of glasses.

Numbers feel safe. When we see a specific ratio or a hard integer attributed to a credible source, we tend to stop questioning it. It feels like a fact, a fixed point we can use to build a strategy or justify a budget cut. But there is a dangerous gap between a number being mathematically accurate and it being meaningful.

This is currently happening with the "crawl to refer" ratio in the AI space. It is a metric that has moved from technical logs to boardroom slides with alarming speed, and in the process, it has lost the very context that makes it useful. When we strip away the "how" and "when" of a number, we aren't simplifying the data, we are distorting it. This connects with Google While Meta Reads the Web when the same signal needs a clearer operating decision. The same pattern also shows up in Working Framework, where the practical question is how the signal becomes visible.

The Math Is Not The Problem

To understand where this breaks, we first have to understand what we are actually measuring. The crawl to refer ratio is a simple comparison: how many pages did an AI platform's crawler take from your site, versus how many actual human visitors did that platform send back to you via a link?

In the old world of search, this was a symbiotic trade. A crawler indexed your site, and in exchange, it sent you traffic. AI has disrupted this. Many AI systems now provide the answer directly in the interface, meaning they still take the content but rarely send the user back to the source. The ratio is the most direct way to quantify this imbalance. A ratio of 5 to 1 is manageable; a ratio of 70,000 to 1 suggests a one sided extraction.

The interesting part is that the formula itself is transparent. Cloudflare, which popularized this metric, laid it out clearly: take the total HTML requests from a platform's user agents and divide them by the HTML requests that carried a Referer header from that platform. It is a reproducible calculation that anyone with server logs can perform.

The failure isn't in the math. It is in what happens after the number leaves the technical documentation. When a formula is this simple, people assume the result is a universal constant, forgetting that the inputs are highly volatile.

Expert Interpretation: The tradeoff here is between simplicity and accuracy. A single ratio is easy to communicate, but it collapses a complex behavioral relationship into a single digit. When you see these numbers, you should inspect whether the metric is being used as a diagnostic tool or a weapon to justify a pre determined decision.

The Hidden Variables in the Denominator

When you see different people quoting different ratios for the same AI platform, it is usually because they are using different "windows" of time. For example, one report might show Anthropic at 70,900 to one based on a single week in June, while another report from the same month shows a different figure because the date range shifted slightly.

Time windows matter immensely because crawler behavior is not linear. A platform might run a massive training pass for a few days, spiking the "crawl" side of the ratio, and then go quiet for a month. If you look at a rolling 28 day average, you get one story. If you look at a specific week of heavy indexing, you get another. Cloudflare noted that Google's ratio shifted by nearly 20 percent in a single week simply because of a change in how GoogleBot was crawling.

Then there is the issue of which bots are being counted. Most analyses group "training crawlers" and "user request crawlers" together under one platform name. This is a critical distinction. A training crawler is designed to ingest data at scale and will never send a referral. A user request crawler fetches a page because a human is currently asking a question and might actually click a citation.

By blending these two, the ratio becomes a measure of "platform activity" rather than "platform value." If you treat them as one, you are attributing the behavior of a vacuum cleaner to the behavior of a librarian.

Expert Interpretation: This is where most analysts fail. They treat the ratio as a static property of the AI company rather than a snapshot of a specific behavior. If you are making decisions based on these numbers, you must ask if the data distinguishes between training and inference traffic. If it doesn't, the number is effectively useless for predicting future traffic.

The Gap Between Referrals and Reality

The most significant distortion comes from the fact that the ratio doesn't actually track "referrals", it tracks "referrals that announced themselves."

For a visit to be counted in this ratio, the request must include a Referer header. However, many native apps (like the Claude app or other AI mobile interfaces) do not send this header. This means a huge portion of the traffic these platforms send back to publishers is invisible to the calculation. The ratio, therefore, likely overstates the imbalance. We know the imbalance exists, but we don't actually know by how much because the "referral" side of the equation is undercounted.

there is a layer of human judgment involved in what gets excluded. Cloudflare, for instance, excludes certain Google network traffic because they view "prefetching" (where a browser predicts a link you might click) as non human consumption. While this is a logical choice, it is still a choice. Another analyst could include that traffic and arrive at a completely different number without being "wrong" in a mathematical sense.

Expert Interpretation: The danger here is the "invisible denominator." When the data source admits that the distortion is "unclear by how much," the number ceases to be a measurement and becomes an estimate. You should inspect whether your organization is treating an estimate as a hard fact.

How Context Vanishes During Retelling

The path from a technical blog post to a corporate slide deck is a process of aggressive compression. The original source usually includes the caveats: the specific date range, the warning about native apps, and the distinction between bot types.

But as the number is shared, the caveats are stripped away because they complicate the headline. The first person to summarize the data drops the native app warning. The second person drops the date range to make the number more "quotable." The third person sees a clean integer attributed to a credible source like Cloudflare and repeats it with total confidence.

Within four steps, a nuanced technical observation has been transformed into a bare integer. It looks like a fact, but it has been stripped of the only things that made it usable. This isn't necessarily fraud or error; it is just the natural tendency of information to simplify as it moves up the chain of command.

Expert Interpretation: This is a systemic risk in data reporting. The more "credible" the original source, the less likely people are to question the number once it has been detached from its context. The tradeoff is speed of communication versus accuracy of insight.

The Real World Cost of Bad Data

If these numbers were just for academic curiosity, the lack of context wouldn't matter. But they are being used to drive high stakes business decisions.

Publishers are using these ratios to decide which AI crawlers to block via robots.txt. If a publisher sees a staggering crawl to refer ratio, they may decide to shut the door entirely. However, if that ratio is inflated because it blends a training bot with a citation bot, the publisher might be blocking the very mechanism that could have sent them traffic in the future.

Similarly, marketing teams are using these figures to argue that AI referral traffic is a dead end and not worth pursuing. Once a company decides that a channel is "worthless" based on a flawed number, it is very difficult to reopen that conversation. The decision becomes a part of the corporate record, and the flawed data is rarely revisited.

When you make a semi permanent policy based on a snapshot of a training pass from a single week in June, you aren't managing risk, you are gambling with your visibility.

Expert Interpretation: The risk here is "path dependency." A wrong decision made today based on a stripped down number creates a trajectory that is hard to change. Before blocking a crawler or abandoning a channel, you must verify if the data represents a permanent state or a temporary spike.

Developing a Filter for "Usable" Numbers

This problem isn't unique to AI crawlers; it applies to almost every metric we track. A number without its window, its grouping, and its collection boundary is not actually a number. It is a shape that looks like one.

The window (the time period), the grouping (what was bundled together), and the boundary (what was excluded) are not "extra details" or "caveats." They are the components that build the number. When you separate a figure from these three things, you haven't simplified the data; you've made it unusable while keeping it looking usable. This is more dangerous than a number that is obviously wrong, because obviously wrong numbers get challenged.

The next time someone presents you with a measurement, you can run a simple, low cost test. Ask three questions:

What specific time period does this cover? What exactly was grouped together to produce this result? Where did the collection stop (what was excluded)?

If the person providing the number cannot answer these immediately, they don't actually know what the number means. They are simply quoting a shape. By insisting on these three answers, you aren't being difficult; you are ensuring that the decisions your organization makes are based on reality rather than a compressed retelling of a technical log. A useful companion note is Not the Answers), because it looks at a nearby part of the same system.

Expert Interpretation: The goal is to move from passive consumption of data to active interrogation. The tradeoff is that this takes more time and can be socially awkward in a meeting. However, the cost of a few uncomfortable questions is significantly lower than the cost of a permanent strategic error.

Comments

Comments are reviewed before they are published. Links are not allowed inside comments.

Only your name, optional LinkedIn profile, and comment will be shown.