In early August, an search engine optimisation referred to as Anastasia Kourou seen queries in her Search Console report that didn’t seem like searches: “Sure.” “Sure go on.” “Sure, pricing.”
She posted a screenshot and requested John Mueller whether or not Search Console was monitoring what individuals say to AI.

His reply confirmed it.
Search Console consists of AI Overviews and AI Mode information in the common efficiency report, and Google’s documentation explains the mechanism. A follow-up question inside AI Mode counts as a brand new query, and the whole lot in the response will get attributed to it. When somebody tells the AI “sure go on” and your web page seems in what comes again, Search Console data an impression to your web page towards the question “sure go on.”
Search Engine Roundtable covered the thread on August 6, and Ross Tavendale requested the query everybody was circling. If Search Console is recording individuals’s responses to AI Mode, how will we reverse engineer it?

No one answered him.
This put up is my reply.
I pulled each one among these fragments out of 16 months of my very own Search Console information, labored out that they arrive in seven recognizable varieties, and constructed the classifier into my free Search Console MCP so you may run it on your web site with one immediate. Sorting them is what makes the leak usable.
The entire thing comes down to 4 concepts.
- Google’s new AI report hides queries, however the queries leak into the report you already use.
- The leaked fragments are available in seven recognisable varieties, and a classifier can kind each one.
- What you may see is a flooring. Most of the dialog is in impressions Google anonymises.
- Joined with the new Generative AI report’s export, the fragments flip right into a page-level image.
I discovered 1,127 queries and 20,300 impressions throughout 16 months on my web site, towards thousands and thousands of strange impressions. Small, and each row of it is one thing an actual session mentioned or ran.
The Report Everybody Requested For Is Lacking Its Queries
Google launched Generative AI performance reports in June and, as of August 11, they’re reside for everybody. The report reveals how typically your pages seem inside AI Overviews and AI Mode. It has five data views. Impressions, pages, nations, gadgets, dates.
Queries and clicks are the two things it leaves out, and there’s no API assist both. I checked that final half on my very own property simply to make sure. The Search Analytics API’s type parameter nonetheless ends at googleNews, the searchAppearance dimension returns nothing AI-related, and the BigQuery bulk export schema doesn’t have an AI column. The export button in the UI is the solely method this information leaves Google.
So, the report tells you the way a lot AI visibility you could have and refuses to say for what. (I believe everyone knows why lol) In the meantime the strange efficiency report, the one you’ve been studying for years, has been choosing up AI dialog fragments the entire time. No one filters them out as a result of formally they’re simply queries.
Concept 1: Your Question Report Is Half Dialog Log
AI Mode seems like a chatbot. Beneath, each message is processed as a Google search, together with the follow-ups, and Google folds all of it into the web search type alongside the basic 10 blue hyperlinks. Your question report now holds two various things. Searches individuals typed, and fragments of conversations individuals had with a mannequin that occurred to present your web page.
The place information proves it’s the second factor. My web site reveals common place 4.5 for the question “sure.” On the open internet, that rating is inconceivable; “sure” belongs to songs and grammar websites. Inside an AI response, it is smart, as a result of Google’s documentation says links in an AI Overview inherit the position of the whole block, and AI Mode citations get counted below the identical guidelines as soon as they scroll into view. Place 4.5 on a reply phrase means my hyperlink sat inside the reply block, not on a outcomes web page.
That’s the leak. The query is whether or not the fragments will be advised aside from regular queries at scale.
Concept 2: The Fragments Come In 7 Sorts
I pulled 16 months of queries from my property and categorised the whole lot that couldn’t be a typed search. Seven varieties got here out, and every one has a distinct origin.
Notice: I haven’t included all of my queries for apparent causes, only a pattern. Once you run this on your personal GSC account, you’ll get all your information.
Reply Artifacts
Naked replies: “sure,” “certain,” “actually?,” “present me.” An individual answered the AI mid-conversation; the reply was processed as a search, and your web page appeared in the response. Positions right here come from the reply block, not a outcomes web page.

Pivot Observe-Ups
Mid-conversation comparisons: “what about resend?,” “what about gemini,” “how about in chinese language?” The particular person has a solution in entrance of them and asks the AI to check another, and the various they identify is the one they care about.

Conversational Questions
Questions addressed to somebody relatively than typed at a search field: “are you able to jailbreak meta raybans,” “how do i promote it,” “is it free.” The giveaway is grammar that solely works with a listener, which will get its personal part beneath.

Tracker Probes
Artificial prompts from AI visibility instruments, run on a schedule. Two signatures in my information. Prompts ending “. my location is usa.” and prompts in the type “consider the [company] on [facet].” No one searches the identical sentence each day for 2 months, so these are software program.

Agent Harnesses
A machine’s full directions, logged entire: “search the internet for… return the 3 most related outcomes you truly discovered … do not invent outcomes or urls.” Someplace an engineer wrote a immediate template, and Google filed it as a question.

Pasted Strings
Error messages and spreadsheet headers searched as-is, by individuals or pipelines. My information features a rank tracker’s full CSV column header as a single question.

Lengthy Uncategorized
Ten or extra phrases with no different marker. Some are quoted sentences, some are brokers, some are individuals. The classifier recordsdata them for overview as an alternative of guessing, which is the correct amount of confidence for this bucket.

The classifier is a ladder of checks, utilized so as, and a question stops at the first one it matches.
Take “what about resend?” It isn’t on the reply listing, so it passes the first rung. It matches the second, a pivot, “what about” adopted by a brief noun. Classification achieved, and the row tells me the relaxation.
One impression on my newsletter cost post, place 10, one click on. Someplace, somebody was asking an AI about e-newsletter instruments, requested “what about resend?” obtained my self-hosted setup in the response, and clicked by way of. That’s an individual mid-decision, and the question names the various they cared about.
There’s no machine studying in any of this. The patterns are a curated listing anybody can learn and disagree with, which is the level. Each classification is explainable.
How Is This Totally different From A Lengthy-Tail Question?
The apparent objection to all of this is that lengthy queries existed before AI. [how to get not provided keywords in google analytics] is 9 phrases, and no person mentioned it to a chatbot. The break up comes down to who the question is addressed to.
Once I first confirmed this to my co-founder Andy, that was his query.
An extended-tail question is an in depth request addressed to no person. It has its personal topic and names its personal instruments.
A conversational question is addressed to somebody, and 4 alerts give it away.
| Sign | Instance from my information | Why it might’t be lengthy tail |
|---|---|---|
| Instructing an assistant | “give me step-by-step” | No one instructs a search field |
| First particular person context | “i’m utilizing lmstudio” | You don’t transient Google about your setup |
| Dangling pronouns | “is it free”, “does it work” | “It” has no referent in the question. The referent lives in a dialog |
| Politeness | “please make clear” | No one says please to an enter subject |
Any of these 4, at any size, and the question is conversational.
Size alone is the weak sign, so it will get quarantined.
Ten or extra phrases with query syntax will get categorised as conversational, as a result of typed queries common two to 4 phrases and nearly no person sorts 11.
Ten or extra phrases with no different marker goes to the overview pile as an alternative of right into a declare. And something at 9 phrases or fewer with none of the alerts is handled as an strange search, which is why the not supplied question above by no means enters the dataset.
One inform validates the boundary, and I’m not utilizing it but. Repetition. That not supplied string has 1,853 impressions on my web site as a result of hundreds of individuals kind the identical long-tail question.
Conversational strings nearly by no means repeat; most of mine sit at one to three impressions, as a result of no two individuals phrase a follow-up identically. Impressions per string as a typed versus spoken sign is the apparent second model of this classifier.
What 16 Months Of My Information Exhibits
Begin with the timeline. Reply artifacts on my web site, by month. Zero impressions from April to November 2025. A primary flicker in December. Then March 2026 switches the class on, and it’s run at 20 to 30 impressions a month since.

ProTip: Go to Search console → Search outcomes → Add filter → Question → Choose customized.
Add this regex → ^(sure|yeah|okay|okay|certain)[?!.,]*$
You possibly can see all of the artifacts.
A question class that didn’t exist for eight months after which turns into persistent is a habits change with a date on it. “Yes” alone has 110 impressions, six clicks, and a mean place 4.5 throughout the window. Six individuals replied to Google’s AI after which clicked by way of to my web site off the again of their very own “sure.”
The pivots line up with my posts one for one. [what about claude] hit my WebMCP guide.
[what about xcode?] hit the Xcode post. [what about wayback machine] hit the Wayback guide.
Each is a reader asking the AI to examine one thing towards the factor my web page covers.
The conversational bucket is the greatest, and it has a signature. 559 queries, 8,834 impressions, 13 clicks.
That ratio is what being learn inside solutions seems like.
My greatest single row is [do ai crawlers like gptbot support content negotiation for markdown] at 2,998 impressions, and my favourite is [what ai search monitoring tools let me combine ga4 session data, gsc click data, and ai citation rates in a single analysis so i can understand the full search picture?].
Sixty-five impressions of somebody asking an AI for the product class I used to be constructing whereas they requested.
Then there are the machines.
124 of my queries are tracker probes with 2,902 impressions, and the side matrix aimed toward my Roam overview ran each day for 59 consecutive days.
The agent harnesses whole 2,181 impressions, and one among them ran roughly 2,160 instances throughout 5 days in July, the Dubai climate immediate from the listing above.
Each run surfaced my ChatGPT teardown in its outcomes, presumably as a result of the put up is about looking the internet for sources. I nonetheless don’t know whose hallucination checker that was. If it was yours, the climate in Dubai is scorching. (41 levels right now, ugh)
Three rows for the highway. Somebody pasted a full rank tracker CSV header, 17 columns of it, and it earned 146 impressions as a question.
Another person pasted X Help’s total rejection e mail, the one that claims the username is obtainable, and Google filed it towards my deal with put up at place 1.3. And a single impression exists for [you didnt give me the link], a consumer complaining at the AI, recorded by Google, filed towards my Wayback put up.
Google recorded an individual dropping an argument with a robotic haha. Basic!
Concept 3: What You Can See Is The Tip
Earlier than you run this on your web site, the important limitation. Google anonymizes uncommon queries, and conversations are nearly by definition uncommon strings. No one phrases a follow-up the method you do. So the fragments that survive into the report are the repeated ones, and the bulk of the dialog pool is hidden.
My BigQuery export reveals that over the final 59 days, 57.7% of my internet impressions carry no question string in any respect. 454,720 anonymized towards 333,651 seen. The 1,127 categorised queries are the seen tip of that pool. Learn each quantity on this put up as a flooring.
3 Fast Caveats Earlier than You Run This
The classifier judges how a question seems, not what the particular person meant. [is it agent ready] seems like a dangling pronoun aimed toward an assistant, and it in all probability is, however someone might have typed it at my agent readiness guide intentionally. Single rows are proof, not statistics.
AI Overviews and AI Mode can’t be separated at the question degree. Google studies each below internet search, so a fraction tells you a dialog occurred, not which floor hosted it.
And the sample library is English. A conversational question in Tamil or German solely will get caught if it journeys the size rung or a software signature. The one Spanish probe in my information was caught by its “. my location is spain.” suffix, not its phrases. The multilingual mannequin at the finish of this put up is constructed to shut that hole. It isn’t out but.
Run It On Your Web site
The classifier is genai_conversation_queries, a free software inside my Search Console MCP, launched right now as v2.4.0.
Should you don’t have the MCP but, one command units it up.
npx -y suganthan-gsc-mcp setup
The wizard indicators you into Google, verifies the reference to a reside name, helps you to decide your property from an inventory, and writes the Claude config for you. The full setup guide covers Claude Desktop, Claude Code, and the guide routes with screenshots. It’s free and open supply, and the server runs on your machine, so your information strikes between you and Google and passes by way of no person else’s servers.
Should you already use it, the default config begins the server by way of npx, which picks up the newest printed model.
Restart Claude, and also you’re on v2.4.0. Should you put in the one-click desktop bundle as an alternative, obtain the new one from the releases page.
Both method, the subsequent step is one sentence.
Run genai_conversation_queries on my web site
One name pulls 16 months of your queries by way of Google’s regex filter, classifies each match down the ladder, attaches the touchdown pages, and returns the seven buckets with the month-to-month artifact timeline. BigQuery isn’t required; there are no new permissions past the Search Console entry you already granted, and your information goes to Google and again like each different software in the server.
Should you run the BigQuery version of the MCP, v4.1.0 provides gsc_genai_conversation_queries, the identical classifier over your bulk export. That twin has no API row limits, which issues on giant websites, and it studies the anonymised break up, so that you get your personal iceberg proportion subsequent to your fragments.
John Mueller’s Tip (New)
I awakened to this good remark from John Mueller on my LinkedIn put up about it. He instructed organising the BigQuery information export, as a result of it would floor extra of those queries.

He’s proper for big websites, and the cause is how a lot information every route out of Search Console arms over. The UI export stops at 1,000 rows per desk. The API goes far deeper however tops out round 50,000 rows a day per search kind. The majority export has no row cap in any respect. On a giant property, the tail the API drops is the place these fragments reside, since conversational strings are uncommon and nearly by no means repeat.
On a web site my measurement, the routes agree. I checked on August 14. My busiest day in the final 4 weeks held 988 distinct seen queries, about 2% of the each day API ceiling. The classifier discovered 625 dialog queries through the API versus 621 by way of the export, with the export lagging solely as a result of its information stops two days earlier. Small properties get an identical solutions from both route. If yours is large, run the BigQuery twin, which is the level John was making.
The pool from Concept 3 stays out of attain both method. Anonymized impressions sit in the export as rows with no question string, so BigQuery tells you the way large the hidden pool is and retains the strings to itself.
If in case you have already linked your Search Console property with BigQuery, then you need to use my MCP to pull all of the queries with the identical immediate.
You possibly can entry my BigQuery MCP here.
Concept 4: Add The Generative AI Report Export
The software you simply ran reveals what individuals ask. The Generative AI report reveals how a lot of your visibility sits inside AI solutions, web page by web page. Becoming a member of them offers you each numbers for each web page.
The be part of wants one guide step from you, and right here’s why. Google nonetheless doesn’t present this information through the API, which I verified earlier on this put up, so the software has no method to fetch it. The export button in the UI is the solely method.
5 steps, about two minutes.

- In Search Console, open Efficiency, then Generative AI.
- Set the date vary to the full 16 months.
- Press Export, high proper, and select Obtain CSV. It downloads as a zipper.
- Unzip it. The file the be part of wants is Pages.csv, your AI impressions per web page.
- Connect Pages.csv to the Claude chat the place you ran the software, and ask.
Map this export towards my genai_conversation_queries outcomes. For every
web page, present its AI impressions subsequent to the fragments Google recorded on it.
Now each web page has two numbers. How a lot of its visibility is inside AI options, and which precise dialog fragments Google recorded towards it.

The overlap offers you a directional trace I haven’t seen wherever else. (Everyone knows all the search engine optimisation instruments are going to magically provide you with this concept quickly LOL.)
Reply artifacts and pivots come from AI Mode particularly, as a result of that’s the place follow-up conversations occur, whereas the export counts AI Overviews and AI Mode collectively. So a web page with excessive AI visibility and wealthy fragments skews in the direction of AI Mode conversations, and a web page with excessive AI visibility and no fragments skews in the direction of one-shot AI Overview citations. It’s an inference, not a measurement, but it surely’s the solely wedge anybody has right into a break up Google received’t present.
When Google provides an API for this report, I’ll replace the MCP to totally automate this course of.
What To Do With What You Discover
Every bucket desires a distinct response, and two of them need you to do nothing in any respect.
Pivot follow-ups are content material directions. [what about resend?] on my e-newsletter put up means the comparability readers need isn’t in the put up. Writing the Resend part solutions a query individuals are already asking an AI about my web page.
Conversational questions with impressions and no clicks present which pages get learn inside solutions. For these, title tweaks do little or no, as a result of the particular person by no means sees a title. What travels into an AI response is the passage that solutions the query, the first paragraph below a heading, the desk, the named reality. That’s the half to work on, and it’s the identical conclusion I preserve reaching from the ChatGPT side of this analysis.
Reply artifacts are corroboration, and a warning. They show pages reside inside multi-turn conversations, they usually’ll wreck a naive evaluation if left in. A key phrase software that doesn’t find out about this class will finally advocate optimizing for the question “sure.” Exclude the machine buckets, the probes and harnesses, from any alternative evaluation, and deal with artifacts as a sign relatively than demand.
And the probes deserve a minute of your consideration on their very own. If you run an AI visibility tracker, a few of these each day prompts are yours, independently logged by Google, which makes your Search Console a free audit of what your tracker truly runs. Should you don’t run one, third events are sweeping subjects you rank for, each day, and you’ll watch them do it.
Which is the common level about all seven buckets. These are prompts from actual classes the place Google’s AI reached to your pages. Each query fan-out tool, mine included, generates artificial questions and hopes they resemble actuality. This is actuality, small pattern, flooring and all. Use it as a information to how individuals phrase issues at your content material, not as a quantity dataset.
The Reply To The Query
Ross requested how we reverse-engineer Search Console recording individuals’s responses to AI Mode.
The reply turned out to be a sample library, a strict order of checks, and the discovery that the recording consists of greater than individuals.
It consists of the AI trade’s personal equipment too, tracker probes and agent harnesses displaying up in everybody’s question studies.
Run the software, learn your fragments, and go take a look at what your web site says in the passages the AI retains selecting. The conversations are already occurring. Google’s been taking minutes.
Why Google Isn’t Exhibiting Question Or Click on Information
I do know this is a small dataset with leaked fragments. Even so, there are barely any clicks, regardless of having first rate impressions and excellent common positions.
Which solutions why Google isn’t together with this in Generative AI studies.
I doubt they’ll ever add it, however solely time will inform.

The Multilingual Mannequin
The sample listing has edges. It solely works in English, and on the border between a conversational question and an strange lengthy one, it makes rule-based guesses. Each issues need machine studying relatively than extra regex, so I’ve already educated a mannequin.
It understands languages the patterns can’t, together with [kannst du mir das bitte zeigen], and it handles the border instances the guidelines guess at. (Don’t ask what it is as a result of I image a German politely asking me if I would like some espresso.)

I’m not releasing it but. It’s at the moment educated on some actual and artificial information.
For instance, on a recipe web site, [can you freeze cooked chicken] is an everyday search that appears like a dialog. So I’m testing it towards different niches and making it higher first.
The plan is a free software the place you drop your export in and get your buckets again, with the classification working in your browser so your queries by no means depart your machine. (As common, respecting your privateness.)
It’s coming after I’m pleased with the output.
Extra Sources:
This put up was initially printed on Suganthan.
Featured Picture: Ball SivaPhoto/Shutterstock
Disclaimer: This article is sourced from external platforms. OverBeta has not independently verified the information. Readers are advised to verify details before relying on them.