How to track your brand in ChatGPT
To track whether ChatGPT recommends your brand, send the questions your buyers ask to the ChatGPT model through the OpenAI Responses API with the web_search tool enabled, then check each answer and its url_citation annotations for your brand names and domains — every day, on a prompt set you hold constant. ChatGPT has the largest reach of any AI assistant, which makes it the single most important answer surface to measure.
Last updated 2026-07-28
How do you query ChatGPT for measurement?
Through the Responses API with the built-in web_search tool switched on. That combination is what produces a grounded answer plus url_citation annotations — a structured list of the pages the model actually retrieved, rather than URLs written into the prose.
The distinction matters more than it sounds. An OpenAI completion without web_search answers from training data and will happily produce citations for pages that do not exist, which corrupts a visibility score invisibly. AI Visibility Tracker runs ChatGPT through the Responses API with web search on, reads citations only from the annotations, and reads the answer text separately for brand and competitor names. Why that separation matters is covered in why web-grounded measurement is non-negotiable.
Which model should represent "ChatGPT"?
The flagship consumer tier — the model a normal ChatGPT user actually gets. Measuring with a cheaper mini model measures a product almost nobody uses, and the resulting number is not comparable to anything.
This is a real temptation because grounded flagship calls are the most expensive part of a run. AI Visibility Tracker resolves it by fixing the measurement model to a consumer-representative choice with no free-text override, and by controlling cost elsewhere instead: prompt tiers so only core prompts run daily, same-day deduplication, a hard answer-token ceiling, and a per-run budget cap that skips a run whose projected cost exceeds it.
What does ChatGPT actually cite?
For discovery questions it leans heavily on third-party comparison articles, review directories and recent roundups; for brand-specific questions it reads your own site and major coverage. Because each answer synthesizes only a handful of retrieved pages, presence in those specific sources — not general ranking strength — decides whether you appear.
The practical consequence is that your own site is the wrong lever for "best X for Y" prompts. A vendor asserting it is the best is exactly the claim synthesis discounts; a third-party comparison naming you is the claim it repeats. How to get cited by ChatGPT works through the moves that follow from this.
What is ChatGPT-specific about reading the results?
- Highest reach of any assistant — if you weight engines by audience rather than treating them equally, ChatGPT carries the most weight in your overall score.
- Grounded and ungrounded answers differ sharply, and the difference is not random: ungrounded answers over-favour brands with a large training-data footprint. Always measure with search on.
- Answers vary between runs even with identical inputs. Daily sampling on a fixed prompt set is the only honest way to read it; a single answer is an anecdote.
- Real users' answers are personalised by memory, custom instructions and location. Your API measurement is deliberately the neutral baseline — it will not match any individual user's experience, and should not try to.
- A brand can be recommended by name with no link attached. Checking only the citation list undercounts ChatGPT visibility noticeably.
How much does daily ChatGPT tracking cost?
It depends on prompt count and model tier, but for a core set of ten to fifteen prompts it is typically a few cents per day in API usage — you pay the provider directly, at cost. Grounded calls cost more than bare completions because they carry search fees and read retrieved content.
AI Visibility Tracker shows a projected cost before every run and enforces a per-run budget cap (default $5) that skips a run projected to exceed it, so spend cannot run away while you sleep. The app itself is $29 once, no subscription, with a 7-day free trial; there is no per-prompt or per-seat meter on top.
Step by step
- 1Write down the questions your buyers actually ask — discovery ("best X for Y"), comparisons ("A vs B"), alternatives ("A alternatives"), and brand checks ("is A worth it"). Ten to fifteen core prompts is enough to start.
- 2Build a brand profile: every name your brand goes by, plus the domains that belong to you. This is what a citation is matched against, so spelling variants matter.
- 3Send each prompt to the engine with live web search enabled — an ungrounded answer reflects stale training data and can cite pages that do not exist.
- 4Check the answer text AND the structured citation list for your brand and your competitors. A mention without a link and a link without a mention are both visibility.
- 5Repeat daily on the same prompt set, and read the trend rather than single answers.
- 6Record which domains are cited on the prompts you lose. That list, ranked by frequency, is your action plan.
Frequently asked questions
Can I just ask ChatGPT whether it recommends my brand?
Not reliably. Answers vary run to run, and rephrasing the question changes the result — so a single check tells you about one sample, not about your visibility. Meaningful tracking asks the same fixed prompt set every day with web search on, then reads the rate over weeks. See what is Share of AI Answer for the formula.
Do I need API access, or can I do this in the chat app?
You need API access with web grounding enabled. The chat apps personalise answers with memory, history and location, which is the opposite of what a measurement baseline needs — and manual checking stops scaling past a handful of prompts. One OpenRouter key can carry every engine, or you can use each vendor's own API key directly.
Does ChatGPT tell you why it recommended a brand?
Not directly, but its citations are strong evidence. The pages it retrieved are the pages that shaped the answer, so the domains cited on prompts where a competitor wins and you do not are a reliable map of where the recommendation came from.