AI visibility is not one thing; it is three sequential doors. First, reach: can the engine get into your site? Second, recognize: does it know you as an entity? Third, believe: does it trust what it says about you?
Most sites I audited have shut at least one. At the reach door, three quarters give AI bots no explicit access, and some block GPTBot, ClaudeBot and PerplexityBot outright, telling AI “do not read me” while asking why ChatGPT never recommends them.
At the recognize door, roughly 87% have no Wikidata entry, the most basic way for an engine to know a brand as an entity. At the believe door, half have no independent reputation signal (third-party review, award, industry mention), so even when the engine says something, it finds no proof behind it.
Let me be clear: this does not mean your content is bad. Content and technicals are relatively fine on most sites. The problem is whether these three doors are open. The good news: the list is short and the same for everyone; the first to open the doors pulls ahead.
Study details (transparency matters)
Scope: through the GEO Score Card program I have reviewed hundreds of sites. The percentages below come from the sites I audited in depth with the full scorecard. That audit set spans Turkey and abroad (US, UK, Germany, Estonia, Romania, South Korea), across sectors: health and medical tourism, e-commerce, B2B software, real estate, local services, media, law, manufacturing, travel.
Method: I crawled each site site-wide with Screaming Frog, checked robots.txt, llms.txt, ai.txt and AGENTS.md with Chrome, read PageSpeed Insights’ agentic-browsing layer, pulled market data from Semrush, and ran live tests in ChatGPT, Perplexity, Gemini and Google’s AI surfaces.
Honesty note: if I could not clearly observe a signal on a site, I dropped it from that item’s denominator rather than guessing. The rates are a photo of the set I audited, not of the whole industry.
Door 1: Reach. Can the engine get in?
The most basic door, and the most surprising. About three quarters of the sites I audited give AI bots no explicit access in robots.txt. A subset does something worse: it blocks GPTBot, ClaudeBot and PerplexityBot outright.
Then the same brand asks why AI never recommends it. It is like locking the door, peering through the window, and asking why nobody comes in.
Most of this is not deliberate; an old security or bandwidth rule nobody revisited. Before any AI-visibility work, the first job is to open your own robots.txt and make sure you are not accidentally turning away GPTBot, ClaudeBot, PerplexityBot, OAI-SearchBot and Google-Extended. Also confirm you are well indexed in Bing, because ChatGPT grounds heavily on Bing’s index.
A note on llms.txt, since the field overhypes it. On most sites I audited it was missing or broken. But be honest: llms.txt is not a Google ranking factor and real AI-bot usage of it is very low today. Do not make it your main deliverable. Treat it as a low-cost, low-risk extra layer for some non-Google engines; the door is opened by robots.txt access.
Door 2: Recognize. Does the engine know you as an entity?
Reaching you is not enough; the engine must recognize you as an entity. Here is the biggest shared gap: roughly 87% of the sites I audited have no Wikidata entry.
Wikidata is a primary entity database feeding Google, ChatGPT, Gemini and Copilot. If your brand is not there, the engine either never mentions you or confuses you with another entity.
The other locks on this door are shut too: about three quarters have no Wikipedia article, and roughly 30% lack even a basic Organization schema.
The fix is not expensive. A Wikidata item is free; add your founding year, founder and social profiles, then link it in your site’s Organization schema sameAs array, and you tell the engines “this is me, one entity, not scattered fragments.” Wikidata is the quiet weapon here: the item most brands never touch, and the highest-return one.
Door 3: Believe. Does the engine trust what it says?
Say the engine reached you and recognized you. One more door: does it believe what it says?
AI is no longer a signpost; it is a narrator. Users form a judgment from the sentence the engine writes, without ever clicking. If there is no proof behind that sentence, both the user and the engine tag it an “empty claim.”
In my data this door is half shut: half the sites have no independent reputation signal. No Trustpilot, no industry review, no award, no third-party validation. About two thirds also lack a citation-ready Q&A (FAQ) structure.
The rule is simple and most brands miss it: proof beats posture. The page that gets cited is not the one saying “we are the leader,” but the one documenting “we were included in this independent list, we won this award, we produced this data.” Replace your “we’re the best” lines with independent proof; what the engine trusts is not an adjective, it is a document.
Counter-thesis: “so many gaps means this is hard”
Do not read the long gap list as “GEO is complex.” The opposite.
The gaps being this shared is the most hopeful part, because the same short list keeps repeating: robots.txt access, a Wikidata entry, basic schema, one independent reputation signal. This is not rocket science.
If the industry average is this low, the reason is not difficulty; it is that these known items were simply never addressed. We are early, not late.
Is Turkey behind? No, it’s a maturity problem
The first thing I wanted to know: are local sites behind, or is this global? The data was blunt. Foreign sites have the same gaps.
An Estonian software firm has no Wikidata, a German travel site has a broken llms.txt, a few US sites have no schema at all.
So “we are not ready for AI” is not a Turkish excuse; it is a maturity stage. Even the most-ready sites owe it to their own maturity, not their country. For emerging-market brands that is encouraging: the competition abroad has not solved this either; the first to do it well pulls ahead.
What to do this week: the door to open
Check all three doors in half an hour; open the two most critical right away.
Reach: open yoursite.com/robots.txt. If you see “Disallow” for GPTBot, ClaudeBot, PerplexityBot, OAI-SearchBot or Google-Extended, remove those lines.
Recognize: search your brand on Wikidata. No record means this is the week’s highest-return task. Create a free item with the correct founding year, founder and official site, then link it in your Organization schema sameAs.
Believe: ask AI about your brand (“what is [brand], is it trustworthy”). Check whether there is proof behind the answer. If not, aim to earn at least one independent reputation signal: a real review, an industry mention, a third-party validation.
Optional add-on: if you have no llms.txt, add a simple one, but do not mistake it for the main job; the door is opened by robots.txt access.
One sentence: before you try to appear in AI, make sure the engine can reach you, recognize you, and believe what it says about you. Fancy content comes later; these three doors come first.
Join the benchmark
This study is not a one-off. I keep updating the GEO roadmap and refining the scorecard, so it is now a benchmark. Apply and you get a free GEO roadmap for your site, and you see where you stand across these three doors:
English: stradiji.com/geo-score-card
Turkish: stradiji.com/tr/geo-skor-karti
Every application makes the benchmark stronger. Thank you in advance.