Table of Contents
Summary
  • Domain extension does not predict citation rate. Otterly.AI's May 2026 study of 1,028,959 cited URLs found .COM, .ORG, .IO, and .AI all averaged 1.7 citations per URL.
  • Profound's August 2024 to June 2025 dataset put .COM at 80.41% of ChatGPT citations. Rate and share answer different questions.
  • Clean URLs do matter. URLs without query strings averaged 2.1 citations against 1.6 for URLs carrying parameters.
  • Different assistants draw on different sources. Wikipedia accounts for 7.8% of ChatGPT's citations; Reddit accounts for 6.6% of Perplexity's.
summary img

More people now ask an AI assistant about a product before they ever see a search results page. That has produced a lot of advice about domains and AI visibility, much of it confident and very little of it sourced.

This article works from the studies that actually exist. We look at what four datasets say about domain extensions, URL structure, page types, and where each assistant gets its information. Where the data is clear, we empathize. Where it runs out, we say that too, rather than filling the gap with a guess.

If you are choosing a domain, consolidating a portfolio, or trying to work out why one assistant describes your company differently from another, this is what the evidence supports.

 

First Step: AI Bots Crawl Your Page

None of this works if the bots cannot reach you.

And it is not one bot. OpenAI runs three, each doing a different job:

  • GPTBot collects content used to train models.
  • OAI-SearchBot builds the index behind ChatGPT search results.
  • ChatGPT-User fetches pages live during a session.

You control each one separately in robots.txt. That means you can let OAI-SearchBot in so you show up in search results, while blocking GPTBot to keep your content out of training.

It also means blocking one tells you nothing about the other two. This is the part people get wrong.

Those three bots map onto three different things that usually get lumped together as "AI visibility":

  • Training. Content collected in bulk and associated with your domain over time.
  • Retrieval. Live or indexed fetches from specific hostnames.
  • Citation. The handful of retrieved sources an answer actually shows.

Almost all published advice targets the last one. Block the wrong bot and you have quietly cut off the first two instead, which is a different problem with a different fix.

 

Why Your Domain Works as an Entity Anchor

This section is an argument, not a finding. No published study measures how AI systems bind a brand name to a domain, so treat what follows as reasoning rather than evidence.

Your brand name gets written several ways across your own site. Your legal name differs from your trading name. Social handles vary by platform. A registered domain is unique by design and resolves to one owner, which makes it the most consistent public string pointing at your company.

Semrush's July 2025 study analyzed 5,000 randomly selected keywords and over 150,000 unique citations across four platforms: Google AI Mode, Google AI Overviews, ChatGPT, and Perplexity. It measured how closely each one's citations overlapped with Google's top 10 organic results.

screenshot shwoing overlap between ai citations and top 10 goodle search rankings

Source: Semrush

Platform Domain Overlap with Google Top 10 URL Overlap
Perplexity Over 91% 82%
Google AI Overviews Around 86% Around 67%
Google AI Mode Around 54% Around 35%
ChatGPT Around 45% Around 28%

Perplexity aligned most closely, which suggests it leans heavily on Google's top 10 when choosing what to cite.

AI Overviews also drew significantly on Google's traditional index.

AI Mode had a looser relationship, indicating more independent retrieval.

ChatGPT had the weakest overlap of the four, echoing Semrush's earlier finding that ChatGPT aligns more closely with Bing than with Google.

Notice what happens to the two columns. On every platform, domain overlap runs well ahead of URL overlap. Semrush attributes this to LLMs pulling different pages from the same trusted domains, reaching deeper into subpages, blog posts, and help articles rather than the homepage-style content that tends to rank.

They also found that as a domain's count of top 10 rankings increased, so did its presence in AI citations, with a strong correlation at the domain level. That is the closest thing to direct evidence for the argument in this section: the domain, not the page, is the unit these systems appear to trust.

 

What the The URL AI Citation Study 2026 Data Say

The study by Otterly tested something adjacent, and the result is more useful than a guess about names would have been.

It looked at the page type. Not the domain, not the extension, but what kind of page sits at the URL. Here is what came back:

Page Type Average Citations
Guide pages (/guide/ path) 2.7
Blog posts 2.0
Baseline (all pages) 1.9
News 1.7
Product/Service 1.6
Pricing pages 1.5

Source: Otterly

That is an 80% spread from top to bottom. For comparison, the spread across domain extensions was zero.

 

What to Take From This

Your name matters for how people read it, remember it, and type it. Those are real reasons to pay for a good one.

But if the question is specifically "will this name get me cited more," the honest answer is that nobody knows, and the thing we can measure points somewhere else entirely. A guide published on a plain descriptive domain outperforms a pricing page on a beautiful brandable one. What you publish carries more measurable weight than what you publish it on.

 

Does Your TLD Affect AI Citations

Short answer: no, not as a selection signal.

 

The Study

Otterly.AI ran the largest analysis of this we have seen. They looked at 1,028,959 cited URLs and 1,932,200 individual citations across six platforms:

  • ChatGPT
  • Google AI Overviews
  • Google AI Mode
  • Perplexity
  • Gemini
  • Microsoft Copilot

The data was collected over a 24-hour window in May 2026.

TLD URLs in Sample Average Citations
.COM 528,314 1.7
Other 374,657 1.7
.UK 43,331 3.0
.ORG 32,395 1.7
.NET 12,549 1.6
.IO 11,235 1.7
.AI 8,326 1.7
.CO 6,170 1.6
.EDU 6,095 1.5
.GOV 5,887 1.5

Four of the biggest extensions landed on exactly the same number. The commercial one, the nonprofit one, the tech-startup one, and the AI one.

Worth pausing on the bottom two rows. We tend to assume .EDU and .GOV carry automatic credibility. In this dataset they came in slightly below .COM.

 

One extension broke the pattern

.UK domains averaged 3.0 citations across 43,331 URLs, nearly double everything else.

The study authors said they could not explain it. Their window was 24 hours, so it’s hard to consider this as a rule.

Without prompt-level data, they could not tell whether .UK domains are genuinely favoured, whether UK-focused queries happened to dominate that day, or whether a handful of high-volume prompts skewed the average.

 

Why Other Studies Seem to Disagree

You will see figures elsewhere showing .COM utterly dominating AI citations. Those figures are real. They are also answering a different question.

Citation rate asks: does the extension predict whether a page gets cited? In the Otterly data, no. All four major extensions averaged 1.7.

Citation share asks: which extensions make up most of the citations out there? Profound, tracking 680 million citations between August 2024 and June 2025, put .COM at 80.41% of ChatGPT citations and .ORG at 11.29%. In the Otterly sample, .COM made up 51.3% of URLs by count alone.

Both are true. .COM dominates the totals because most of the web lives on .COM. That is a volume effect, not a preference.

 

What to Take From This

Pick your extension for human reasons, like:

  • How it reads on a business card
  • Whether people will mistype it.
  • What your buyers expect to see.
  • What you can defend.
  • Do not pick it for machine visibility.

The data you find in studies we mentioned, do not support those logical reasons.

 

What the Data Say About URL Structure

The Otterly tested 15 different URL attributes such as path depth, URL length, etc.. Almost all of them came back empty, which means that URL structure isn’t important for the AI visibility that much.

The Correlations

Attribute Correlation with Citations
Path depth +0.002
URL length -0.025
Domain length -0.007
Hyphen count -0.013

Every one of these is functionally zero.

In plain terms: a page sitting five folders deep got cited about as often as one at the root. A 120-character URL performed like a 40-character one. Hyphens made no difference. Neither did a short domain.

If a folder restructure is sitting in your roadmap under the heading of AI visibility, take it out.

URL structure still matters in traditional SEO, so make sure you pay attention to it.

 

The One Thing That Did Register

Clean URLs beat messy ones.

URLs without query strings averaged 2.1 citations. URLs carrying parameters or tracking tags averaged 1.6. That is a 24% gap, measured across 430,628 query-bearing URLs.

This is the single URL-level change the data actually supports, and it’s not a complicated job:

  • Set rel="canonical" pointing at the clean version of each page.
  • Keep tracking parameters out of internal links.
  • Strip them from any URL you expect other people to share or reference.

 

One More Thing Worth Knowing

Before you set targets around citation counts, look at how the numbers are distributed.

The median URL in the study was cited exactly once. The mean was 1.9. And 15.8% of URLs accounted for half of all citations.

That is a power law, the same shape you see in backlinks and social shares. Most pages get cited once and never again. A small group gets cited constantly.

So "we got cited" is not really the milestone. Getting a page into that repeatedly-cited group is.

 

Which Sources Each Platform Draws On

Being recognised by one assistant does not mean being recognised by the others. They build their picture of the world from different places.

 

The Headline Split

Profound's dataset shows each platform leaning in a different direction:

Platform Leading Source Share of That Platform's Citations
ChatGPT Wikipedia 7.8%
Google AI Overviews Reddit 2.2%
Perplexity Reddit 6.6%

Look only at each platform's ten most-cited sources, and you’ll find that Wikipedia accounts for 47.9% within ChatGPT's top ten. Reddit accounts for 46.7% within Perplexity's.

In other words, when ChatGPT reaches for a well-known source, roughly half the time it reaches for Wikipedia. When Perplexity does the same, roughly half the time it reaches for Reddit.

 

The Four Sources Everyone Uses

Some sources turn up regardless of industry.

Search Engine Land analysed 800+ domains across 11 sectors in October 2025, using responses from Google AI Mode, Perplexity, and ChatGPT search. Four domains appeared in the top 50 cited URLs of every single sector:

Source Approximate Mentions Across All Sectors
Reddit ~66,000
Wikipedia ~25,000
YouTube ~19,000
Forbes ~10,000
LinkedIn ~9,000
Quora ~8,000

If your brand has no presence on any of these, that is a structural gap. It does not matter what industry you are in.

 

Rankings Still Feed Citations

The same analysis found organic keyword breadth correlated with AI visibility at 0.41. Backlinks came in lower, at 0.37.

Neither number is decisive on its own. But the direction is clear enough: ranking for a wide range of queries gives these systems more chances to find you. Traditional search visibility has not stopped mattering. It has become an input rather than the destination.

 

What to Take From This

Three questions, in order:

  1. Where do your buyers actually research? If they use ChatGPT, reference-style coverage does more for you than community activity. If they use Perplexity, the reverse is closer to true.
  2. Are you present in the universal four? Reddit, Wikipedia, YouTube, and Forbes show up everywhere. Absence there is not something on-site work can fix.
  3. Are you ranking broadly, not just highly? Breadth correlated more strongly than backlinks. A wide footprint gives these systems more entry points.

 

How to Audit Which Domain AI Treats as Your Brand

These steps are procedural rather than evidence-based. They tell you what your current configuration is; they do not promise a citation outcome.

  1. Confirm your canonical host. Check what DNS, redirects, and canonical tags declare right now.
  2. Confirm crawler access. Verify AI crawlers can reach your primary pages, and check each separately. OpenAI's three crawlers are controlled independently, so allowing one says nothing about the others. Check CDN and WAF rules, not just robots.txt. Note that OpenAI states robots.txt changes can take around 24 hours to affect search results.
  3. Confirm identifier consistency. Your domain, brand name, and legal name should appear identically across your site, structured data, and third-party profiles.
  4. Query the platforms. Ask each assistant who your company is and record which hostname comes back.

If you don’t know where to start, Semrush offers a great course for beginners. You can learn more about AI search and test how AI platforms cite and see your brand.

 

What This Means When You Buy a Domain Name

What sits on the domain matters more than the domain. Page type moved citations by 80% from best to worst in the Otterly sample. Extension moved them by nothing at all. A guide on an ordinary name will outperform a pricing page on a great one.

Platform-specific presence matters. A brand documented in Wikipedia and absent from Reddit will read differently to ChatGPT than to Perplexity, given how differently the two source information.

Ready to get your website live and start working on the content that will get you AI cited? Use our free website builder to create your website.

 

FAQ

 

Does my domain extension affect whether AI cites my website?

Not as a selection signal. Otterly's May 2026 analysis of over a million cited URLs found .COM, .ORG, .IO, and .AI all averaged 1.7 citations per URL. Citation totals skew heavily toward .COM because most web content lives there, which is a volume effect rather than a preference. The study covered a 24-hour window, so treat it as strong but not final.

 

Should I restructure my URLs to get cited more often?

No. Path depth correlated with citation frequency at +0.002 and URL length at -0.025, both effectively zero. The one structural factor that did register was query strings: URLs carrying parameters averaged 1.6 citations against 2.1 for clean ones. Serve clean canonical URLs and put the effort elsewhere.

 

What kind of page gets cited most?

In the Otterly sample, pages under a /guide/ path averaged 2.7 citations against a 1.9 baseline, the highest of any page type tested. Blog posts came next at 2.0. Pricing pages performed worst at 1.5. The spread from top to bottom was 80%, far larger than anything URL structure produced.

 

Which sources should my brand focus on to get cited more?

Search Engine Land's analysis of 800+ domains across 11 sectors found Reddit, Wikipedia, YouTube, and Forbes present in the top 50 cited URLs of every sector. Beyond those, sector patterns diverge sharply, with healthcare citations concentrating around peer-reviewed and government sources and entertainment around community platforms. That study also found keyword breadth correlated with AI visibility more strongly than backlinks did.

/
AUTHOR
Natasa Vujovic
Marketing SpecialistNatasa is an SEO specialist and content writer at Dynadot, specializing in search optimization, keyword strategy, and domain industry trends. With a strong background in digital marketing, she helps domain investors, entrepreneurs, and businesses understand the critical intersection between SEO and domains. At Dynadot, she creates actionable guides on choosing SEO-friendly domain names, and leveraging new TLDs to increase online visibility.