Does AI-Generated Content Get Cited? What Google Says and What the Data Shows
Yes, AI-written pages can be cited, and Google doesn't penalize content for being AI-written. Google's rule is about purpose and value: generating many pages mainly to manipulate rankings is scaled content abuse, however the pages are made. In Graphite's study, 18% of the articles cited by ChatGPT and Perplexity were AI-generated, against 14% of articles ranking in Google. What matters is the same as for any page: it can be read, it answers a real question, and it is accurate and specific.
Key takeaways
- Google says appropriate use of AI is not against its guidelines, and that using AI to generate content primarily to manipulate rankings is spam. Scaled content abuse applies "no matter how it's created."
- Ahrefs found 86.5% of top-ranking pages contain some AI content, and the correlation between AI share and position was 0.011, effectively zero.
- Graphite found 82% of articles cited by ChatGPT and Perplexity were human-written and 18% AI-generated, and AI articles tended to rank lower in Google.
- Every one of these figures depends on AI detectors, which make errors, and none of the studies could identify human-edited AI drafts.
What does Google say about AI-generated content?
Google's position has been stable since February 2023: it cares about quality, not how content is produced. Its Search Central blog said then that "Appropriate use of AI or automation is not against our guidelines. This means that it is not used to generate content primarily to manipulate search rankings, which is against our spam policies." It added: "Using AI doesn't give content any special gains. It's just content. If it is useful, helpful, original, and satisfies aspects of E-E-A-T, it might do well in Search."
Google's current guidance page, last updated October 1, 2026, says generative AI "can be particularly useful when researching a topic, and to add structure to original content." It warns that using it "to generate many pages without adding value for users may violate Google's spam policy on scaled content abuse," and that it is "critical to manually factcheck and review all AI-generated content for accuracy and trustworthiness before publishing."
What is scaled content abuse?
Google's spam policies define it as "when many pages are generated for the primary purpose of manipulating search rankings and not helping users." It is "typically focused on creating large amounts of unoriginal content that provides little to no value to users, no matter how it's created." The examples include:
- Using generative AI tools to generate many pages without adding value for users.
- Scraping feeds, search results or other content to generate many pages, with little value added.
- Stitching together content from different pages without adding value.
- Creating many pages whose content makes little sense but contains search keywords.
When Google announced the policy in March 2024, it said the policy applies "whether automation, humans or a combination are involved." It expected the update, together with its earlier efforts, to reduce low-quality, unoriginal content in results by 40%, and reported on April 26, 2024 that the reduction was 45%, by its own evaluation. Sites that violate the policy "may rank lower in results or not appear in results at all."
Do AI engines cite AI-generated pages?
Sometimes, at a rate close to how often they rank in Google. The three large studies we found:
| Study | Sample | Finding | Caveat |
|---|---|---|---|
| Ahrefs, July 2025 | 600,000 pages from the top 20 results for 100,000 keywords | 86.5% had some AI content, 4.6% were "pure AI", and the correlation with ranking position was 0.011 | Ahrefs' own detector, Google rankings only |
| Graphite, October 2025 | 31,493 keywords in 10 categories (June 2025 results), plus 100 keywords per category turned into questions for ChatGPT and Perplexity | 14% of Google-ranking articles and 18% of articles cited by ChatGPT and Perplexity were AI-generated, and only 7% of number-one results | Shows the share of citations, not a controlled test of preference |
| Semrush, April 2026 | 42,000 blog pages from the top 10 results for 20,000 keywords, November 2025 | At position 1, pages were about 80% likely to be human-written and about 10% AI-generated, and the gap narrows from position 5 | Blog URLs only, GPTZero detector, Google only |
Read together, AI-written pages are common, they are cited, and the very top of Google skews human. Ahrefs says Google "neither significantly rewards nor penalizes pages just because they use AI," and that "the very highest-ranking pages, i.e. #1, tend to have slightly less AI-generated content." Graphite's 82% to 18% split is the share of cited articles, so on its own it doesn't prove engines prefer human pages, though Graphite concludes that AI content "does not perform well in Google Search or LLMs."
Lab research points in a different direction. A KDD 2024 paper found neural retrieval models tend to rank LLM-generated documents higher, and a PNAS paper found a consistent tendency for LLMs, in binary-choice tests, to prefer options described by LLMs over ones described by humans. Neither tested live web citations, so they show a possible bias, not a ChatGPT behavior.
Can anyone reliably tell AI text from human text?
Not perfectly, and every share above rests on a detector. Ahrefs notes that "No AI content detector is perfect" and that detectors "always carry the risk of false positives." Graphite measured a 4.2% false positive rate and a 0.6% false negative rate for the Surfer detector in its own test. A 2023 paper found GPT detectors consistently misclassify non-native English writing as AI-generated. Graphite says it did not evaluate AI content with heavy human editing, and Ahrefs' detector measures how much of a page reads as AI, not how it was made. That matters, because Semrush's survey says 64% of SEOs use a human-led, AI-assisted workflow.
What actually gets a page cited?
AI is not the variable. The conditions are the same for any page:
- It can be read. For ChatGPT, OpenAI says sites opted out of OAI-SearchBot "will not be shown in ChatGPT search answers." For Google's AI features, a page must be indexed and eligible to show with a snippet, with no additional technical requirements.
- It answers a real question. Google's May 2025 guidance says to focus on "unique, non-commodity content." A page that answers the question in its first two sentences is easy to quote.
- It is specific and sourced. In the research that named GEO, adding sources, quotations and statistics raised visibility in generative engine responses by up to 40% in the authors' tests.
- Other sources agree. Mentions on other sites count too.
How should you use AI to write without hurting your visibility?
- Start from facts you own. Give the model your product details, data and quotes, not a blank prompt.
- Check every claim. Google says generative models predict a likely sequence of words and "may contain inaccuracies."
- One page per real question. Don't generate a page for every keyword variation.
- Add something only you have: original data, screenshots, customer examples, a clear opinion.
- Be honest about how it was made. Google suggests disclosing how automation was used where readers might wonder, and warns that fabricated author profiles are "a form of deception."
- Review before publishing. A person should read and approve each page.
This is how Citepoint works. It writes pages from your site, notes and cited sources, passes each page through eight checks, and lets you approve every page before it ships. It is still AI-assisted content, so the same test applies: does the page help the reader?
Frequently asked questions
Will Google penalize my site for using AI to write content?
Not for using AI itself. Google says appropriate use of AI or automation is not against its guidelines. It penalizes content generated primarily to manipulate rankings, such as many pages without value, however they are produced.
Do I need to label AI-generated content?
Google suggests adding information on how content was created where readers might wonder how it was made, and we found no Google source saying a label affects ranking. Google Merchant Center separately requires AI-generated product data to be labeled as AI-generated, and AI-generated images to carry IPTC metadata that marks them as AI-made.
Does ChatGPT cite AI-generated articles?
Yes. In Graphite's study, 18% of the articles cited by ChatGPT and Perplexity were classified as AI-generated. The study measured the share of cited articles, not a preference.
Should I list an AI as the author?
Google says giving AI an author byline is probably not the best way to make clear when AI is part of the process, and warns against fabricated creator profiles. Name the real people who review and stand behind the page.
Is human-edited AI content safe?
Google judges the result, not the process. The studies above couldn't identify human-edited AI drafts, so there is little direct data, but review, original input and fact-checking are what Google's guidance asks for.
Sources
- Google Search Central: Guidance on using generative AI content on your website
- Google Search Central Blog: Google Search's guidance about AI-generated content (February 8, 2023)
- Google Search Central: Spam policies for Google web search
- Google Search Central Blog: What web creators should know about our March 2024 core update and new spam policies
- Google: New ways we're tackling spammy, low-quality content on Search (March 5, 2024)
- Google Search Central: Creating helpful, reliable, people-first content
- Google Search Central Blog: Top ways to ensure your content performs well in Google's AI experiences on Search (May 2025)
- Google Search Central: AI features and your website
- Ahrefs: AI-generated content does not hurt your Google rankings (600,000 pages analyzed)
- Graphite: How does AI-generated content perform in search and answer engines?
- Semrush: Does AI content rank well in search? Survey and data study
- Dai et al.: Neural retrievers are biased towards LLM-generated content (arXiv:2310.20501)
- Laurito et al.: AI-AI bias, large language models favor communications generated by large language models (arXiv:2407.12856)
- Liang et al.: GPT detectors are biased against non-native English writers (arXiv:2304.02819)
- OpenAI: Overview of OpenAI crawlers
- Aggarwal et al., GEO: Generative Engine Optimization (arXiv:2311.09735)