Pith. sign in

REVIEW 7 cited by

Generative Engine Optimization: How to Dominate AI Search

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2509.08919 v1 pith:5ZSFY6RB submitted 2025-09-10 cs.IR cs.AIcs.CLcs.LGcs.SI

Generative Engine Optimization: How to Dominate AI Search

classification cs.IR cs.AIcs.CLcs.LGcs.SI
keywords searchgenerativeengineoptimizationanalysisbiascontentcritical
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The rapid adoption of generative AI-powered search engines like ChatGPT, Perplexity, and Gemini is fundamentally reshaping information retrieval, moving from traditional ranked lists to synthesized, citation-backed answers. This shift challenges established Search Engine Optimization (SEO) practices and necessitates a new paradigm, which we term Generative Engine Optimization (GEO). This paper presents a comprehensive comparative analysis of AI Search and traditional web search (Google). Through a series of large-scale, controlled experiments across multiple verticals, languages, and query paraphrases, we quantify critical differences in how these systems source information. Our key findings reveal that AI Search exhibit a systematic and overwhelming bias towards Earned media (third-party, authoritative sources) over Brand-owned and Social content, a stark contrast to Google's more balanced mix. We further demonstrate that AI Search services differ significantly from each other in their domain diversity, freshness, cross-language stability, and sensitivity to phrasing. Based on these empirical results, we formulate a strategic GEO agenda. We provide actionable guidance for practitioners, emphasizing the critical need to: (1) engineer content for machine scannability and justification, (2) dominate earned media to build AI-perceived authority, (3) adopt engine-specific and language-aware strategies, and (4) overcome the inherent "big brand bias" for niche players. Our work provides the foundational empirical analysis and a strategic framework for achieving visibility in the new generative search landscape.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Prominence-Stratified Failure Modes in Retrieval-Augmented Commercial Recommendation: A 37,000-Run Audit

    cs.IR 2026-05 unverdicted novelty 7.0

    A large-scale audit of AI commercial recommendations reveals tier-specific failure modes: L1 brands reach recommendations but convert at 25-41%, L2 convert highest at 37-52%, L3 is an inflection point, and L4/L5 brand...

  2. How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews

    cs.IR 2026-04 unverdicted novelty 7.0

    AI Overviews and Gemini retrieve substantially different sources than traditional Google search (Jaccard similarity <0.2), favor Google-owned content, appear for 51.5% of queries especially controversial ones, and are...

  3. Deep-Research Agents Can Be Poisoned via User-Generated Content

    cs.CR 2026-05 unverdicted novelty 6.0

    Deep-research agents have a concentrated attack surface because they repeatedly retrieve the same UGC pages, allowing a single poisoned page to affect citations and entity promotion across query clusters in systems li...

  4. From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms

    cs.IR 2026-04 unverdicted novelty 6.0

    A measurement study of 602 prompts across ChatGPT, Google AI Overview, and Perplexity finds that citation selection breadth and absorption depth diverge, with high-influence pages being longer, structured, and evidence-rich.

  5. How Large Language Models Source Brand Reputation Across Languages and Markets

    cs.IR 2026-06 unverdicted novelty 5.0

    LLMs cite third-party domains for 85.7% of brand attributions, with Wikipedia dominant in most languages, a long-tailed domain distribution, and market-specific shifts such as YouTube and HR sites in Poland.

  6. Can It Reach the Generator? Investigating the Survival of Prompt-Injection Attacks in Realistic RAG Settings

    cs.CR 2026-05 unverdicted novelty 5.0

    Gradient-based and instruction-override prompt injections largely fail to survive retrieval and reranking in realistic RAG systems, while only LLM-driven injections remain effective end-to-end, and all attacks are det...

  7. MaxShapley: Towards Incentive-compatible Generative Search with Fair Context Attribution

    cs.LG 2025-12 unverdicted novelty 5.0

    MaxShapley computes fair document attributions in generative QA by reducing Shapley value calculation to polynomial time via a max-sum utility, matching exact Shapley quality on HotPotQA, MuSiQUE, and MS MARCO while u...