Pith. sign in

REVIEW 5 cited by

News Source Citing Patterns in AI Search Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2507.05301 v1 pith:BMIWSHVR submitted 2025-07-07 cs.IR cs.CLcs.CY

News Source Citing Patterns in AI Search Systems

classification cs.IR cs.CLcs.CY
keywords newssearchsystemssourcespatternscitationcitationscited
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

AI-powered search systems are emerging as new information gatekeepers, fundamentally transforming how users access news and information. Despite their growing influence, the citation patterns of these systems remain poorly understood. We address this gap by analyzing data from the AI Search Arena, a head-to-head evaluation platform for AI search systems. The dataset comprises over 24,000 conversations and 65,000 responses from models across three major providers: OpenAI, Perplexity, and Google. Among the over 366,000 citations embedded in these responses, 9% reference news sources. We find that while models from different providers cite distinct news sources, they exhibit shared patterns in citation behavior. News citations concentrate heavily among a small number of outlets and display a pronounced liberal bias, though low-credibility sources are rarely cited. User preference analysis reveals that neither the political leaning nor the quality of cited news sources significantly influences user satisfaction. These findings reveal significant challenges in current AI search systems and have important implications for their design and governance.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Generative Engine Optimization at Scale: Measuring Brand Visibility Across AI Search Engines

    cs.IR 2026-06 unverdicted novelty 7.0

    Large-scale analysis of AI engine responses shows brand visibility in three tiers (73%, 44%, 11%) with corporate sites and best-of listicles as top cited sources.

  2. Synthetic Sources?: Auditing Generative Search Engine Citations for Evidence of AI-Generated Sources

    cs.IR 2026-05 unverdicted novelty 6.0

    Audit of ChatGPT, Copilot, Gemini and Perplexity finds ~16% of cited sources are AI-generated across 712 queries on politics, health and environment.

  3. From Citation Selection to Citation Absorption: A Measurement Framework for Generative Engine Optimization Across AI Search Platforms

    cs.IR 2026-04 unverdicted novelty 6.0

    A measurement study of 602 prompts across ChatGPT, Google AI Overview, and Perplexity finds that citation selection breadth and absorption depth diverge, with high-influence pages being longer, structured, and evidence-rich.

  4. How Large Language Models Source Brand Reputation Across Languages and Markets

    cs.IR 2026-06 unverdicted novelty 5.0

    LLMs cite third-party domains for 85.7% of brand attributions, with Wikipedia dominant in most languages, a long-tailed domain distribution, and market-specific shifts such as YouTube and HR sites in Poland.

  5. Divergent Recommendations, Convergent Diagnoses: Cross-Provider Failure-Mode Convergence in AI Commercial Recommendation

    cs.CY 2026-05 unverdicted novelty 4.0

    Two major AI providers diverge in which brands they recommend but converge on classifying the failure reasons, especially for low-prominence brands.