Pith. sign in

REVIEW 14 cited by

A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2303.04226 v1 pith:J6E2KI3H submitted 2023-03-07 cs.AI cs.CLcs.LG

classification cs.AIcs.CLcs.LG
keywords aigccontentmodelschatgptcomprehensivegenerationgenerativeintent
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recently, ChatGPT, along with DALL-E-2 and Codex,has been gaining significant attention from society. As a result, many individuals have become interested in related resources and are seeking to uncover the background and secrets behind its impressive performance. In fact, ChatGPT and other Generative AI (GAI) techniques belong to the category of Artificial Intelligence Generated Content (AIGC), which involves the creation of digital content, such as images, music, and natural language, through AI models. The goal of AIGC is to make the content creation process more efficient and accessible, allowing for the production of high-quality content at a faster pace. AIGC is achieved by extracting and understanding intent information from instructions provided by human, and generating the content according to its knowledge and the intent information. In recent years, large-scale models have become increasingly important in AIGC as they provide better intent extraction and thus, improved generation results. With the growth of data and the size of the models, the distribution that the model can learn becomes more comprehensive and closer to reality, leading to more realistic and high-quality content generation. This survey provides a comprehensive review on the history of generative models, and basic components, recent advances in AIGC from unimodal interaction and multimodal interaction. From the perspective of unimodality, we introduce the generation tasks and relative models of text and image. From the perspective of multimodality, we introduce the cross-application between the modalities mentioned above. Finally, we discuss the existing open problems and future challenges in AIGC.

Discussion (0). Sign in to comment.

Forward citations

Cited by 14 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 404 citations worldwide. Full citation record

  1. Unlock the Potential of Fine-grained LLM Serving via Dynamic Module Scaling

    cs.DC 2025-07 conditional novelty 7.0 of 10

    CoCoServe dynamically replicates and migrates individual LLM modules across GPUs, claiming up to 46% cost reduction and 1.16x-4x throughput gains over HFT and vLLM.

  2. An Exploration of Agentic Information Fusion for Test Maintenance Prediction

    cs.SE 2026-07 conditional novelty 6.0 of 10

    MAST fuses static call graphs, BM25 lexical similarity, and LLM semantic summaries via multi-agent fusion and post-check to localize tests needing maintenance after production changes, outperforming a semantic-only ba...

  3. Private Seeds, Public LLMs: Realistic and Privacy-Preserving Synthetic Data Generation

    cs.CR 2026-04 unverdicted novelty 6.0 of 10

    RPSG generates realistic synthetic replicas of private text by combining private seeds with public LLMs and a formal differential privacy mechanism in candidate selection.

  4. Tab-MIA: A Benchmark Dataset for Membership Inference Attacks on Tabular Data in LLMs

    cs.CR 2025-07 conditional novelty 6.0 of 10

    Tab-MIA shows LLMs fine-tuned on tabular data are vulnerable to membership inference attacks, with AUROC up to 97.7% after three epochs and encoding format strongly affecting leakage.

  5. CorrDetail: Visual Detail Enhanced Self-Correction for Face Forgery Detection

    cs.CV 2025-07 conditional novelty 6.0 of 10

    CorrDetail achieves state-of-the-art face forgery detection by training a vision-language model to self-correct error-laden visual questions and combining it with a dedicated visual branch.

  6. "I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products

    cs.HC 2025-06 conditional novelty 6.0 of 10

    GAI tools' content moderation policies are comprehensive in scope but thin on user reporting and appeals, and Reddit users report frequent frustration with opaque moderation decisions.

  7. Can Smaller LLMs do better? Unlocking Cross-Domain Potential through Parameter-Efficient Fine-Tuning for Text Summarization

    cs.CL 2025-09 reject novelty 5.0 of 10

    PEFT adapters trained on high-resource summarization domains can improve Llama-3-8B's summaries on unseen domains, but the reported gains are weakened by test-set selection and missing significance tests.

  8. Generative AI for Testing of Autonomous Driving Systems: A Survey

    cs.SE 2025-08 conditional novelty 5.0 of 10

    A systematic survey that organizes 91 studies of generative AI for autonomous driving testing into six scenario-based tasks and catalogs 27 limitations.

  9. Modeling Human Responses to Multimodal AI Content

    cs.AI 2025-08 unverdicted novelty 5.0 of 10

    A 154K-post study reports that humans identify AI content best when text and images are both present and inconsistent, and offers metrics plus an LLM agent for human-aligned responses.

  10. Large Language Model Agent for Structural Drawing Generation Using ReAct Prompt Engineering and Retrieval Augmented Generation

    cs.LG 2025-07 conditional novelty 5.0 of 10

    A six-stage LLM agent chain converts natural-language descriptions of three beam types into Python code for AutoCAD drawings, with per-step success rates between 77% and 100% over 100 runs.

  11. DatasetAgent: A Novel Multi-Agent System for Auto-Constructing Datasets from Real-World Images

    cs.CV 2025-07 reject novelty 5.0 of 10

    DatasetAgent is an LLM-powered multi-agent pipeline that automatically constructs image classification, detection, and segmentation datasets from web images, with modest downstream gains shown but weak experimental controls.

  12. FlashDP: Private Training Large Language Models with Efficient DP-SGD

    cs.LG 2025-07 conditional novelty 5.0 of 10

    FlashDP fuses per-sample gradient computation, norm calculation, clipping, and noise addition into a cache-friendly block-wise all-reduce workflow that avoids explicit per-sample gradient storage and redundant recomputation.

  13. Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems

    cs.CR 2025-07 conditional novelty 3.0 of 10

    A systematic review of deepfake detection finds a pervasive lack of adversarial robustness evaluation across all modalities and calls for resilient, modality-agnostic detectors.

  14. Artificial intelligence for representing and characterizing quantum systems

    quant-ph 2025-09 unverdicted novelty 1.0 of 10

    A review organizes AI-based quantum system characterization into ML, deep learning, and language model paradigms, covering property prediction and implicit state reconstruction.

Pith tools