Pith. sign in

REVIEW 3 cited by

History, Development, and Principles of Large Language Models-An Introductory Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.06853 v3 pith:K56RU5YW submitted 2024-02-10 cs.CL

classification cs.CL
keywords languagellmsmodelssurveybackgroundknowledgeprinciplesdevelopment
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Language models serve as a cornerstone in natural language processing (NLP), utilizing mathematical methods to generalize language laws and knowledge for prediction and generation. Over extensive research spanning decades, language modeling has progressed from initial statistical language models (SLMs) to the contemporary landscape of large language models (LLMs). Notably, the swift evolution of LLMs has reached the ability to process, understand, and generate human-level text. Nevertheless, despite the significant advantages that LLMs offer in improving both work and personal lives, the limited understanding among general practitioners about the background and principles of these models hampers their full potential. Notably, most LLM reviews focus on specific aspects and utilize specialized language, posing a challenge for practitioners lacking relevant background knowledge. In light of this, this survey aims to present a comprehensible overview of LLMs to assist a broader audience. It strives to facilitate a comprehensive understanding by exploring the historical background of language models and tracing their evolution over time. The survey further investigates the factors influencing the development of LLMs, emphasizing key contributions. Additionally, it concentrates on elucidating the underlying principles of LLMs, equipping audiences with essential theoretical knowledge. The survey also highlights the limitations of existing work and points out promising future directions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. AMCR: A Framework for Assessing and Mitigating Copyright Risks in Generative Models

    cs.LG 2025-08 reject novelty 4.0 of 10

    AMCR combines prompt sanitization, attention-based partial infringement detection, and a similarity-minimizing fine-tuning loss to reduce copyright infringement in text-to-image generation.

  2. Educational-Psychological Dialogue Robot Based on Multi-Agent Collaboration

    cs.CL 2024-12 reject novelty 4.0 of 10

    A multi-agent dialogue system for education and psychological counseling, with benchmark results only for the educational component.

  3. Foundation Models for Astrophysics

    astro-ph.IM 2026-08 conditional novelty 3.0 of 10

    Astronomical 'foundation models' largely reuse transformers and self-supervised pretraining, but evidence of transfer to new instruments, populations, or tasks remains rare; the paper argues such evidence, not archite...

Pith tools