Pith. sign in

REVIEW 3 major objections 6 minor 2 cited by

Large Language Models for Agent-Based Modelling: Current and possible uses across the modelling cycle

T0 review · 3 major / 6 minor · reviewed 2026-08-06 · deepseek-v4-flash

Pith's one-line read This paper claims that LLM use in agent-based modelling has so far clustered almost entirely in the implementation phase, with 91% of reviewed papers using LLMs there, and that the rest of the modelling cycle remains largely unexploited.

desk verdict A useful, transparent phase-by-phase map of LLM use in ABM, but the '91% implementation' concentration is a claim about a narrow Scopus-only corpus and should be scoped accordingly. read the letter →

arxiv 2507.05723 v1 pith:EDZWP5VN submitted 2025-07-08 cs.MA

classification cs.MA
keywords LargeLanguageModelsAgent-BasedModellingCycleRapidLiteratureReviewSocialSimulationLLM-poweredagentsproblemformulationverificationandvalidation
verification ladder T0 review T1 audit T2 compute T3 formal

The pith

A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.

The reading

The paper sets out to answer two questions: where and how are large language models actually used in agent-based modelling today, and where and how could they be used across the entire modelling cycle. A rapid review of 22 papers found that 91% of current uses sit in the implementation phase, mostly as LLM-powered agents that reason, deliberate, or communicate, while only two papers touch other phases. The remainder of the paper walks through the modelling cycle phase by phase, and for each phase lists concrete potential LLM uses, associated pitfalls, and mitigations. The authors conclude that LLMs are best seen not as suppliers of finished model components but as interpretive tools that provoke rethinking of system conceptualisations.

What carries the argument

The organising device is the agent-based modelling cycle, taken from reference [41] and elaborated with reference [34]. The cycle divides modelling into problem formulation, system analysis, conceptualisation, implementation, verification, validation, interpretation and communication, and documentation. The paper uses this structure twice: first as a coding scheme for the literature review to locate where LLMs are used, and then as a template for a systematic opportunity map. The same structure also generates the pitfalls and mitigations, since each phase imposes different demands on text, transparency, and domain knowledge.

What would settle it

A comparable systematic search across several other bibliographic databases using the paper's own inclusion criteria, finding more than a small number of implemented LLM uses in problem formulation, system analysis, or interpretation phases, would directly weaken the claim that 91% of current use is in implementation; the search is easy to run because the paper gives its full query in an annex.

Watch

Extended reading notes

Core claim

The central discovery is empirical and structural. After coding 22 papers retrieved from a single literature database, the authors report that 20 of them (91%) use LLMs at the implementation stage, and that this use is almost always to power agents with reasoning, decision-making, or communication abilities. Only one coded paper focuses on code generation and one on interpreting model results. From this concentration the paper argues that the field has left most of the modelling cycle unexploited, and it offers a phase-by-phase map of opportunities, from problem formulation and system analysis through conceptualisation, verification, validation, interpretation, and documentation, each with risks and mitigations.

Load-bearing premise

The whole analysis rests on a literature review that searched one database with one query as of March 2025, together with the authors' own group discussions for the forward-looking claims; if the search missed substantial work, or the group's judgment skews toward their own interests, both the 91% concentration result and the priority map could shift.

Editorial extensions

If this is right

  • If current use is this concentrated, then the largest set of untested LLM applications lies outside implementation, and the paper's phase-by-phase map functions as a research agenda.
  • Modelers who treat LLM outputs as provisional and triangulate with experts and data get a way to reduce the hallucination and bias risks the paper catalogues.
  • The paper's reframing, where LLMs act as interpretive provocateurs rather than model suppliers, changes what an LLM-assisted modelling workflow is for, shifting value toward problem framing and communication.
  • Documentation practices like the ODD protocol would need to record where and how LLMs assisted, a direct call the paper makes for implementation-phase code.
  • The duality of symbolic ABM logic and data-driven LLM semantics suggests the two paradigms can complement rather than replace each other.

Reading between the lines

Editorial extensions of the paper, not claims the author makes directly.

  • The empirical concentration may be partly a publication artefact: implementation with LLM agents is the most demonstrable and citable use, while LLM support for problem formulation or validation is harder to showcase in a conference paper.
  • The same single-database design, if applied to other simulation communities, might show a similar implementation-heavy pattern, and the paper's cycle framework could be reused for those cross-community audits.
  • A testable extension would be to benchmark an LLM-assisted modelling workflow against a traditional one across a full cycle, measuring time-to-model, number of errors, and interpretive quality; the paper does not report such measurements but its map implies the need for them.
  • The paper's mitigations imply new reporting infrastructure: a standard way to disclose LLM involvement per phase, which could be piloted in venues requiring structured documentation protocols.
Share X Bluesky LinkedIn Reddit HN

Signed reviews

No signed human review yet.

Editorial analysis

A structured set of objections, weighed in public.

Desk editor's note, referee report, and a circularity audit.

Referee Report

3 major / 6 minor

Summary. The manuscript reports a Rapid Literature Review (RLR) of 22 Scopus-indexed papers that describe implemented uses of large language models in agent-based modelling. It finds that 20 of the 22 papers (91%) use LLMs in the implementation phase, mostly as LLM-powered agents for reasoning, decision-making, or communication, with one paper focused on code generation and one on interpretation of results. The paper then develops a structured map of possible uses across the ABM cycle—problem formulation, system analysis, conceptualization, implementation, verification, validation, interpretation and communication, and documentation—pairing each phase with potential pitfalls and mitigations. The forward-looking map is based on monthly group discussions among the co-authors rather than on a systematic community survey, and the paper ends with a critical reflection on the symbolic versus data-driven tension between ABM and LLMs.

Significance. The paper's main value is its structured taxonomy of opportunities, challenges, and mitigations for LLM use across the entire ABM cycle, which is more comprehensive than prior surveys focused on single stages or single LLM capabilities. The descriptive RLR is transparently reported: the search query is given in the annex, inclusion/exclusion criteria are stated, dual coding with discussion is described, and the 22 coded papers are listed. The 91% implementation concentration, if reliable, is a clear and actionable observation for the community. The forward-looking sections are explicitly framed as critical reflection rather than empirical evidence, and the discussion of LLM-powered agents contains a substantive analysis of the tension between symbolic and data-driven AI. The principal limitations are the narrow empirical base for the concentration claim and the unquantified reliability of the phase coding.

major comments (3)
  1. [Section 4 / Annex / Section 5.1] The central descriptive claim that current LLM use in ABM is concentrated in implementation (20/22, 91%) depends entirely on the corpus retrieved from a single database (Scopus) with a March 2025 cutoff. Section 2 of the same paper cites relevant works that are not in the coded set, including pre-prints and works addressing model design, simulation tasks, and result interpretation (e.g., [14], [17], [40], [2], and [38]). If such works, or other papers not indexed in Scopus, had been included, the phase distribution could shift materially. The paper should either recompute the distribution after expanding the corpus to arXiv and other databases, or explicitly rescope the conclusion to "the Scopus-indexed implemented-use corpus" rather than claiming that "the inherent potential of LLMs has not yet been fully leveraged throughout the ABM cycle" (end of Section 5.1). This rescoping is load-bearing because the forward-looking map is motivated by the perceived gap.
  2. [Section 4 / Table 1] The reliability of the phase coding is asserted but not quantified. The method section says two coders "checked each other's coding and discussed it extensively," but no codebook, disagreement rate, or intercoder agreement statistic is reported. Since the single most important number in the paper is the 20/22 phase classification, the absence of a reproducibility measure for that classification is a substantive gap. Please add a summary of coding disagreements and their resolution, or a quantitative reliability measure such as Cohen's kappa per phase code, and make the full codebook available in the annex if space permits.
  3. [Section 5.1 / Section 5.2 (Implementation)] The reporting of the main count is internally ambiguous. The text says "The most common use (n=20, 91%) involves implementation" and then lists code generation [31] as the main focus of one paper and interpretation [30] as the main focus of another. However, Section 5.2 defines the Implementation phase as including code generation. The paper should clarify whether the 20 implementation papers include or exclude code-generation papers. If code generation is considered part of implementation, the count for that phase would be 21/22 rather than 20/22; if it is excluded, the definition of "implementation" in Section 5.1 differs from the one used in Section 5.2, which confuses the central statistic.
minor comments (6)
  1. [Section 2] The sentence listing limitations of previous studies contains an apparent contradiction: it says previous studies "do not follow a straightforward structure" and then says they "follow a straightforward structure that is not specifically tailored for ABM." One of these clauses is likely a typo and should be corrected.
  2. [Section 5.2 (Verification)] The enumerated list of LLM-assisted verification tasks jumps from item (2) to item (5); the numbering should be corrected to (1), (2), (3), (4).
  3. [Annex] The search query as printed contains unnatural spacing and line breaks (e.g., "prompt e n g i n e e r i n g"), which makes it difficult for readers to reproduce. A clean, copy-pasteable version of the query should be provided.
  4. [Section 4] The paper reports that the RLR used a single database and a March 2025 cutoff but does not explicitly discuss the implications of these choices for the fast-moving LLM literature. A short limitation note, even one or two sentences, would help calibrate readers' expectations.
  5. [Acknowledgments] The phrase "and and" appears in the acknowledgment for Vivek Nallur's grants; this should be corrected.
  6. [Section 5.2] The forward-looking map is based on structured group discussions among the co-authors, but the paper does not describe the discussion protocol (e.g., number of sessions, how suggestions were aggregated, how disagreements were resolved). Adding a brief description would improve transparency for a contribution that is partly a collective expert opinion.

Circularity Check

0 steps flagged · score 0.0 of 10

No significant circularity: the 91% implementation figure is a coded observation, and the self-cited ABM cycle is an organizational framework rather than a load-bearing derivation.

full rationale

This paper is a rapid literature review and expert synthesis, not a derivation with fitted parameters or constructed predictions. The central descriptive claim, 'The most common use (n=20, 91%) involves implementation' (Section 5.1), is a count of the 22 papers coded during the RLR (Section 4 and annex), so it is an observed distribution rather than a quantity forced by the framework. The coding scheme did use 'where in the ABM cycle (cf. [41]) were LLMs used' (Section 4), and the cycle itself comes from [41], a co-author's prior work; this is a minor self-citation, but it is not load-bearing because the phase distribution is not entailed or defined by that citation, and the paper does not invoke any uniqueness theorem or forbid alternative phase schemes. The inclusion criterion requiring an 'implemented use of LLMs in connection with ABM' does not by construction force the implementation phase, since an implemented use could occur in any phase. The forward-looking Section 5.2 is explicitly a co-author group-discussion synthesis with stated pitfalls and mitigations, not a prediction derived from the review inputs. The Scopus-only, March-2025-cutoff corpus is a coverage limitation that could affect the phase distribution, but that is a methodological threat to external validity, not circularity. No equation or fitted parameter is reused as an output, so no circular reduction can be exhibited.

Assumptions & free parameters 0 free parameters · 4 assumptions · 0 invented entities

The paper is a review and synthesis; it introduces no free parameters, fitted values, or new postulated entities. The central claims rest on the chosen ABM cycle framework, the RLR method, the LLM capability taxonomy, and the group discussion process, all of which are domain assumptions or ad hoc methodological choices.

assumptions (4)
  • domain assumption The ABM cycle as defined by Siebers and Klügl [41] and Nikolic and Ghorbani [34] is the correct organizing structure for mapping LLM use.
    Section 3.1 states the phase descriptions are based on these studies; the entire Section 5.2 is organized by these phases.
  • domain assumption The Rapid Literature Review methodology from [7] yields a sufficiently valid and comprehensive evidence base for the descriptive claims.
    Section 4 adopts RLR from [7]; the single-database Scopus search is a limitation acknowledged by the authors but still foundational to the 'current use' claims.
  • domain assumption The taxonomy of LLM capabilities from [48] adequately represents the range of LLM abilities relevant to ABM.
    Section 3.2 lists usage areas from [48]; the 'How' proposals in Section 5.2 are framed around these categories.
  • ad hoc to paper Co-author group discussions (monthly meetings from June 2024) provide a sound basis for identifying possible future uses.
    Section 4 describes the group discussion method; the representativeness of this volunteer group is not empirically established.

how reviews work

0 comments
Cite this review

Pith. "Pith review of Large Language Models for Agent-Based Modelling: Current and possible uses across the modelling cycle." pith.science (2026). https://pith.science/paper/EDZWP5VN

@misc{pith2026250705723,
  author       = {Pith},
  title        = {Pith review of: Large Language Models for Agent-Based Modelling: Current and possible uses across the modelling cycle},
  year         = {2026},
  howpublished = {\url{https://pith.science/paper/EDZWP5VN}},
  note         = {Machine review of arXiv:2507.05723}
}
read the original abstract

The emergence of Large Language Models (LLMs) with increasingly sophisticated natural language understanding and generative capabilities has sparked interest in the Agent-based Modelling (ABM) community. With their ability to summarize, generate, analyze, categorize, transcribe and translate text, answer questions, propose explanations, sustain dialogue, extract information from unstructured text, and perform logical reasoning and problem-solving tasks, LLMs have a good potential to contribute to the modelling process. After reviewing the current use of LLMs in ABM, this study reflects on the opportunities and challenges of the potential use of LLMs in ABM. It does so by following the modelling cycle, from problem formulation to documentation and communication of model results, and holding a critical stance.

Figures

Figures reproduced from arXiv: 2507.05723 by the authors.

Figure 1
Figure 1. Social simulation study life cycle. From [41] nor for ABM); they follow a straightforward structure that is not specifically tai￾lored for ABM, but for simulation modeling more broadly; or the structure they use does not cover the entire ABM cycle. Our study aims to address these limita￾tions by using the structure of the ABM cycle and exploring all the phases of this cycle and by accounting for the variety of LLMs … view at source ↗

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Subjective-Graph LLM Agents for Simulating Uncertainty in Classroom Social Perception

    cs.AI 2026-03 conditional novelty 6.0 of 10

    Subjective-graph LLM agents on 12 real classrooms accumulate collective ranking error from 0.066 to 0.124 over six exams despite repeated score anchors.

  2. A Large Language Model-Driven Agent-Based Modeling Framework with Multi-Round Communication for Simulating Vaccine Opinion Dynamics

    cs.MA 2026-07 conditional novelty 4.0 of 10

    An LLM-driven agent-based model with multi-round dialogue reproduces non-linear social influence patterns in vaccination opinion dynamics, with memory increasing resistance and prompt diversity increasing adoption.

Reference graph

Works this paper leans on

52 extracted references · 49 canonical work pages · cited by 2 Pith papers

  1. [40]

    arXiv preprint arXiv:2405.08032 (2024)

    Siebers, P.O.: Exploring the potential of conversational ai support for agent-based social simulation model design. arXiv preprint arXiv:2405.08032 (2024)

  2. [14]

    Frydenlund, E., Martínez, J., Padilla, J.J., Palacio, K., Shuttleworth, D.: Modeler in a box: how can large language models aid in the simulation modeling process? Simulation 100(7), 727–749 (2024)

  3. [17]

    In: 2023 Winter simu- lation conference (WSC)

    Giabbanelli, P.J.: Gpt-based models meet simulation: how to efficiently use large- scale pre-trained language models across simulation tasks. In: 2023 Winter simu- lation conference (WSC). pp. 2920–2931. IEEE (2023)

  4. [2]

    System Dynamics Review 40(3), e1773 (2024)

    Akhavan, A., Jalali, M.S.: Generative ai and simulation modeling: how should you (not) use large language models like chatgpt. System Dynamics Review 40(3), e1773 (2024)

  5. [38]

    In: Conference of the European Social Simulation Association

    Polhill, G.,Borit,M.,Elsenbroich, C.,Verhagen, H.,Wijermans, N.:Ethical dimen- sions to empirical applications of agent-based social simulation. In: Conference of the European Social Simulation Association. p. FORTHCOMING. Springer (2025)

  6. [31]

    In: 2024 Winter Simulation Con- ference (WSC)

    Martínez, J., et al.: Enhancing gpt-3.5’s proficiency in netlogo through few-shot prompting and retrieval-augmented generation. In: 2024 Winter Simulation Con- ference (WSC). IEEE (2024)

  7. [30]

    Future Internet15(12), 375 (2023)

    Lynch, C.J., et al.: A structured narrative prompt for prompting narratives from large language models: Sentiment assessment of chatgpt-generated narratives and real tweets. Future Internet15(12), 375 (2023)

  8. [1]

    In- ternational Journal of Social Research Methodology25(4), 517–540 (2022)

    Achter, S., Borit, M., Chattoe-Brown, E., Siebers, P.O.: Rat-rs: a reporting stan- dard for improving the documentation of data use in agent-based modelling. In- ternational Journal of Social Research Methodology25(4), 517–540 (2022)

Show all 52 references
  1. [3]

    Journal of Artificial Societies and Social Simulation25(4) (2022)

    Anzola, D., Barbrook-Johnson, P., Gilbert, N.: The ethics of agent-based social simulation. Journal of Artificial Societies and Social Simulation25(4) (2022)

  2. [4]

    In: 2023 15th International Congress on Advanced Applied Informatics Winter (IIAI-AAI-Winter)

    Aoki, N., Mori, N., Okada, M.: Analysis of llm-based narrative generation using the agent-based simulation. In: 2023 15th International Congress on Advanced Applied Informatics Winter (IIAI-AAI-Winter). IEEE (2023)

  3. [5]

    In: Simulating social phenomena, pp

    Axelrod, R.: Advancing the art of simulation in the social sciences. In: Simulating social phenomena, pp. 21–40. Springer (1997)

  4. [6]

    Proceedings of the national academy of sciences99(suppl_3), 7280– 7287 (2002)

    Bonabeau, E.: Agent-based modeling: Methods and techniques for simulating hu- man systems. Proceedings of the national academy of sciences99(suppl_3), 7280– 7287 (2002)

  5. [7]

    https://doi.org/10.1007/978-1-0716-1566-9

    Bouck, Z., Straus, S., Tricco, A.: Meta-Research Methods and Protocols: Methods and Protocols (2021). https://doi.org/10.1007/978-1-0716-1566-9

  6. [8]

    In: Proceedings of Agents

    Carley, K.M.: Simulating society: The tension between transparency and veridical- ity. In: Proceedings of Agents. pp. 2–2 (2002)

  7. [9]

    In: Proceedings of the 2024 CHI Conference on Human Factors in Computing Systems (2024)

    Chen, J., et al.: Learning agent-based modeling with llm companions: Experiences of novices and experts using chatgpt & netlogo chat. In: Proceedings of the 2024 CHI Conference on Human Factors in Computing Systems (2024)

  8. [10]

    In: Proceedings of the Annual Meeting of the Cognitive Science Society

    Chuang, Y.S., et al.: Simulating opinion dynamics with networks of llm-based agents. In: Proceedings of the Annual Meeting of the Cognitive Science Society. vol. 46 (2024)

  9. [11]

    Minding Norms: Mechanisms and Dynamics of Social Order in Agent Societies p

    Edmonds, B.: Agent-based social simulation and its necessity for understanding so- cially embedded phenomena. Minding Norms: Mechanisms and Dynamics of Social Order in Agent Societies p. 34 (2013) Large Language Models for Agent-Based Modelling 15

  10. [12]

    In: International Conference on Advances in Social Networks Analysis and Mining

    Ferraro, A., et al.: Agent-based modelling meets generative ai in social network simulations. In: International Conference on Advances in Social Networks Analysis and Mining. Springer Nature Switzerland, Cham (2024)

  11. [13]

    Jasss-The journal of artificial societies and social simulation20(4), 2 (2017)

    Flache, A., Mäs, M., Feliciani, T., Chattoe-Brown, E., Deffuant, G., Huet, S., Lorenz, J.: Models of social influence: Towards the next frontiers. Jasss-The journal of artificial societies and social simulation20(4), 2 (2017)

  12. [15]

    Humanities and Social Sciences Communications11(1), 1–24 (2024)

    Gao, C., Lan, X., Li, N., Yuan, Y., Ding, J., Zhou, Z., Xu, F., Li, Y.: Large language models empowered agent-based modeling and simulation: A survey and perspectives. Humanities and Social Sciences Communications11(1), 1–24 (2024)

  13. [16]

    System Dynamics Review40(1), e1761 (2024)

    Ghaffarzadegan, N., et al.: Generative agent-based modeling: An introduction and tutorial. System Dynamics Review40(1), e1761 (2024)

  14. [18]

    In: Artificial societies, pp

    Gilbert, N.: Emergence in social simulation. In: Artificial societies, pp. 134–143. Routledge (2006)

  15. [19]

    Ecological modelling221(23), 2760– 2768 (2010)

    Grimm, V., Berger, U., DeAngelis, D.L., Polhill, J.G., Giske, J., Railsback, S.F.: The odd protocol: a review and first update. Ecological modelling221(23), 2760– 2768 (2010)

  16. [20]

    HHAI 2024: Hybrid Human AI Systems for the Social Good pp

    Gürcan, Ö.: Llm-augmented agent-based modelling for social simulations: Chal- lenges and opportunities. HHAI 2024: Hybrid Human AI Systems for the Social Good pp. 134–144 (2024)

  17. [21]

    In: Proceedings of the Thirty-Third International Joint Conference on Artificial Intelligence, IJCAI-24

    Hu, Y., et al.: An llm-enhanced agent-based simulation tool for information prop- agation. In: Proceedings of the Thirty-Third International Joint Conference on Artificial Intelligence, IJCAI-24. Jeju, Republic of Korea (2024)

  18. [22]

    In: International Congress on Information and Com- munication Technology

    Ilagan, J.B., et al.: Exploratory customer discovery through simulation using chat- gpt and prompt engineering. In: International Congress on Information and Com- munication Technology. Springer Nature Singapore, Singapore (2024)

  19. [23]

    arXiv preprint arXiv:2406.18841 (2024)

    Jiao, J., Afroogh, S., Xu, Y., Phillips, C.: Navigating llm ethics: Advancements, challenges, and future directions. arXiv preprint arXiv:2406.18841 (2024)

  20. [24]

    In: Proceedings of the 19th International Conference on the Foundations of Digital Games (2024)

    Kelly, J., Mateas, M., Wardrip-Fruin, N.: Paradise: An experiment extending the ensemble social physics engine with language models. In: Proceedings of the 19th International Conference on the Foundations of Digital Games (2024)

  21. [25]

    Khatami, S.: Ai-enhanced abm development: Facilitating agent-based modeling using artificial intelligence (2025)

  22. [26]

    arXiv preprint arXiv:2504.03274 (2025)

    Larooij, M., Törnberg, P.: Do large language models solve the problems of agent- based modeling? a critical review of generative social simulations. arXiv preprint arXiv:2504.03274 (2025)

  23. [27]

    In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) (2024)

    Li, N., et al.: Econagent: Large language model-empowered agents for simulat- ing macroeconomic activities. In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) (2024)

  24. [28]

    In: 34th Congress of the Inter- national Council of the Aeronautical Sciences (2024)

    Lovaco, J., et al.: Large language model-driven simulations for system of systems analysis in firefighting aircraft conceptual design. In: 34th Congress of the Inter- national Council of the Aeronautical Sciences (2024)

  25. [29]

    Physics of Life Reviews (2024) 16 L

    Lu, Y., Aleta, A., Du, C., Shi, L., Moreno, Y.: Llms and generative agent-based models for complex systems research. Physics of Life Reviews (2024) 16 L. Vanhée et al

  26. [32]

    Minaee, S., Mikolov, T., Nikzad, N., Chenaghlu, M., Socher, R., Amatriain, X., Gao, J.: Large language models: A survey (2025), https://arxiv.org/abs/2402.06196

  27. [33]

    In: Findings of the Association for Computational Linguistics: ACL 2024 (2024)

    Mou,X.,Wei,Z.,Huang,X.J.:Unveilingthetruthandfacilitatingchange:Towards agent-based large-scale social movement simulation. In: Findings of the Association for Computational Linguistics: ACL 2024 (2024)

  28. [34]

    In: 2011 international conference on networking, sensing and control

    Nikolic, I., Ghorbani, A.: A method for developing agent-based models of socio- technical systems. In: 2011 international conference on networking, sensing and control. pp. 44–49. IEEE (2011)

  29. [35]

    In: Proceedings of the 23rd In- ternational Conference on Autonomous Agents and Multiagent Systems (AAMAS) (2024)

    Niu, T., Zhang, W., Zhao, R.: Solution-oriented agent-based models generation with verifier-assisted iterative in-context learning. In: Proceedings of the 23rd In- ternational Conference on Autonomous Agents and Multiagent Systems (AAMAS) (2024)

  30. [36]

    Frontiers in artificial intelligence 6, 1199350 (2023)

    Orrù, G., Piarulli, A., Conversano, C., Gemignani, A.: Human-like problem-solving abilities in large language models using chatgpt. Frontiers in artificial intelligence 6, 1199350 (2023)

  31. [37]

    Applied Sciences14(5), 2074 (2024)

    Patil, R., Gudivada, V.: A review of current trends, techniques, and challenges in large language models (llms). Applied Sciences14(5), 2074 (2024)

  32. [39]

    In: Proceedings of the 29th conference on Winter simulation (1997)

    Robinson, S.: Simulation model verification and validation: increasing the users’ confidence. In: Proceedings of the 29th conference on Winter simulation (1997)

  33. [41]

    Simulating social complexity: a handbook pp

    Siebers, P.O., Klügl, F.: What software engineering has to offer to agent-based social simulation. Simulating social complexity: a handbook pp. 81–117 (2017)

  34. [42]

    Policy Insights from the Behavioral and Brain Sciences5(2), 240–246 (2018)

    Sun, R.: Cognitive social simulation for policy making. Policy Insights from the Behavioral and Brain Sciences5(2), 240–246 (2018)

  35. [43]

    Takata,R.,Masumori,A.,Ikegami,T.:Spontaneousemergenceofagentindividual- itythroughsocialinteractionsinlargelanguagemodel-basedcommunities.Entropy 26(12), 1092 (2024)

  36. [44]

    PET clinics16(4), 449–469 (2021)

    Toosi, A., Bottino, A.G., Saboury, B., Siegel, E., Rahmim, A.: A brief history of ai: how to prevent another winter (a critical review). PET clinics16(4), 449–469 (2021)

  37. [45]

    In: Proceedings of the 19th International Conference on the Foundations of Digital Games (2024)

    Treanor, M., Samuel, B., Nelson, M.J.: Prototyping slice of life: Social physics with symbolically grounded llm-based generative dialogue. In: Proceedings of the 19th International Conference on the Foundations of Digital Games (2024)

  38. [46]

    In: 2024 IEEE International Conference on Agents (ICA)

    Vidler, A., Walsh, T.: Tradertalk: An llm behavioural abm applied to simulating human bilateral trading interactions. In: 2024 IEEE International Conference on Agents (ICA). IEEE (2024)

  39. [47]

    In: Agentic Markets Workshop at ICML 2024 (2024) Large Language Models for Agent-Based Modelling 17

    Wu, Z., et al.: Shall we team up: Exploring spontaneous cooperation of competing llm agents. In: Agentic Markets Workshop at ICML 2024 (2024) Large Language Models for Agent-Based Modelling 17

  40. [48]

    ACM Transactions on Knowledge Discovery from Data18(6), 1–32 (2024)

    Yang, J., Jin, H., Tang, R., Han, X., Feng, Q., Jiang, H., Zhong, S., Yin, B., Hu, X.: Harnessing the power of llms in practice: A survey on chatgpt and beyond. ACM Transactions on Knowledge Discovery from Data18(6), 1–32 (2024)

  41. [49]

    In: Proceedings of the 2024 7th International Conference on Math- ematics and Statistics (2024)

    Zaslavsky, I., et al.: Enhancing spatially-disaggregated simulations with large lan- guage models. In: Proceedings of the 2024 7th International Conference on Math- ematics and Statistics (2024)

  42. [50]

    Systems13(1), 29 (2025)

    Zhang, L., et al.: Llm-aidsim: Llm-enhanced agent-based influence diffusion simu- lation in social networks. Systems13(1), 29 (2025)

  43. [51]

    In: International Symposium on Knowl- edge and Systems Sciences

    Zheng, W., Tang, X.: Simulating social network with llm agents: An analysis of in- formation propagation and echo chambers. In: International Symposium on Knowl- edge and Systems Sciences. Springer Nature Singapore, Singapore (2024)

  44. [52]

    chat gpt

    Zhou, X., et al.: Is this the real life? is this just fantasy? the misleading success of simulating social interactions with llms. In: Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing (2024) Annexes Rapid Literature Review query ( TITLE - ...

Pith tools

Reviewed August 6, 2026 · model on record in the stance chip above.