Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:21:09.834127Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2504.18572.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:21:09.834127Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 98efdbdd-028e-4080-90a2-8aa8efd840de · outbound
BELL: Benchmarking the Explainability of Large Language Models Natural language processing: State of the art, current trends and challenges
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e0e9772e-993e-4570-8202-8fb6fb26c4fe · outbound
BELL: Benchmarking the Explainability of Large Language Models Multilingual machine translation with large language models: Empirical results and analysis, 2023
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90a70ab8-e798-4eca-8c9c-189eaee12bc1 · outbound
BELL: Benchmarking the Explainability of Large Language Models Wordcraft: story writing with large language models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a5a355d9-353a-46cb-8f84-ef2e8518e0b0 · outbound
BELL: Benchmarking the Explainability of Large Language Models https://medium.com/whatnot - engineering/enhancing-search-using-large-language-models-f9dcb988bdb9
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 98c32250-c170-4b33-96a1-35ef5b150813 · outbound
BELL: Benchmarking the Explainability of Large Language Models Code Llama: Open Foundation Models for Code
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36ee1f7b-e6a1-41ab-bba3-43f89a885d4d · outbound
BELL: Benchmarking the Explainability of Large Language Models Bloomberggpt: A large language model for finance, 2023
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0703a95a-74aa-4f88-b36c-c5598e2258bf · outbound
BELL: Benchmarking the Explainability of Large Language Models The impact of large language models on scientific discovery: a preliminary study using gpt-4, 2023
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 226de855-8c00-4e3b-bf8b-1c17a7b76954 · outbound
BELL: Benchmarking the Explainability of Large Language Models Pllama: An open -source large language model for plant science, 2024
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fc13c7a3-9bbd-49ca-972c-4255cda2c17d · outbound
BELL: Benchmarking the Explainability of Large Language Models Artgpt-4: Artistic vision-language understanding with adapter-enhanced minigpt-4, 2023
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 157f1900-e870-4cf3-bce3-25772c0224bf · outbound
BELL: Benchmarking the Explainability of Large Language Models Taoli llama
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 769cd097-00a7-4cb4-9f01-9acaa6065ac0 · outbound
BELL: Benchmarking the Explainability of Large Language Models Marinegpt: Unlocking secrets of “ocean” to the public, 2023
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 51c88f23-8d03-44d4-88b4-fa72ec605517 · outbound
BELL: Benchmarking the Explainability of Large Language Models Disc-lawllm: Fine-tuning large language models for intelligent legal services, 2023
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ebb52783-b724-402d-82d1-b57fc0eb3ad0 · outbound
BELL: Benchmarking the Explainability of Large Language Models Large language models and political science
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a57423f-20f1-46af-a1e8-9f67963cf49a · outbound
BELL: Benchmarking the Explainability of Large Language Models Alpacare:instruction-tuned large language models for medical application, 2023
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc4c6b75-4d0d-4462-ac74-944233082852 · outbound
BELL: Benchmarking the Explainability of Large Language Models Davison, Quanzheng Li, Yong Chen, Hongfang Liu, and Lichao Sun
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4a889158-793c-47c8-b413-7485049af38a · outbound
BELL: Benchmarking the Explainability of Large Language Models Factuality challenges in the era of large language models, 2023
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 91025a1d-f9c2-4db6-9fc9-e5e0f3305169 · outbound
BELL: Benchmarking the Explainability of Large Language Models Unraveling the link between translations and gender bias in llms, 2023
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 409c95cb-a209-4132-87c8-0a51d80203b8 · outbound
BELL: Benchmarking the Explainability of Large Language Models Jailbroken: How Does LLM Safety Training Fail?
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e367df9f-18f4-490b-9bdd-909e86452bb0 · outbound
BELL: Benchmarking the Explainability of Large Language Models Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1c91cfb2-46f9-4b80-aaa6-d13192a878c7 · outbound
BELL: Benchmarking the Explainability of Large Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f865b43-d8de-4d45-b35a-3738878826ed · outbound
BELL: Benchmarking the Explainability of Large Language Models Maieutic Prompting: Logically Consistent Reasoning with Recursive Explanations
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1580d568-7c04-4c3c-a295-00c439476217 · outbound
BELL: Benchmarking the Explainability of Large Language Models Large Language Models are Zero-Shot Reasoners
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 59ae4b4f-5cb9-49b5-9711-7ebd9097a4f4 · outbound
BELL: Benchmarking the Explainability of Large Language Models Soft-prompt Tuning for Large Language Models to Evaluate Bias
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f70a547d-3c60-4929-9829-d92d18e01aad · outbound
BELL: Benchmarking the Explainability of Large Language Models Chi, Quoc V
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d8f37bf0-8bda-488b-a542-de327ac31989 · outbound
BELL: Benchmarking the Explainability of Large Language Models Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f329d8a-3870-4189-8e0a-5fd9d567fb28 · outbound
BELL: Benchmarking the Explainability of Large Language Models Graph of Thoughts: Solving Elaborate Problems with Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05dd6804-b44b-4a78-a7bf-291dd05c07a5 · outbound
BELL: Benchmarking the Explainability of Large Language Models Beyond Chain-of-Thought, Effective Graph-of-Thought Reasoning in Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00ae8f80-cef1-4fe5-9ab7-00638f63aca0 · outbound
BELL: Benchmarking the Explainability of Large Language Models Le, and Ed H
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 027cb75c-6c49-4b96-8792-f2d4886c5164 · outbound
BELL: Benchmarking the Explainability of Large Language Models Self -consistency improves chain of thought reasoning in Infosys Responsible AI Office language models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 29e7abaa-4027-4808-bdcc-512f4c2c9064 · outbound
BELL: Benchmarking the Explainability of Large Language Models Star: Bootstrapping reasoning with reasoning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae236c1a-ef04-456c-8d9c-a5246a769a16 · outbound
BELL: Benchmarking the Explainability of Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 634891e2-62f2-4fc5-b9c3-0a1ab0457a66 · outbound
BELL: Benchmarking the Explainability of Large Language Models Thread of Thought Unraveling Chaotic Contexts
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a816261-d817-489b-b40f-4152c574f2fd · outbound
BELL: Benchmarking the Explainability of Large Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8dd44d4-d7fa-4777-a87a-b59641704a50 · outbound
BELL: Benchmarking the Explainability of Large Language Models Logic-of-Thought: Injecting Logic into Contexts for Full Reasoning in Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a54889c-6f82-4d57-862f-9d0e39eb0df4 · outbound
BELL: Benchmarking the Explainability of Large Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ad1ca8c4-2239-4269-84eb-d24dde5154a3 · outbound
BELL: Benchmarking the Explainability of Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed2573d-b198-4991-a8f1-c06cf1518245 · outbound
BELL: Benchmarking the Explainability of Large Language Models Orca: Progressive Learning from Complex Explanation Traces of GPT-4
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
No inbound Pith citation observations are available.