Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:53:29.689406Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 5 of 5 outbound references and 5 inbound Pith citation observations for arXiv:2501.18062.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:53:29.689406Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:01:16.726294Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T08:56:25.216379Z
5 of 5 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e72439a9-86fe-476f-aab3-5a609cab0709 · outbound
FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models FinQA: A Dataset of Numerical Reasoning over Financial Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 474df27b-e50d-4e34-bc2a-f91db76556c0 · outbound
FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models FinanceBench: A New Benchmark for Financial Question Answering
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a1c8f9d-32ed-492d-a6d9-ba31d633346c · outbound
FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a430616-9e27-4c79-a038-68fdfff5f118 · outbound
FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models LegalBench: A Collaboratively Built Benchmark for Measuring Legal Reasoning in Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a0ac1f2-210a-47c6-9806-17bf56a4be43 · outbound
FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models On Leakage of Code Generation Evaluation Datasets
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5cbb368-61f5-41b7-a401-2147d79c0faa · inbound
On Path to Multimodal Historical Reasoning: HistBench and HistAgent FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fbc8455-1593-4cd7-8fff-6dce4cc607e4 · inbound
Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 12905610-f501-471c-8f08-6dbd8eade104 · inbound
BizCompass: Benchmarking the Reasoning Capabilities of LLMs in Business Knowledge and Applications FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 13123e87-942a-4f52-a066-c3c8fae96dff · inbound
LATTICE: Evaluating Decision Support Utility of Crypto Agents FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 63f6b8a0-ee55-448e-b67a-77b42fa61a60 · inbound
Fin-Bias: Comprehensive Evaluation for LLM Decision-Making under human bias in Finance Domain FinanceQA: A Benchmark for Evaluating Financial Analysis Capabilities of Large Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.