Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:26:02.765845Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 0 inbound Pith citation observations for arXiv:2412.11988.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:26:02.765845Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
11 of 11 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a060630d-61b3-487b-97a4-ce393b1bba62 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7975dc92-c889-4679-b8b5-7caf0e115050 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c790df0f-0c01-491a-8b9f-dae81159b7bf · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation AI now beats humans at basic tasks — new benchmarks are needed, says major report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5a36e48c-08d1-43dc-88e7-8e1e1e71d5e8 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a97c27ac-a0a6-4193-bf31-d979cd9bc06f · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation StylePTB: A Compositional Benchmark for Fine-grained Controllable Text Style Transfer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0211efbe-614f-420d-aa31-5eda3a7a4614 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation Fine-grained Text Style Transfer with Diffusion-Based Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3d5ea51-3a4e-49b9-951d-11f840bf0365 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation C., Shoham, Y., Wald, R., and Clark, J
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7b3f1a3b-871c-4b00-951f-6764fc93bd23 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation Learning to reason with LLMs
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 54d87bdb-d81b-4287-bd93-3599c098793b · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9f1481b-f470-4768-be29-cc5edd664e29 · outbound
SciFaultyQA: Benchmarking LLMs on Faulty Science Question Detection with a GAN-Inspired Approach to Synthetic Dataset Generation Crowdsourcing Multiple Choice Science Questions
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09f2343c-70f9-4924-a5ac-bca3add8e83a · outbound
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.