Pith. sign in

Paper Citation Record · LEDGER

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models

As of 11 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2501.12975.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12975 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T16:40:09.392349Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved16
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 15fa9b3a-7eaa-4d45-b452-e9e442a5443b · outbound

This paper cites Chain-of-Verification Reduces Hallucination in Large Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Chain-of-Verification Reduces Hallucination in Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.340378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.340378Z digest=sha256:9e09dd6f3bd9b45d05a3f8f181793cdbb943b2d042b8268108fef3bd8d2b8457

Observation 5da9f60d-ff73-4e2e-b34a-3adec2ec0f74 · outbound

This paper cites A Survey on In-context Learning.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models A Survey on In-context Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.344554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.344554Z digest=sha256:b05f751797d3dfbf1377a9949e88c516439080188e38414783fffd67189ae26b

Observation e9289ce8-72fb-4181-8d4a-3ff64b1d6337 · outbound

This paper cites Gao, Y ., Xiong, Y ., Gao, X., Jia, K., Pan, J., Bi, Y ., Dai, Y ., Sun, J., Wang, M., and Wang, H.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Gao, Y ., Xiong, Y ., Gao, X., Jia, K., Pan, J., Bi, Y ., Dai, Y ., Sun, J., Wang, M., and Wang, H

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.348776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.348776Z digest=sha256:9277e2817e523cee68a795088e8b685270ee759af23c57b90f81e27bc31d102e

Observation 659bcc24-1057-48b1-a378-43a21bcc63ad · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.352726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.352726Z digest=sha256:d6c04be23b61eed635e8ecf03c1b18a47bb13f3e8d8d0dd15529365b1730a9ec

Observation f7b5f089-0863-4120-8b03-72760b7a8ce2 · outbound

This paper cites doi: 10.1145/3703155.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models doi: 10.1145/3703155

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.356950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.356950Z digest=sha256:87cd6c0b0174cbc5597a317e930031b7da53e6dbbf38db001607901e6c5e77a5

Observation c6592de1-66ad-4bdc-9c27-a63986b4ced5 · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.364771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.364771Z digest=sha256:3861d2e57aba52a62241f21ab4e30efbd4b13f9e7e7dc336704b25d246ae8fd2

Observation 89572614-f4c1-4db0-86e6-4783b6ba3a7a · outbound

This paper cites Small Language Models: Survey, Measurements, and Insights.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Small Language Models: Survey, Measurements, and Insights

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.368585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.368585Z digest=sha256:60815a5fc95e589a8eb1a82b80ebfbabaf75724ac1a83af4125d2cd417de6fe7

Observation 35e53049-a6af-4509-a787-4ef9e04d3c45 · outbound

This paper cites FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.372674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.372674Z digest=sha256:0738d6e9a7bae4e9377c82cf0fa4c0516470b18d04f2897c31d27712d0613c0a

Observation 4e023ab7-ff79-4da2-89da-1495394e556b · outbound

This paper cites Generating Benchmarks for Factuality Evaluation of Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Generating Benchmarks for Factuality Evaluation of Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.376607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.376607Z digest=sha256:a09e3c6db68254f6f5f6e94ed4a8678c8ece667d73b9fcbddec14ddc12cdf32d

Observation 7e709871-fc63-4101-b293-ecf2ec627b46 · outbound

This paper cites A Survey of Small Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models A Survey of Small Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.380671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.380671Z digest=sha256:6f1e7757b7e7c49021be8137def69fa500266a66b1e003929ded169cc563a2ab

Observation a431e742-8563-47d5-ab8d-2daa168b6991 · outbound

This paper cites FACTOID: FACtual enTailment fOr hallucInation Detection.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models FACTOID: FACtual enTailment fOr hallucInation Detection

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.384433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.384433Z digest=sha256:2d059d9e893c945566c5765eaeadffcc68d38de64b5ff2f548e86513b022419a

Observation 8c711cbc-4cd2-49f6-8a08-597120a21ff5 · outbound

This paper cites A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models

Reference 16

Resolution
malformed identifier
no resolver link, observed 2026-08-10T16:40:09.388276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.388276Z digest=sha256:6986b4414b96d1ba2bd42bc071df13f05600980dc1a731f8ac420c2af85ffcd5

Observation de7a8bef-c064-429d-95ab-70410a3a1bfc · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.392349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.392349Z digest=sha256:c7e4729a14f2d725668e643feb5371731fbf7c2d3f4bf6e265bd65592e4c5630

Observation 92a2703a-e9dc-4386-87c4-8f8c1c178919 · outbound

This paper cites HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.360884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.360884Z digest=sha256:80a3bc337701be447ba176f6c19e54236eff750ee2e825dd5fcb32f514c5d926

Observation e74b7ef8-6a1a-4ddd-b55c-9a1d8f73b4d3 · outbound

This paper cites Language Models are Few-Shot Learners.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Language Models are Few-Shot Learners

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.326384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.326384Z digest=sha256:1d719b1d9cf004ce5e8cc89965231bdbfabb395624368bd35b42b1683b6fdb05

Observation 20f66010-b418-42cd-8b81-fdaa9f467c28 · outbound

This paper cites Evaluating Hallucinations in Chinese Large Language Models.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models Evaluating Hallucinations in Chinese Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.335541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.335541Z digest=sha256:cee20b03ae2fe472593b4451c9da5d2076609969a92f35f490df4b0d9b427eb2

Observation a44d3d6f-8e05-4c79-89b6-e3062ef48755 · outbound

This paper cites FactCHD: Benchmarking Fact-Conflicting Hallucination Detection.

OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models FactCHD: Benchmarking Fact-Conflicting Hallucination Detection

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T16:40:09.331361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:40:09.331361Z digest=sha256:fb6aba4e95bd84e0deed37c9f2e0b9eab9079a85f2469372249e8ab08eee1178

Pith citing papers

No inbound Pith citation observations are available.