Pith. sign in

Paper Citation Record · LEDGER

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge

As of 6 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2604.05593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.05593 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:38:11.595077Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T05:36:21.058156Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-29T05:43:08.668932Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact32
  • verified fuzzy6
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5a8acef9-8e2b-4f2b-a601-843d8d83fa04 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.652023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:06385eb279ec75f41c24e56dc63276c40aaef6504ba8fd5fd0c62d8d4e794c67

Observation fbc64fda-7fd1-4891-bc00-bc92861188fe · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.663685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:9df40ce40152519b5b51ade19eb2a8cc960f69220bceb70d4f3fcfc901353ed9

Observation 0f2971c2-a22c-4237-afd8-bf5a00c56cb2 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.655760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:6bcd2f5274335417429c6e1eb94b6d0ee8d345a4a80d42805b790a705467e6d3

Observation 08946746-33d5-4b39-9ed2-eed6361133b7 · outbound

This paper cites Cacioppo, Louis G.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Cacioppo, Louis G

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T04:30:37.637626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:33def59719c7901582eeca13f9693713cd5aa7e8d8acdb4aba72d3186880509c

Observation bf718d33-67ed-448e-98b1-47bf3966f37b · outbound

This paper cites Humans or LLM s as the Judge? A Study on Judgement Bias.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Humans or LLM s as the Judge? A Study on Judgement Bias

Reference 5

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.627560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:d57506a9136004499de3a1fa40038039eccf4a2bcdb81e6287596b96b9d7544d

Observation 305a6179-b5ad-4a44-8370-51fe04218837 · outbound

This paper cites Proceedings of the National Academy of Scienceshttps://doi.org/10.1073/pnas.2412015122 (2025) doi:10.1073/pnas.2412015122.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Proceedings of the National Academy of Scienceshttps://doi.org/10.1073/pnas.2412015122 (2025) doi:10.1073/pnas.2412015122

Reference 6

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.624821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:2a106247ffc5ec6200a947238d85bc7a746a0df7e8a712fd20c8ae3f2f4d4325

Observation 552f7a16-69f6-4ba6-a009-d0645ffc7510 · outbound

This paper cites In: Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge In: Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T19:40:45.639719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:1fe2d8566db1ac75d689ee682b1ab0ac68b3643e34e28dfab6156de1208ab911

Observation f13d996f-bf89-42ad-b3ea-39c1ba4ac376 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.648702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:779aa9cc1ff17fc51c1d29eb81632076c4447dcf4ab33419fa506473b49b6e12

Observation aa8216d6-e3c8-4ff3-8e62-994a6fcf227a · outbound

This paper cites Cognitive Bias in Decision-Making with LLM s.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Cognitive Bias in Decision-Making with LLM s

Reference 9

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.634912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:fb2a685fdce7da562205ebca0433f50bbc09cc775f4da6f5672509965ceac63e

Observation 9771e6ba-845e-4c12-b2d6-a503fae7d33b · outbound

This paper cites We Need Structured Output.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge We Need Structured Output

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:40:45.632186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:50c8ef03f0acc2d374385de3fe440460b83d68b82642fa53364e687b869641e6

Observation cf8cad75-8f11-4885-a9c5-461ee9a0fefb · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.642170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:adc181a219cb2953acd5b63c2f42c9b2b729c26cb290856e66684c57bed43824

Observation 6f2ae0d9-334f-441a-b87f-c5eed7a82274 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.659841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:48677b2aac77d8879ba464afaf6163bdc8993cc2f56043e0d8fc3a4e46d17b50

Observation 55e87f8e-ea62-40da-8b71-d0126673ed3e · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.645438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:1b4d7fec209a6c4e2c47b87f4040ac7f76c0b549952cb836c22aad46fd909e05

Observation db4b10cf-1fbe-42fc-ae2d-a1756478861e · outbound

This paper cites A Survey on LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge A Survey on LLM-as-a-Judge

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:40:51.786196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:29cf776b3ebafb82e472ed8e8b48c13906021e934f898128203d113141598de7

Observation bf075786-5d17-432b-869a-6fcad4375361 · outbound

This paper cites Rating Roulette: Self-Inconsistency in.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Rating Roulette: Self-Inconsistency in

Reference 15

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.660492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:3681d756c9359d01fdce60b9cf5b14b6f8ad0d45c469895701ba5e5b72c9e182

Observation 50601554-e28d-432e-96e1-580e2c34ff1c · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.660667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:85ed5be4b84d3b7f45dee91b76f2ce74b183904c0b2346c8649cd6b3b64e85a5

Observation f4f4b314-965b-4a69-8b55-9fcd77f57253 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 17

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.677197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:85f03f1d9e7c855a8feb42181796b436b3796dddc6a87a521b6435b458cf6dd1

Observation b5c7d5c1-8204-4339-9c72-88fb52b3e571 · outbound

This paper cites InProceedings of the 2019 CHI Conference on Human Factors in Computing Systems.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge InProceedings of the 2019 CHI Conference on Human Factors in Computing Systems

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T19:40:45.681187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:b10ed3f8692f88fd90fb879ada8091fbea1a1779e1500b2f024be0876603e244

Observation a4db0292-b962-4748-b391-63151e5eddc8 · outbound

This paper cites A Survey on Human Preference Learning for Large Language Models.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge A Survey on Human Preference Learning for Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.769687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:01b0ee56ecaf765c63c54f3767fd0b93567f20c10012263f881d5538b5af8afc

Observation fd4b2506-0a61-42f4-aa5a-f64a21745a56 · outbound

This paper cites Johnson, Jennifer E.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Johnson, Jennifer E

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T04:30:37.649394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:210396d8fa69254bfe5dd78c9b4e0b158a6b2acfda742ef70f0aaf495c6a78ab

Observation 4c15ba4b-d167-4603-b233-3e8697d7e006 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.664282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:a5a660c4293c110cbfaed1282bebd5f685a944afc7a855bb89ab8b6aac512b8c

Observation c201a1e6-efd2-4a2b-aa98-f8ad1342f647 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 22

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.663277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:1b4a6d5c3996bf05d20f9dd21bd3fca62ff5cdb92bb12859e79281b58ccc8e65

Observation 8d11da68-f84b-42d3-9592-8b7e910a030b · outbound

This paper cites URL https: //www.pnas.org/doi/abs/10.1073/pnas.2415 697122.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge URL https: //www.pnas.org/doi/abs/10.1073/pnas.2415 697122

Reference 23

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.647583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:de8bb5f27fc5cca292afa1036d08c76343ebca2967bb98a186e9201fcb4a5e07

Observation 824cc18f-9be0-4592-a031-9de30e033a9d · outbound

This paper cites From generation to judgment: Opportunities and challenges of LLM-as-a-judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge From generation to judgment: Opportunities and challenges of LLM-as-a-judge

Reference 24

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.670863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:2c469a6ec48f5e792b7b4a05b1b0446401d6ae236f249b8cd1b215807bfae415

Observation 2ba0eb17-e0dd-46fd-9695-7ac96219d4d6 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:37.431680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:1c6e8571899b93218b6675eac31da418c0b8909ab5a93f1913e19a7225547b91

Observation 8102e053-c053-4b67-bd3d-d6dc4f24cfd6 · outbound

This paper cites Evaluating Scoring Bias in LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Evaluating Scoring Bias in LLM-as-a-Judge

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:03:47.800935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:23e2f48ce88ccd6816c639acce69bcade1bd3e48daca306314c9b39580e491e8

Observation 37ec9b1c-1c4c-444f-9ef0-5ce82bf4230c · outbound

This paper cites LLMs Cannot Reliably Judge (Yet?): A Comprehensive Assessment on the Robustness of LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge LLMs Cannot Reliably Judge (Yet?): A Comprehensive Assessment on the Robustness of LLM-as-a-Judge

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.670602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:798b4593394771e5248e4defa7521f6dd60405172ad65513a7ef674644c1cb80

Observation 83b4cb02-b702-4301-965f-acc244a000e3 · outbound

This paper cites Designing for Responsible Trust in AI Systems: A Communication Perspective.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Designing for Responsible Trust in AI Systems: A Communication Perspective

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:40:45.654523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:ede51c1feb67847fcf33a9074c81af25e279228f5bfaf1b8e7674a4db0f62c48

Observation 840b0dea-dd41-4f33-b352-41ec181f9447 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T22:55:50.972702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:c7fc99cf51981399d3aaf70ad7ed14355d948d9aaf6757100cb389ddcb91e8c4

Observation fd366746-7250-4011-809d-00549e858da4 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 30

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.650505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:89b137bba0a2692dac74eaa73602dd519ca516e4ab10db038ddaad2937b01ccc

Observation ab3ffc62-1552-4f01-8234-af9be4012561 · outbound

This paper cites InProceedings of the AAAI Conference on Artificial Intelligence, volume 38, pages 4171–4179.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge InProceedings of the AAAI Conference on Artificial Intelligence, volume 38, pages 4171–4179

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.731097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:10c7d9c9996dccb34a4a09d8cc9a04ab7105d2ba3a11835e1f65d327364215ea

Observation 1d431efc-21fa-468b-bc09-d3da4aae071d · outbound

This paper cites GPT-4 Technical Report.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge GPT-4 Technical Report

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:40:51.706199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:f485b98b6ae534e51156b3a15834f46eb1cd6168b9c3992ef53dda1e9bb415f3

Observation 5ddbabef-7c17-4039-bb25-c78e7c0ae631 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.656709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:00de0bea84bd6a721a25a17d86934b653804ab343dc19f62e594cb589e69c3a9

Observation 0c9b5d3e-e8c1-4c3b-8c83-c515b4f10e6f · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.652761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:e55a582a48efcfdc00b06941ad8332b499b282573756afd6e38e9c07edbb7bc8

Observation f3936e39-9b4d-481a-b4b7-67d57aa04e70 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.674263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:9787f2bc5943d3001aa6dd559a811dbd6363b725eb8b0264ecdadbf6ca2f3f49

Observation aa976e1d-07a5-41c1-be22-8e31410a2c53 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-05-16T04:30:37.645923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:66702c7641c7a408e8605ef25616021cdd7a79d52e49ca37c5cb09aa16f8ddce

Observation 2ef216c8-9d43-425e-bef8-14fa8868509c · outbound

This paper cites Rowley, Frances C.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Rowley, Frances C

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T04:30:37.667522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:2c524a89c500136e1fcd76b2a42f9eab2a760c49c2348f3e62dd34f0cc2d3683

Observation d02b100b-3dd7-4709-b705-b94181f5ab37 · outbound

This paper cites Ali, Angèle Christin, Andrew Smart, and Riitta Katila.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Ali, Angèle Christin, Andrew Smart, and Riitta Katila

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T19:40:45.667455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:bf2520cea7d2ddda263b3e3c943dc8b32e8094279ece5bbec651bde7147bee2d

Observation b2143566-bc8f-46b0-babc-aadb24da9869 · outbound

This paper cites Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.745869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:1e61cfc1a12f92588f586534dd7a6d4283455e9c658eb8765693d90e4d512f7d

Observation 2066a129-7a5d-489e-a368-b95074777215 · outbound

This paper cites Analyzing Uncertainty of LLM -as-a-Judge: Interval Evaluations with Conformal Prediction.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Analyzing Uncertainty of LLM -as-a-Judge: Interval Evaluations with Conformal Prediction

Reference 40

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.657596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:a38eab0a52855981d50238f86c6df7f9acaa3f6d7add9e2980d399bc6e57c698

Observation d573bc3d-09e6-463b-bc25-74df479dc81f · outbound

This paper cites Judging the judges: A systematic study of position bias in llm-as-a-judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Judging the judges: A systematic study of position bias in llm-as-a-judge

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.699055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:5c655ba1b0e177319abc9ec9802ae1e6ddea806396ef63c15e6ce06cc3bef836

Observation 4d2d3d9e-46ad-4bc4-8ec8-5f727af6279c · outbound

This paper cites Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.666192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:72cce09d52dc81b34c03abcba39c24c74c0705b178de279c3ac944c3dc36443c

Observation 9be947f1-4a2d-4551-b3cc-b779460aa926 · outbound

This paper cites doi: 10.1162/tacl_a_00685.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge doi: 10.1162/tacl_a_00685

Reference 43

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.674149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:5f365ce15e8c446507065357e6a8dd91b32e9b70a9a2201a123dc79bbde10175

Observation c37bc6d8-22b6-47f0-9ab2-6df4471574f2 · outbound

This paper cites Assessing Judging Bias in Large Reasoning Models: An Empirical Study.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Assessing Judging Bias in Large Reasoning Models: An Empirical Study

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.715805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:8b3570cb299e588da5d74c92b8fad2cd845687796ca4d855d313acef01830c72

Observation f3917525-caf0-4018-abc8-d69acf93800c · outbound

This paper cites Trust- judge: Inconsistencies of LLM-as-a-judge and how to alleviate them.arXiv preprint arXiv:2509.21117,.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Trust- judge: Inconsistencies of LLM-as-a-judge and how to alleviate them.arXiv preprint arXiv:2509.21117,

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.710379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:9f3400c0dedad76ee40cac60b2d23acd0b2b82ad9d749e417b2c0904a6d349d1

Observation e231de27-cbaa-40e7-b90a-c2de4c084f23 · outbound

This paper cites Self-Preference Bias in LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Self-Preference Bias in LLM-as-a-Judge

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:46:32.590484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:5ed8a862bbf1497de93d28dfc76890e8378fc36f46310f786fb70ad2bb856ad6

Observation 16700c6f-f671-4db3-b04e-fe4bb98e3a23 · outbound

This paper cites Attention is not not explanation.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Attention is not not explanation

Reference 47

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.644803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:62d1b783e49f3aae1836ae92aaa08b79ed5595ee1a6586b91c9daf3e65338ce4

Observation ca39f562-e034-4905-adc6-2f78af4da0f4 · outbound

This paper cites Torr, Bernard Ghanem, and Guohao Li.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Torr, Bernard Ghanem, and Guohao Li

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T04:30:37.642434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:9afd68571da76b12907320c3b65fdb7dcb3fea6dfb296cd0a0df004257d6fa99

Observation 5ff2e9ce-e592-4a0a-877e-c351d3384938 · outbound

This paper cites CHQ-Summ: A Dataset for Consumer Healthcare Question Summarization.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge CHQ-Summ: A Dataset for Consumer Healthcare Question Summarization

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.694620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:9e06246bcf3ce7640a8896b2a015a000c824736c24d192e4dea022078869247f

Observation 8769697d-4564-4c24-9fbe-20e191c5a429 · outbound

This paper cites an unresolved cited work.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Unresolved cited work

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:51.759792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:991fc40e032eee7544f53761d4f794c256f951ebd448afdc3afa8e2e3c170778

Observation ece987b0-a795-4e6f-a376-41d4c4f320ad · outbound

This paper cites Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:00:24.757706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:6e96948d4d65fd5209e323fc1f89a83b8f394468a6b90dbda33b3cf50161849d

Observation 796a9868-9c95-4778-80af-8e52c33ad9da · outbound

This paper cites , date =.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge , date =

Reference 52

Resolution
verified exact
doi, observed 2026-05-10T19:40:45.642098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:49fae3fda03881f64c34a3ccfaf7ae5d25c22c879b920e25b757bd73db136d20

Observation aee7455e-b161-4f5d-b82c-55b366002df0 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T22:40:51.686239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:1a87a08caacab37a540b3e85ef67d410b879f36aacff6e57773618821a3a0189

Observation 598d470d-9110-4cbb-8512-7b008e8e7461 · outbound

This paper cites online" 'onlinestring :=.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge online" 'onlinestring :=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T04:30:37.671038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:e47ded59cc0a4676c74987e846ecb369751370595d882c9be03ba3f31e0c7171

Observation 289700b1-1d06-46c5-bcaf-1ebc888d6758 · outbound

This paper cites write newline.

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge write newline

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T04:30:37.677639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T19:38:11.595077Z digest=sha256:d90a8596f1f87e83e0c353fbf7c82059f1ad8a99357017d691a59cc1d56d5ea6

Pith citing papers

Observation 376b07e2-fcd9-4555-b82e-6f56eeb7a457 · inbound

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs cites this paper.

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-06-29T05:43:08.670468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T05:36:21.058156Z digest=sha256:c534913cd9dadd3451d79a5697b10eb10ea9da64df024b13ea372bc124d653af