Pith. sign in

Paper Citation Record · LEDGER

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges

As of 19 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2608.12097.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.12097 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:20:37.105573Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ea13104-9e1f-4c2e-ae98-96780e216643 · outbound

This paper cites AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:36.998626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:36.998626Z digest=sha256:eb75fbb8561b7d9611ee5ac608f47bffea92442e88bdbb8174edde224cc4502b

Observation 5688d16f-8212-4af7-a5c8-9b31db5606f3 · outbound

This paper cites From Rubrics to Reliable Scores: Evidence-Grounded Text Evaluation with LLM Judges.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges From Rubrics to Reliable Scores: Evidence-Grounded Text Evaluation with LLM Judges

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.008172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.008172Z digest=sha256:76c94d534f6062726b65cccbea651dc5e6f76aa1e80e348152e23ef2a2fba3c0

Observation 93835941-1fb3-4ded-b88a-dff9c610b23b · outbound

This paper cites Seungone Kim, Jamin Shin, Yejin Cho, Joel Jang, Shayne Longpre, Hwaran Lee, Sangdoo Yun, Seongjin Shin, Sungdong Kim, James Thorne, and Minjoon Seo.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Seungone Kim, Jamin Shin, Yejin Cho, Joel Jang, Shayne Longpre, Hwaran Lee, Sangdoo Yun, Seongjin Shin, Sungdong Kim, James Thorne, and Minjoon Seo

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.017082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.017082Z digest=sha256:ff758f0addd7bc96c468ae684852149c0a3c32169142dd952aacf7251a596280

Observation b148c48d-9f69-42fc-8dac-a5e8271fb0cd · outbound

This paper cites URL https://aclanthology.org/2025.emnlp-main.796/.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https://aclanthology.org/2025.emnlp-main.796/

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.025326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.025326Z digest=sha256:d22524c1eca48a3ee80b75be2617e8aa35f5f048787ae02559c4e30ba6c19ebd

Observation 3707d863-dfb9-4e4d-aa18-063615f46400 · outbound

This paper cites URL https://aclanthology.org/2025.acl-long.513/.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https://aclanthology.org/2025.acl-long.513/

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.029615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.029615Z digest=sha256:380894d2ecf886307068e301e855aa5e98234a6e04465c0a6aa5a9662965c3f0

Observation 0477cb5e-356e-4009-8406-894d54e79533 · outbound

This paper cites an unresolved cited work.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:20:37.609232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T00:20:37.038213Z digest=sha256:f2ded0f864d1147c756825ac8d7b49bd0e595a8c371e02b6d3f7296cf70dc5cf

Observation d88b83c3-bc5b-4bd7-9434-d51f03172e6e · outbound

This paper cites URL https: //aclanthology.org/2023.emnlp-main.153/.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https: //aclanthology.org/2023.emnlp-main.153/

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.042254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.042254Z digest=sha256:122c1f0508ca7fcfed387cf4d747df4da53ec282a7b69e87ae53f1d4a6b3b826

Observation 285b3382-0df7-481e-8342-2f5761dcfe33 · outbound

This paper cites Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.051760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.051760Z digest=sha256:664f3593f2b53c300888c0039693e917dd45cd6f40d66f8f178581042fed2045

Observation d8d0665d-5685-4def-b29f-ac063fd0e4ef · outbound

This paper cites Rickard Stureborg, Dimitris Alikaniotis, and Yoshi Suhara.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Rickard Stureborg, Dimitris Alikaniotis, and Yoshi Suhara

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.061638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.061638Z digest=sha256:5023ff5de729abc27f47f380f84a9bc80b7dcda438c83768d54317bace06321e

Observation 807b6c1a-b1ab-4632-9b10-a9d5c2a3236a · outbound

This paper cites Large Language Models are Inconsistent and Biased Evaluators.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Large Language Models are Inconsistent and Biased Evaluators

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.066383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.066383Z digest=sha256:bed8654697c9e54bcef80a905f075e595d0c29ec71c5bbde1bb3fbda260620cf

Observation f4c296c7-5823-43b2-a017-8784d214ab55 · outbound

This paper cites Replacing Judges with Juries: Evaluating LLM Generations with a Panel of Diverse Models.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Replacing Judges with Juries: Evaluating LLM Generations with a Panel of Diverse Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.071621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.071621Z digest=sha256:3a6e8f5885af67c94c86990ecd1bc031924ecfd467250cd9212c1b5e647cceaa

Observation 1ca9de8b-fcc1-4300-96e5-f49afa891341 · outbound

This paper cites Large Language Models are not Fair Evaluators.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Large Language Models are not Fair Evaluators

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.076335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.076335Z digest=sha256:af2c0f9049bbe6aefe87b9a78ef5410128660148b024aeed5367013603ef4550

Observation d65a90db-1711-4dd1-98a1-b23cc6dcd4d0 · outbound

This paper cites URL https://papers.nips.cc/paper_files/paper/2024/hash/ 02fd91a387a6a5a5751e81b58a75af90-Abstract-Datasets_and_Benchmarks_Track.html.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https://papers.nips.cc/paper_files/paper/2024/hash/ 02fd91a387a6a5a5751e81b58a75af90-Abstract-Datasets_and_Benchmarks_Track.html

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.080570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.080570Z digest=sha256:5398ac0653e21bb32ab621f51ff7c51d9238d65e9274ae6a98e1257e4f6fda59

Observation 5b3ce199-c30c-4473-b838-23e0d9ca61f4 · outbound

This paper cites Seonghyeon Ye, Doyoung Kim, Sungdong Kim, Hyeonbin Hwang, Seungone Kim, Yongrae Jo, James Thorne, Juho Kim, and Min- joon Seo.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Seonghyeon Ye, Doyoung Kim, Sungdong Kim, Hyeonbin Hwang, Seungone Kim, Yongrae Jo, James Thorne, Juho Kim, and Min- joon Seo

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:20:37.594252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T00:20:37.084631Z digest=sha256:e2997a0f420c9338ac142370d0abd5285854db2b8ff5009893de71f549f5e1bf

Observation 2a897c6f-6631-45ff-af95-11fbf81595d0 · outbound

This paper cites Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:20:37.581032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T00:20:37.091043Z digest=sha256:265f67d0810c80aa1ccd227a039a4eca39249a247651f6be8e7fadad78a94174

Observation c3999291-ec4f-4527-9de1-1017005f2ebe · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.095446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.095446Z digest=sha256:efe2a9391dc0695b28efc6aeeaa8b34c69c9b2ca49cf5796b80b93c4565cc250

Observation d8f3a760-f420-45f4-82ea-70260393199a · outbound

This paper cites URL https://aclanthology.org/2026.acl-long.1439/.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https://aclanthology.org/2026.acl-long.1439/

Reference 28

Resolution
verified exact
doi, observed 2026-08-16T00:20:37.140527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T00:20:37.100966Z digest=sha256:c91535bfbca27f4ea11f8f1b57cde5d25225dd3098507fbad940f1ae16d6ad5f

Observation 9c657517-5da5-48d7-8421-6ce748f4d94a · outbound

This paper cites an unresolved cited work.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-16T00:20:37.566899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T00:20:37.105573Z digest=sha256:7220875d8f425f450b2e17ba6f85eed66f0d07881ae904599a2f7f7491ec45b4

Observation ae35214c-d880-4aa9-a780-122b69c03bfd · outbound

This paper cites URL https://aclanthology.org/2022.acl-long.229/.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https://aclanthology.org/2022.acl-long.229/

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.033567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.033567Z digest=sha256:256b0db2535eee9b66b71bfe629afcf9ae46b782b16252db3acb2545553a7201

Observation 7528c025-86d9-461e-8c46-c72c310779bd · outbound

This paper cites ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:36.979575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:36.979575Z digest=sha256:533d5da528060602adb800ea2e4d6b020a403f5243ff44f64c35217b25f52c1e

Observation c6115cf0-3004-4596-8ea6-13abc11fca11 · outbound

This paper cites TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:36.988654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:36.988654Z digest=sha256:059cb5971eefd198b603db5abeb227947252a9cdd2013f5d409aeceab7719387

Observation dbac2e59-ee3f-4c35-9adb-020f1891a3c1 · outbound

This paper cites URL https://aclanthology.org/2025.naacl- long.303/.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges URL https://aclanthology.org/2025.naacl- long.303/

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-16T00:20:37.021116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:20:37.021116Z digest=sha256:f88cccb4800b030ef333e5d684e00f7c4a56a905eae0b63bd71c30efdca5cb40

Observation dcbede6f-92d7-4af4-8c65-1e54be86a6d2 · outbound

This paper cites Accessed 2026-07-25.

Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges Accessed 2026-07-25

Reference 2026

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:20:37.622043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T00:20:36.983518Z digest=sha256:b7c5db6956c829bd9bba510565c706d83e8405cea1e62fb17926ba6dda9dba9a

Pith citing papers

No inbound Pith citation observations are available.