Pith. sign in

Paper Citation Record · LEDGER

A Comprehensive Assessment of Dialog Evaluation Metrics

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2106.03706.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2106.03706 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:08:11.541632Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T09:17:40.306983Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d1937a42-aa18-4a11-b9c3-e79e2415899b · inbound

Towards Automatic Evaluation of Task-Oriented Dialogue Flows cites this paper.

Towards Automatic Evaluation of Task-Oriented Dialogue Flows A Comprehensive Assessment of Dialog Evaluation Metrics

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T19:44:04.030853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:44:04.030853Z digest=sha256:46ba1e0e16b23d4f4c8d247014950b63e361a638446abb2f0920cb23b20a1539

Observation fc316826-c338-488b-88c9-a5f96bb579d8 · inbound

Towards Understanding the Robustness of LLM-based Evaluations under Perturbations cites this paper.

Towards Understanding the Robustness of LLM-based Evaluations under Perturbations A Comprehensive Assessment of Dialog Evaluation Metrics

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T17:09:54.164611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:09:54.164611Z digest=sha256:714a1d007fa59d7a4f564ea3a96ce57d00656d8c5d148b70f8f00dcc790110af

Observation 74478500-8350-4ab8-af79-28af5d0e775e · inbound

clem:todd: A Framework for the Systematic Benchmarking of LLM-Based Task-Oriented Dialogue System Realisations cites this paper.

clem:todd: A Framework for the Systematic Benchmarking of LLM-Based Task-Oriented Dialogue System Realisations A Comprehensive Assessment of Dialog Evaluation Metrics

Reference 2763

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:11.541632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:11.541632Z digest=sha256:b535d7c47b3dd89c1cb61ba53da95b870d030768ea594e4a2afc9513a1ac0710

Observation 8b89eae5-98eb-45d6-941c-60f05c24201c · inbound

RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation cites this paper.

RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation A Comprehensive Assessment of Dialog Evaluation Metrics

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:17:40.308681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T09:13:55.204404Z digest=sha256:0ef249cf706415a851a5bd171fb43b7df323e7e836b8a16095d535ac14ce4887