Pith. sign in

Paper Citation Record · LEDGER

BLEURT: Learning Robust Metrics for Text Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2004.04696.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2004.04696 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:27:34.047055Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:28:18.568194Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4748c6e0-df3c-4d25-a780-8acaf550ed37 · inbound

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate cites this paper.

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate BLEURT: Learning Robust Metrics for Text Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:03:18.843161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T13:03:18.765496Z digest=sha256:22072a665298fa39844016ebbefd3941104939e21030fceb771551716f5d9a47

Observation e641fbff-8315-4057-81d6-d7b574e2e94e · inbound

Secure LLM Fine-Tuning via Safety-Aware Probing cites this paper.

Secure LLM Fine-Tuning via Safety-Aware Probing BLEURT: Learning Robust Metrics for Text Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:11:35.885855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T13:07:09.402763Z digest=sha256:afecc6c57d1eabcbe1d18a0bd85f76b55d7392ef26fc66cc379de46282ac69a9

Observation c1d45d57-7267-4b9a-9917-1c0c43ed79b8 · inbound

Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection cites this paper.

Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:39.845311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:50:39.845311Z digest=sha256:f565d41a13619a726dd8a1dcdd2749defd4dccee912d58e92eaa7ed797ccd78c

Observation 8abaae69-2244-44d3-836c-e249d9e61625 · inbound

From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data cites this paper.

From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:35:03.812275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:35:03.812275Z digest=sha256:d7d3f1f46b0f569f14c40eaa56d3b31822982b1000a3de720b350eb49b91dda7

Observation cc69d0a9-6165-46c7-a44f-931163779e6d · inbound

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations cites this paper.

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations BLEURT: Learning Robust Metrics for Text Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:43.738463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:23:43.738463Z digest=sha256:75060ed8f8b5155d76684f98e2b76a6a75569d34242e2dce9979a1cfd053938b

Observation 7146a338-e57b-45a4-b569-ff02429e57e1 · inbound

RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation cites this paper.

RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation BLEURT: Learning Robust Metrics for Text Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:31:58.726456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:31:58.726456Z digest=sha256:0564f89da2ab354a979b78ee537b90abef6308811986c2c64f3959d746a80e39

Observation 15cac342-1ad2-4f07-87aa-8ba3e0230c58 · inbound

Federated In-Context Learning: Iterative Refinement for Improved Answer Quality cites this paper.

Federated In-Context Learning: Iterative Refinement for Improved Answer Quality BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:38.740034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:42:38.740034Z digest=sha256:f1dd470fe07c3e8ee9ab5bb7ea4879abad1d70af4466761874d46bce22611a41

Observation 3d5da4d0-b840-4da3-a59f-969e4b5f2476 · inbound

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy cites this paper.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEURT: Learning Robust Metrics for Text Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.689211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.689211Z digest=sha256:0f246fcbe51f2c73ae64fd0a34e4eddfa47dc2d01e8d4a4e466dcb2a73296541

Observation c7b411a6-e578-450d-9441-c3438dc892d6 · inbound

BioPars: A Pretrained Biomedical Large Language Model for Persian Biomedical Text Mining cites this paper.

BioPars: A Pretrained Biomedical Large Language Model for Persian Biomedical Text Mining BLEURT: Learning Robust Metrics for Text Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:29:46.327486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:29:46.327486Z digest=sha256:4f7045b3595430138dd314b487df63bf50eb93ffd42cf3a67354c2503838a60a

Observation 531449bd-5594-4489-b6d9-a368e974c07b · inbound

Preserving Privacy, Increasing Accessibility, and Reducing Cost: An On-Device Artificial Intelligence Model for Medical Transcription and Note Generation cites this paper.

Preserving Privacy, Increasing Accessibility, and Reducing Cost: An On-Device Artificial Intelligence Model for Medical Transcription and Note Generation BLEURT: Learning Robust Metrics for Text Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:09.350124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:09.350124Z digest=sha256:512b3f7781199fb4eebe8bce16729e8f9472dab56c5953194135d122e691d524

Observation cf087154-cc14-41c2-9010-7e91381fee32 · inbound

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework cites this paper.

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework BLEURT: Learning Robust Metrics for Text Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:23:31.866770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:23:31.866770Z digest=sha256:916af0e84137ef6cb09da653c0775df215ac9e1be2507927f7464f9b600d1c5d

Observation 96b25d36-b084-4fa1-b2d8-fd095ae79f24 · inbound

Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation cites this paper.

Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T17:38:26.260688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:38:26.260688Z digest=sha256:d5aa5a8d3428ded923034f136e7b07b4acb6c3f4ed627dde5fbd29b7dbe3eaa5

Observation f2ea7918-8e3f-4ba1-b3bb-08b825cf6a36 · inbound

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice cites this paper.

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T14:52:43.086257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:52:43.086257Z digest=sha256:e797e252d0519dc432765ac9eb83c244f556418d5c91740e1d542fb053374d5d

Observation a60ece5c-845a-4e46-bbb4-36f0a2e19cf3 · inbound

LaQual: An Automated Framework for LLM App Quality Evaluation cites this paper.

LaQual: An Automated Framework for LLM App Quality Evaluation BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T16:25:19.860298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:25:19.860298Z digest=sha256:1ca96b7961e1f008624f2c4b3c283ee3d348fc54096f6057b6c06ba3fd0c3e5f

Observation ad36b03d-27b5-4a72-a015-b029487b2161 · inbound

ArgCMV: An Argument Summarization Benchmark for the LLM-era cites this paper.

ArgCMV: An Argument Summarization Benchmark for the LLM-era BLEURT: Learning Robust Metrics for Text Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T15:43:40.909985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:43:40.909985Z digest=sha256:6b80f94b216c0e48b5e1347c3d48c42c640b42adff8c51999b84dabb9dcf8953

Observation 74ae44ef-0794-4b9c-ae4a-fddaf8dfc278 · inbound

How Small Transformation Expose the Weakness of Semantic Similarity Measures cites this paper.

How Small Transformation Expose the Weakness of Semantic Similarity Measures BLEURT: Learning Robust Metrics for Text Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T23:35:04.226541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:35:04.226541Z digest=sha256:50773fd12d599d0afeafebc43cd14d1b5b3eb5ebd1e5f3c59a2e63cc526fe8ff

Observation eaf1a47d-1aab-4f24-9415-c1ccb5c31bf4 · inbound

TabReX : Tabular Referenceless eXplainable Evaluation cites this paper.

TabReX : Tabular Referenceless eXplainable Evaluation BLEURT: Learning Robust Metrics for Text Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T21:28:34.066805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T21:24:44.941536Z digest=sha256:ef23c2ea0d4f4e5408def45074c54298d51bbc12118647a0d7ddbd8f3b23d067

Observation cea93c32-3842-4d25-ac0c-82dba72e5fe8 · inbound

On the Factual Consistency of Text-based Explainable Recommendation Models cites this paper.

On the Factual Consistency of Text-based Explainable Recommendation Models BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T15:50:19.414438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T15:45:20.394551Z digest=sha256:af856fdf3cdad2bb8ae71892385fb9423dbfb3400d7fa271437b5d630a546bfb

Observation c8b1f316-d9e5-45ad-af54-0cc5541e8488 · inbound

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications cites this paper.

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications BLEURT: Learning Robust Metrics for Text Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T06:46:26.342276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:46:26.342276Z digest=sha256:5a4ecfdcea6b28a04ff20ebb050e5bb2977046051179df079264003a9ea5fd6d

Observation b8193dd6-cd75-4ed3-bd60-ec4bec786bcd · inbound

MMP-Refer: Multimodal Path Retrieval-augmented LLMs For Explainable Recommendation cites this paper.

MMP-Refer: Multimodal Path Retrieval-augmented LLMs For Explainable Recommendation BLEURT: Learning Robust Metrics for Text Generation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:33:02.624066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T17:28:28.480792Z digest=sha256:4ba0e5cd1fed41fb17c2010b9a4a24e7819ecdfc022da849a9ff042e93acc5d9

Observation d7e2c477-f15d-4e71-b780-edce5a8353b2 · inbound

Rank, Don't Generate: Statement-level Ranking for Explainable Recommendation cites this paper.

Rank, Don't Generate: Statement-level Ranking for Explainable Recommendation BLEURT: Learning Robust Metrics for Text Generation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:33:02.593278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T17:28:37.922342Z digest=sha256:15bc37fbfc130cb07c0005ba1d56253d8ea266e82e035a97bcf9ce75065fc7ee

Observation b0f3b587-3df3-47b9-9161-e0bebb062aa6 · inbound

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation cites this paper.

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation BLEURT: Learning Robust Metrics for Text Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:36:12.965564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:54:53.603591Z digest=sha256:d5aa3c09677221132c3692d187d6909f4d1b66c546e288c302d42d7562bf45cc

Observation 4be422e9-e4ab-4220-9a0c-38f850a9f6ab · inbound

Calibrating Model-Based Evaluation Metrics for Summarization cites this paper.

Calibrating Model-Based Evaluation Metrics for Summarization BLEURT: Learning Robust Metrics for Text Generation

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:41:36.937721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T06:36:55.334742Z digest=sha256:761b392805b08d956ba01e1e4a82d4b633053ccc65e90a8932eabf40e166653a

Observation 2cfb4fc8-9479-4ca2-945f-003181474048 · inbound

An Explainable Approach to Document-level Translation Evaluation with Topic Modeling cites this paper.

An Explainable Approach to Document-level Translation Evaluation with Topic Modeling BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:11:06.101358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T23:11:59.716468Z digest=sha256:7de1f4a131005be1c262c28c598bce5f1890e84dab444060dedf51fbf4d4921d

Observation 73d497f4-c305-4a16-b118-66c0be8c6b3f · inbound

Evaluating Non-English Developer Support in Machine Learning for Software Engineering cites this paper.

Evaluating Non-English Developer Support in Machine Learning for Software Engineering BLEURT: Learning Robust Metrics for Text Generation

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:26:09.484138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T09:13:24.365852Z digest=sha256:aa0f8e0a85811d6003c1a5b3bc309f51be0017e739c97213522bb88537f1224e

Observation 2199a87b-f254-4d2d-8ad9-0e1e54d1157e · inbound

LLM-Based User Personas for Recommendations at Scale cites this paper.

LLM-Based User Personas for Recommendations at Scale BLEURT: Learning Robust Metrics for Text Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.569695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T08:07:27.059673Z digest=sha256:e95ca5ec5f0e15698226a39a39b18f49575948e076584f1e8ee1fd2da26e4517

Observation be8025a3-ef3f-4b10-90b5-46bfdb523378 · inbound

LLM-Based User Personas for Recommendations at Scale cites this paper.

LLM-Based User Personas for Recommendations at Scale BLEURT: Learning Robust Metrics for Text Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T11:48:20.628483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:48:20.628483Z digest=sha256:73b310dd439b2321d6120fd3b88711e9c29673c57436ed91452c82645298843a

Observation a3d58939-6ea6-4f89-91f3-e7be978f9b63 · inbound

TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluation cites this paper.

TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluation BLEURT: Learning Robust Metrics for Text Generation

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-08T04:27:34.047055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T04:27:34.047055Z digest=sha256:cd0be1f152b69588cc74ee30103550d5f73885e96ae193efdff79b822bd3fe4a