Pith. sign in

Paper Citation Record · LEDGER

BLEURT: Learning Robust Metrics for Text Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2004.04696.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2004.04696 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:23:28.458217Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T13:28:18.568194Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4748c6e0-df3c-4d25-a780-8acaf550ed37 · inbound

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate cites this paper.

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate BLEURT: Learning Robust Metrics for Text Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:03:18.843161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T13:03:18.765496Z digest=sha256:3e242fb17b1e8c9c6657eb1b2370bc35e2feaeec28c4f3573f5d05edf4758788

Observation daac12ba-004d-42f7-9d22-8bc25331a672 · inbound

TruthFlow: Truthful LLM Generation via Representation Flow Correction cites this paper.

TruthFlow: Truthful LLM Generation via Representation Flow Correction BLEURT: Learning Robust Metrics for Text Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T22:23:28.458217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:23:28.458217Z digest=sha256:5d3bd4334aa24fa2a85857aeaeff0ea4c3b3dff322a833b1bcd436b793b06dcc

Observation 553b8061-9c76-4f4c-97d7-80890ab36607 · inbound

Learning to Substitute Words with Model-based Score Ranking cites this paper.

Learning to Substitute Words with Model-based Score Ranking BLEURT: Learning Robust Metrics for Text Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T17:24:13.119093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:24:13.119093Z digest=sha256:13df55840dbd653b9da4e3a02e495b23eac206a8c7a5ec5e5c82523e7941de3f

Observation e641fbff-8315-4057-81d6-d7b574e2e94e · inbound

Secure LLM Fine-Tuning via Safety-Aware Probing cites this paper.

Secure LLM Fine-Tuning via Safety-Aware Probing BLEURT: Learning Robust Metrics for Text Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:11:35.885855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T13:07:09.402763Z digest=sha256:6beb30bb02c773b8a9851cadc4a8ab11a27d8db81c551c34ab9cba5532f89abd

Observation c1d45d57-7267-4b9a-9917-1c0c43ed79b8 · inbound

Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection cites this paper.

Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:39.845311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:50:39.845311Z digest=sha256:26279f25cc66507cbea84ad4e28db2602d981f63e98e910c5ac89e5b5d5480eb

Observation 8abaae69-2244-44d3-836c-e249d9e61625 · inbound

From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data cites this paper.

From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:35:03.812275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:35:03.812275Z digest=sha256:0139328d0529e06fecdb66f3ff4c61d17b96c016ab61cea5c1c926e389933643

Observation cc69d0a9-6165-46c7-a44f-931163779e6d · inbound

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations cites this paper.

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations BLEURT: Learning Robust Metrics for Text Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:43.738463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:23:43.738463Z digest=sha256:a6842e5862e072a0d266fafa9814f61eab77ecf829f41f640cc66c6a909e8db3

Observation 7146a338-e57b-45a4-b569-ff02429e57e1 · inbound

RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation cites this paper.

RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation BLEURT: Learning Robust Metrics for Text Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:31:58.726456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:31:58.726456Z digest=sha256:d7a7d7dbe44996cf18ac15ec1b2b7aa8f90a8a96fd9d933197ff5cd2ad9db88b

Observation 15cac342-1ad2-4f07-87aa-8ba3e0230c58 · inbound

Federated In-Context Learning: Iterative Refinement for Improved Answer Quality cites this paper.

Federated In-Context Learning: Iterative Refinement for Improved Answer Quality BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:38.740034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:42:38.740034Z digest=sha256:1223bb4292b649d912f4a6dcfd86b008a698307e8fd3756d626b366f77c0493c

Observation 3d5da4d0-b840-4da3-a59f-969e4b5f2476 · inbound

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy cites this paper.

Kvasir-VQA-x1: A Multimodal Dataset for Medical Reasoning and Robust MedVQA in Gastrointestinal Endoscopy BLEURT: Learning Robust Metrics for Text Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:42:02.689211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:42:02.689211Z digest=sha256:477690add4ba4112f3f9a7f24ab574434a620493eefdfe1f64e541ec28b727e1

Observation c7b411a6-e578-450d-9441-c3438dc892d6 · inbound

BioPars: A Pretrained Biomedical Large Language Model for Persian Biomedical Text Mining cites this paper.

BioPars: A Pretrained Biomedical Large Language Model for Persian Biomedical Text Mining BLEURT: Learning Robust Metrics for Text Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:29:46.327486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:29:46.327486Z digest=sha256:5f135988324107b3c0414ab08744e0e3cd3cd4fd8a893aaaa9dba9d0c97b8b95

Observation 531449bd-5594-4489-b6d9-a368e974c07b · inbound

Preserving Privacy, Increasing Accessibility, and Reducing Cost: An On-Device Artificial Intelligence Model for Medical Transcription and Note Generation cites this paper.

Preserving Privacy, Increasing Accessibility, and Reducing Cost: An On-Device Artificial Intelligence Model for Medical Transcription and Note Generation BLEURT: Learning Robust Metrics for Text Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:09.350124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:09.350124Z digest=sha256:f048207438ff272c2eb4c0305191a8624503e9dd94c9ec5160847f2bc74134f8

Observation cf087154-cc14-41c2-9010-7e91381fee32 · inbound

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework cites this paper.

Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework BLEURT: Learning Robust Metrics for Text Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:23:31.866770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:23:31.866770Z digest=sha256:fcd10a8bc28ea195f3a560be2b379b22ef426ef3b937547529274f13ac830e17

Observation 96b25d36-b084-4fa1-b2d8-fd095ae79f24 · inbound

Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation cites this paper.

Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T17:38:26.260688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:38:26.260688Z digest=sha256:71978c97a252f621e09549402a78f48c1aa42ec4e37c9636a102c349d936e2c1

Observation f2ea7918-8e3f-4ba1-b3bb-08b825cf6a36 · inbound

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice cites this paper.

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T14:52:43.086257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:52:43.086257Z digest=sha256:5e62b7afcd8ada1ab60ea8849a0587c444fa4793873ace99abd433a30d42a44d

Observation a60ece5c-845a-4e46-bbb4-36f0a2e19cf3 · inbound

LaQual: An Automated Framework for LLM App Quality Evaluation cites this paper.

LaQual: An Automated Framework for LLM App Quality Evaluation BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T16:25:19.860298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:25:19.860298Z digest=sha256:ec0504adf5f7eb8aca0fe1fe251fdf0778abcbac0e750f2c83126f72ccffa454

Observation ad36b03d-27b5-4a72-a015-b029487b2161 · inbound

ArgCMV: An Argument Summarization Benchmark for the LLM-era cites this paper.

ArgCMV: An Argument Summarization Benchmark for the LLM-era BLEURT: Learning Robust Metrics for Text Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T15:43:40.909985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:43:40.909985Z digest=sha256:c824511c2c307b0adb80c2c58b999b9e79501852ed2149209a826318510e675a

Observation 74ae44ef-0794-4b9c-ae4a-fddaf8dfc278 · inbound

How Small Transformation Expose the Weakness of Semantic Similarity Measures cites this paper.

How Small Transformation Expose the Weakness of Semantic Similarity Measures BLEURT: Learning Robust Metrics for Text Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T23:35:04.226541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:35:04.226541Z digest=sha256:90b868a92a64e1c0e9b89e9eea88c3318f98643c613d439613b68158d9754adb

Observation eaf1a47d-1aab-4f24-9415-c1ccb5c31bf4 · inbound

TabReX : Tabular Referenceless eXplainable Evaluation cites this paper.

TabReX : Tabular Referenceless eXplainable Evaluation BLEURT: Learning Robust Metrics for Text Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T21:28:34.066805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T21:24:44.941536Z digest=sha256:7b5b050af3914418f53a21c1ede5c73f662fc4f6e667be198769278f3a203ac8

Observation cea93c32-3842-4d25-ac0c-82dba72e5fe8 · inbound

On the Factual Consistency of Text-based Explainable Recommendation Models cites this paper.

On the Factual Consistency of Text-based Explainable Recommendation Models BLEURT: Learning Robust Metrics for Text Generation

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T15:50:19.414438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T15:45:20.394551Z digest=sha256:606249f08d50934f3dbd36c2abb5161b14b9ca95a4f6ddf6c25e81b7cd1eca7e

Observation c8b1f316-d9e5-45ad-af54-0cc5541e8488 · inbound

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications cites this paper.

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications BLEURT: Learning Robust Metrics for Text Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T06:46:26.342276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:46:26.342276Z digest=sha256:8c0b7cc3c1937a16d08c099e5ffd422931b262c98ea95bf84a08affe44daf06f

Observation b8193dd6-cd75-4ed3-bd60-ec4bec786bcd · inbound

MMP-Refer: Multimodal Path Retrieval-augmented LLMs For Explainable Recommendation cites this paper.

MMP-Refer: Multimodal Path Retrieval-augmented LLMs For Explainable Recommendation BLEURT: Learning Robust Metrics for Text Generation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:33:02.624066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T17:28:28.480792Z digest=sha256:613d8e1c2be552a0b3905a965edb33fdbbce90f3ae8989c14349173576773340

Observation d7e2c477-f15d-4e71-b780-edce5a8353b2 · inbound

Rank, Don't Generate: Statement-level Ranking for Explainable Recommendation cites this paper.

Rank, Don't Generate: Statement-level Ranking for Explainable Recommendation BLEURT: Learning Robust Metrics for Text Generation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:33:02.593278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T17:28:37.922342Z digest=sha256:6c101eb555af0b1d059a0171a3ce87cdfbac708e925dcae8b7eac7dd1deabfb5

Observation b0f3b587-3df3-47b9-9161-e0bebb062aa6 · inbound

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation cites this paper.

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation BLEURT: Learning Robust Metrics for Text Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:36:12.965564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:54:53.603591Z digest=sha256:fd26eabd0755123f032a3059e5fa944c5c0980d737977debe189b19b2574692a

Observation 4be422e9-e4ab-4220-9a0c-38f850a9f6ab · inbound

Calibrating Model-Based Evaluation Metrics for Summarization cites this paper.

Calibrating Model-Based Evaluation Metrics for Summarization BLEURT: Learning Robust Metrics for Text Generation

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:41:36.937721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T06:36:55.334742Z digest=sha256:d9a843d1213d1c25ef6ca793d97df0b1513bc90b119b0606c9940fb3b81ea20d

Observation 2cfb4fc8-9479-4ca2-945f-003181474048 · inbound

An Explainable Approach to Document-level Translation Evaluation with Topic Modeling cites this paper.

An Explainable Approach to Document-level Translation Evaluation with Topic Modeling BLEURT: Learning Robust Metrics for Text Generation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:11:06.101358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T23:11:59.716468Z digest=sha256:fc9e7d0911d397e1736b282f2a3d41265d1b8739a461d4c1d3b86f3136ba4085

Observation 73d497f4-c305-4a16-b118-66c0be8c6b3f · inbound

Evaluating Non-English Developer Support in Machine Learning for Software Engineering cites this paper.

Evaluating Non-English Developer Support in Machine Learning for Software Engineering BLEURT: Learning Robust Metrics for Text Generation

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:26:09.484138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T09:13:24.365852Z digest=sha256:0401903e2bb80bade970b51d31d764912b44fbdd926d9a25ec7b1e71730caf62

Observation 2199a87b-f254-4d2d-8ad9-0e1e54d1157e · inbound

LLM-Based User Personas for Recommendations at Scale cites this paper.

LLM-Based User Personas for Recommendations at Scale BLEURT: Learning Robust Metrics for Text Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.569695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T08:07:27.059673Z digest=sha256:e15800d38d3fd675879caf0ca8dcbfc7c301a69587af779c30d9ced4c6f30405

Observation be8025a3-ef3f-4b10-90b5-46bfdb523378 · inbound

LLM-Based User Personas for Recommendations at Scale cites this paper.

LLM-Based User Personas for Recommendations at Scale BLEURT: Learning Robust Metrics for Text Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T11:48:20.628483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:48:20.628483Z digest=sha256:95418bf7cb12befde3fa1cf518b290ca4ade32d4c460d16195065c538233b59a

Observation a3d58939-6ea6-4f89-91f3-e7be978f9b63 · inbound

TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluation cites this paper.

TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluation BLEURT: Learning Robust Metrics for Text Generation

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-08T04:27:34.047055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T04:27:34.047055Z digest=sha256:4f2decb90fe4fa5a68aecf281abdc5a5c3fc3b3320c5e24df9ca7132dfbb2cd3