Pith. sign in

Paper Citation Record · LEDGER

mT5: A massively multilingual pre-trained text-to-text transformer

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 63 inbound Pith citation observations for arXiv:2010.11934.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.11934 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 63 of 63 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:11:29.436431Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T02:24:27.918351Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3d6c65f2-2767-4b21-b0d0-c4d4fb289b60 · inbound

Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity cites this paper.

Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity mT5: A massively multilingual pre-trained text-to-text transformer

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:57:10.996165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T23:57:10.953617Z digest=sha256:35265d6085db5c35ca899a2bf978ae675a48e636be3244648d8a9296b84dd4c4

Observation 9b7ec0ae-4914-4492-8177-943d9f78b875 · inbound

Deduplicating Training Data Makes Language Models Better cites this paper.

Deduplicating Training Data Makes Language Models Better mT5: A massively multilingual pre-trained text-to-text transformer

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-24T13:39:31.676832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-24T13:36:55.210708Z digest=sha256:3f5945c685d625e95f12c6c921b06e3729bc4fb143026f0bd240ab3d34d81bf9

Observation 8b4bf5cc-87e1-4d14-9b52-ffd82c617f80 · inbound

ST-MoE: Designing Stable and Transferable Sparse Expert Models cites this paper.

ST-MoE: Designing Stable and Transferable Sparse Expert Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 209

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:14:25.934874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T23:14:25.431471Z digest=sha256:bda4b7e55fad2ce8c75f5d8b8f3447e7e0642f86201d7e028a8381749f2c7404

Observation a113adfe-597b-44fc-b4b2-cbe3d27f9285 · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model mT5: A massively multilingual pre-trained text-to-text transformer

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:34:28.206982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:b072a165495e05a17d1491c5886747564509c9d76a43f623906cfcc14d946bdd

Observation 378d607c-89b0-4281-a393-ec096bd40e66 · inbound

Scaling Data-Constrained Language Models cites this paper.

Scaling Data-Constrained Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 134

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:35:21.493501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T01:35:21.150772Z digest=sha256:75623d32ab67b652cae77ca42e9fcb3033394ae12603f01e7105a7b580305e11

Observation 92cf3530-5d7c-4fd4-aff0-c2c114e2e869 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.016371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:9bab83c53ed44aff955a60767f4587bf1dd74973c0cbdf489d70e0582399d59b

Observation a3ccf017-b4d7-4e0c-988f-1ed2bd7df681 · inbound

The Falcon Series of Open Language Models cites this paper.

The Falcon Series of Open Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:46:10.093297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-16T09:46:09.701440Z digest=sha256:0a6a4da8f1b862016961e28410cb09102d54c315bb89b54af93547a0ed00f5a4

Observation 6cfccf6a-7bf8-479d-9655-eb4e1a6e4697 · inbound

M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation cites this paper.

M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation mT5: A massively multilingual pre-trained text-to-text transformer

Reference 63

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:39:03.436766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T22:39:02.540687Z digest=sha256:83e89294929c98e6e1f686060a22ac320a18a63e79c45d6e791a29b05d316dde

Observation 0efa15e2-031e-4383-824e-71f0b8cc19ff · inbound

Large Language Models: A Survey cites this paper.

Large Language Models: A Survey mT5: A massively multilingual pre-trained text-to-text transformer

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:22:54.894359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T15:22:54.023279Z digest=sha256:a5c0b1a35d57a2af10c160d235dae08575818fe230d0c9e362ac2f94650a3598

Observation efd999da-cccc-44db-948f-0d513df43133 · inbound

Vision-Braille: A Curriculum Learning Toolkit and Braille-Chinese Corpus for Braille Translation cites this paper.

Vision-Braille: A Curriculum Learning Toolkit and Braille-Chinese Corpus for Braille Translation mT5: A massively multilingual pre-trained text-to-text transformer

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T23:03:34.716637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T22:58:37.339500Z digest=sha256:dc692f7bd7988b3350cc3ecdefbf7155783df9803afa95a6bc13eaf2c34f2dd8

Observation 08d3f6bf-66d3-4eaa-a3c8-973608e11bf3 · inbound

PaliGemma: A versatile 3B VLM for transfer cites this paper.

PaliGemma: A versatile 3B VLM for transfer mT5: A massively multilingual pre-trained text-to-text transformer

Reference 152

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:10:21.261356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T13:10:19.972353Z digest=sha256:69f95db6d14d52292127f3fc66edb2b3039284e610650b9be9abab5592cd7218

Observation c35a5125-02bf-4f35-8664-0275955e803f · inbound

Gemma 2: Improving Open Language Models at a Practical Size cites this paper.

Gemma 2: Improving Open Language Models at a Practical Size mT5: A massively multilingual pre-trained text-to-text transformer

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:11:16.433596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T12:11:16.326752Z digest=sha256:a60e3faf71d589155dd3b8088958429ff1cca332771c825ee435bf3ce758cb2f

Observation 9ab36c5c-65c7-413a-809e-11efef6a3673 · inbound

Open-Sora Plan: Open-Source Large Video Generation Model cites this paper.

Open-Sora Plan: Open-Source Large Video Generation Model mT5: A massively multilingual pre-trained text-to-text transformer

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.201690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:67dc90477fb095cdcf0ee07206395a677f5572218fb5340f89eb04d7ac61a3be

Observation 08459cb7-3120-447a-9fb4-6d3a0714984d · inbound

Empowering Bengali Education with AI: Solving Bengali Math Word Problems through Transformer Models cites this paper.

Empowering Bengali Education with AI: Solving Bengali Math Word Problems through Transformer Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:11:29.436431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:11:29.436431Z digest=sha256:cacf7c53c8a7aa311bf57723325071321b848f70c69eb351a023a58f642c43ea

Observation aa42bbab-a37c-41f7-9d00-04b4c75d7e5e · inbound

Visual question answering: from early developments to recent advances -- a survey cites this paper.

Visual question answering: from early developments to recent advances -- a survey mT5: A massively multilingual pre-trained text-to-text transformer

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-10T21:46:28.586985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:46:28.586985Z digest=sha256:07d18bb49564453cf36c385ad7a209599d509b2c57cce0376bc5086b69b8e222

Observation 30c50a51-55fc-41c9-a5d8-1855a1f8b24a · inbound

Multilingual Open QA on the MIA Shared Task cites this paper.

Multilingual Open QA on the MIA Shared Task mT5: A massively multilingual pre-trained text-to-text transformer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:28.604951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:43:28.604951Z digest=sha256:08baec145771e77cfeeb101cc51acf97a9f1990943edb7aedf173095a068acef

Observation 63bd1df1-3383-4e9e-85b3-9f7af049e4b7 · inbound

Analysis of Indic Language Capabilities in LLMs cites this paper.

Analysis of Indic Language Capabilities in LLMs mT5: A massively multilingual pre-trained text-to-text transformer

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T15:31:10.796233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:31:10.796233Z digest=sha256:f9996a1fa60e0d0520de84d53f377367cd64dd494cbae666def182c19cf237dd

Observation 19cf54c2-587a-4ec3-b38a-a66ce02e9b12 · inbound

Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction cites this paper.

Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction mT5: A massively multilingual pre-trained text-to-text transformer

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T14:33:23.468892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:33:23.468892Z digest=sha256:669728f78706116a78625f9fc58ce511f12628145bf917b229ea85540d3e033f

Observation 246fcfc3-7ce5-48c6-8cba-65e85858e1b1 · inbound

STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection cites this paper.

STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection mT5: A massively multilingual pre-trained text-to-text transformer

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T14:21:24.253415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:21:24.253415Z digest=sha256:7b8db83673af19bf26a589c33cebf7feab95b343e195a07573a8d1a5d748a2e3

Observation 09b94bc8-dde0-43fd-b8fc-76d482a9eb51 · inbound

Commute Your Domains: Trajectory Optimality Criterion for Multi-Domain Learning cites this paper.

Commute Your Domains: Trajectory Optimality Criterion for Multi-Domain Learning mT5: A massively multilingual pre-trained text-to-text transformer

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T14:26:11.882025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:26:11.882025Z digest=sha256:957c41fdf2c8fba78a39b337e63ff094693a4992792c447692e8e8dabf31fe32

Observation b502ac6b-e0fa-4578-8361-6113c63ae749 · inbound

A Method for Multi-Hop Question Answering on Persian Knowledge Graph cites this paper.

A Method for Multi-Hop Question Answering on Persian Knowledge Graph mT5: A massively multilingual pre-trained text-to-text transformer

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T18:59:58.203120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:59:58.203120Z digest=sha256:f0a78b57bf9fafe1e3db506c9d99a228812ef0dc5657b08e9dd4834efe29daa8

Observation 448d7c2e-5a14-4bbb-82c9-e8a770fc9daa · inbound

Towards the Development of Balanced Synthetic Data for Correcting Grammatical Errors in Arabic: An Approach Based on Error Tagging Model and Synthetic Data Generating Model cites this paper.

Towards the Development of Balanced Synthetic Data for Correcting Grammatical Errors in Arabic: An Approach Based on Error Tagging Model and Synthetic Data Generating Model mT5: A massively multilingual pre-trained text-to-text transformer

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T19:53:30.153016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:53:30.153016Z digest=sha256:8a9594158e641d91603968fbd466fd530ebdcb1329bcd0f5b958a8672cb0ef47

Observation da45e496-bf4d-48ae-91b9-c1f1b5849839 · inbound

Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile cites this paper.

Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile mT5: A massively multilingual pre-trained text-to-text transformer

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T16:36:00.821782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T16:36:00.821782Z digest=sha256:6a5d0698cdc9406e16a0c141d250ff1c4f6acfb4391e4cc8e8c5f7e4bad7e2b9

Observation 0bd3df9e-7d8b-44d7-a9ca-df8be45cc50c · inbound

PiKE: Adaptive Data Mixing for Large-Scale Multi-Task Learning Under Low Gradient Conflicts cites this paper.

PiKE: Adaptive Data Mixing for Large-Scale Multi-Task Learning Under Low Gradient Conflicts mT5: A massively multilingual pre-trained text-to-text transformer

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T16:20:38.426348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:20:38.426348Z digest=sha256:73244f20d296b470d44e9107514c48544375cca4a13e54c8ec57d382199d83b7

Observation 93d28bf7-9618-4927-b599-0a3ece5f364d · inbound

Scaling Pre-training to One Hundred Billion Data for Vision Language Models cites this paper.

Scaling Pre-training to One Hundred Billion Data for Vision Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T12:12:30.815289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:12:30.815289Z digest=sha256:1cb3a81c3fd8c617e853793831f4604e75b683b6a1c6e999b6695e142e1c6317

Observation 00c4a738-7e39-4c19-b7cf-4febec0037d1 · inbound

Training Sparse Mixture Of Experts Text Embedding Models cites this paper.

Training Sparse Mixture Of Experts Text Embedding Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:23.703288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:23.703288Z digest=sha256:56f4f7cf52c74240c5cb1f056e1d4b96438d01eae7437d98e74b007794d7b2b8

Observation 641210eb-c6fd-4c8d-8581-90fb0d8d5c19 · inbound

Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs cites this paper.

Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs mT5: A massively multilingual pre-trained text-to-text transformer

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T21:54:37.229166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T21:54:37.229166Z digest=sha256:2b0596f4f284cd5cd92f734dc6676d440a11a0d4202acf854fe11fc754f9949c

Observation 81ed8a27-e6dc-4e41-adca-98bd0a495925 · inbound

LAGO: Few-shot Crosslingual Embedding Inversion Attacks via Language Similarity-Aware Graph Optimization cites this paper.

LAGO: Few-shot Crosslingual Embedding Inversion Attacks via Language Similarity-Aware Graph Optimization mT5: A massively multilingual pre-trained text-to-text transformer

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:12:20.755310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:12:20.755310Z digest=sha256:4c244b4e4c2803b4ccd4843cf4afa6d22f61c2ed6fb1f3b260b6e353a41baa74

Observation 0aa85d15-362d-4db0-bf46-17688cfd33c4 · inbound

SELF: Self-Extend the Context Length With Logistic Growth Function cites this paper.

SELF: Self-Extend the Context Length With Logistic Growth Function mT5: A massively multilingual pre-trained text-to-text transformer

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.201414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:52:15.201414Z digest=sha256:634b189bc3fcd427fe790df1a663bbcc127d94ae30589ed8e1ec137b606b8c92

Observation b53453b8-f840-4739-b9fe-74a96e924467 · inbound

Mutarjim: Advancing Bidirectional Arabic-English Translation with a Small Language Model cites this paper.

Mutarjim: Advancing Bidirectional Arabic-English Translation with a Small Language Model mT5: A massively multilingual pre-trained text-to-text transformer

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:32.260670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:32.260670Z digest=sha256:f40dc3d68ef62ebc24d7dedf022ec84b29ea80e1734eb38c6aad6dc890f6c547

Observation 7cc003f6-decb-4014-9af3-e8d319b59b4f · inbound

Towards Multi-dimensional Evaluation of LLM Summarization across Domains and Languages cites this paper.

Towards Multi-dimensional Evaluation of LLM Summarization across Domains and Languages mT5: A massively multilingual pre-trained text-to-text transformer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:08:57.258491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:08:57.258491Z digest=sha256:b4ab98758a16275707599466758d5671b853e7c57ae507b4623c8716ac9fed63

Observation 1fa9cc3b-c1a4-4f70-90e2-74942e8be7c9 · inbound

Pi-SQL: Enhancing Text-to-SQL with Fine-Grained Guidance from Pivot Programming Languages cites this paper.

Pi-SQL: Enhancing Text-to-SQL with Fine-Grained Guidance from Pivot Programming Languages mT5: A massively multilingual pre-trained text-to-text transformer

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:58:38.971077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:58:38.971077Z digest=sha256:f0bea2f2530a46605d3e4c0aae8dab26918625ee43662130b8e4546fef0e4ffa

Observation d9529cdc-2f58-43b7-a861-d0f7558cd611 · inbound

Entity Image and Mixed-Modal Image Retrieval Datasets cites this paper.

Entity Image and Mixed-Modal Image Retrieval Datasets mT5: A massively multilingual pre-trained text-to-text transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:45.706545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:45.706545Z digest=sha256:e2b237972c65b997b173412032479e0f99fc2bd44e9a8fdb14af994c44d539f6

Observation 12e7cd48-852d-43fc-ba8f-9c48917f223e · inbound

OneSug: The Unified End-to-End Generative Framework for E-commerce Query Suggestion cites this paper.

OneSug: The Unified End-to-End Generative Framework for E-commerce Query Suggestion mT5: A massively multilingual pre-trained text-to-text transformer

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:52:06.504584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:52:06.504584Z digest=sha256:81a9b611f943d63fcce1f28558cafeae821a5df70c5f50c2b2e819a76ceb31e8

Observation 0465d49b-131e-44a2-b96b-3d19112081b2 · inbound

Cost-Optimal Active AI Model Evaluation cites this paper.

Cost-Optimal Active AI Model Evaluation mT5: A massively multilingual pre-trained text-to-text transformer

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:05.291570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:05.291570Z digest=sha256:e3b9a5bfb8a17586653246ac34f8622625c7113ea79a5687b2954319256bb020

Observation 4bdabd75-5050-4a3a-9d28-f3a3f996cbd8 · inbound

Token Constraint Decoding Improves Robustness on Question Answering for Large Language Models cites this paper.

Token Constraint Decoding Improves Robustness on Question Answering for Large Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:54:44.696499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:54:44.696499Z digest=sha256:90bacd03207a5e0d8f88a8f04a1f3a038b1207677f11ba27d5c8537b430ae599

Observation 87f50de8-15a2-4d19-a1e5-02f159400aad · inbound

Prompt, Translate, Fine-Tune, Re-Initialize, or Instruction-Tune? Adapting LLMs for In-Context Learning in Low-Resource Languages cites this paper.

Prompt, Translate, Fine-Tune, Re-Initialize, or Instruction-Tune? Adapting LLMs for In-Context Learning in Low-Resource Languages mT5: A massively multilingual pre-trained text-to-text transformer

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:49.204744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:14:49.204744Z digest=sha256:a2be9d1fbbc4a1816739f2393644fe3b2109477f9629beb9d8cbdeba1319e488

Observation 03770e90-0e79-4c6c-a2a2-1c2f0daa1e4d · inbound

FineWeb2: One Pipeline to Scale Them All -- Adapting Pre-Training Data Processing to Every Language cites this paper.

FineWeb2: One Pipeline to Scale Them All -- Adapting Pre-Training Data Processing to Every Language mT5: A massively multilingual pre-trained text-to-text transformer

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-06T22:46:42.897626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:46:42.897626Z digest=sha256:543cfd1aa94b8f1f824f4fad676ddad1d07c5d9734994cbc300e6e7e8dfccd46

Observation e45fca83-5c05-4c8a-84d6-664accfa8009 · inbound

Two Spelling Normalization Approaches Based on Large Language Models cites this paper.

Two Spelling Normalization Approaches Based on Large Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:50:02.372032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:50:02.372032Z digest=sha256:7a37d63c909003bbd37696ac882624787613b86be1c596e99e79490be87c08d4

Observation 4b3fa015-7e9d-4ad0-b630-9a2d3d18847a · inbound

Fine-Grained Chinese Hate Speech Understanding: Span-Level Resources, Coded Term Lexicon, and Enhanced Detection Frameworks cites this paper.

Fine-Grained Chinese Hate Speech Understanding: Span-Level Resources, Coded Term Lexicon, and Enhanced Detection Frameworks mT5: A massively multilingual pre-trained text-to-text transformer

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:29.933282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:16:29.933282Z digest=sha256:a67432056025261b71bf663cf34788d4ac348fb6d62e65501d4cac2220058e8b

Observation c7322498-8406-416d-bca1-338d1589208f · inbound

Meta CLIP 2: A Worldwide Scaling Recipe cites this paper.

Meta CLIP 2: A Worldwide Scaling Recipe mT5: A massively multilingual pre-trained text-to-text transformer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:23.768605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:23.768605Z digest=sha256:f2571261669b693a6a02be344c0d0031328a30d9cf004fe7ada9387363781ef2

Observation ad72b08f-cd06-409b-a456-7baf9d73d505 · inbound

Multi-TW: Benchmarking Multimodal Models on Traditional Chinese Question Answering in Taiwan cites this paper.

Multi-TW: Benchmarking Multimodal Models on Traditional Chinese Question Answering in Taiwan mT5: A massively multilingual pre-trained text-to-text transformer

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T05:49:51.616635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:49:51.616635Z digest=sha256:68fcfe76e877cb4a81df5e3b939414cd7875c285f67d13fa02bf3a37d9db7da1

Observation 28171424-a9f8-433f-8bf3-3907bfee3ab2 · inbound

SHAMI-MT: A Syrian Arabic Dialect to Modern Standard Arabic Bidirectional Machine Translation System cites this paper.

SHAMI-MT: A Syrian Arabic Dialect to Modern Standard Arabic Bidirectional Machine Translation System mT5: A massively multilingual pre-trained text-to-text transformer

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T05:07:15.890144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:07:15.890144Z digest=sha256:aaf93afd781ab1f5c262adc2aaeae2576bc01bbfda5118314a6d46f686c9e0ae

Observation 4f42dccd-0a73-4bbb-9ded-6bdd09c91234 · inbound

Multi-task Adversarial Attacks against Black-box Model with Few-shot Queries cites this paper.

Multi-task Adversarial Attacks against Black-box Model with Few-shot Queries mT5: A massively multilingual pre-trained text-to-text transformer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:34.518223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:34.518223Z digest=sha256:5c76ce36d53124d5b1106b35267c42885e563ec14470ad7fc69526f6f78ddead

Observation f76b0769-4721-48fc-8c51-8309f8639c82 · inbound

Evaluating LLMs on Chinese Idiom Translation cites this paper.

Evaluating LLMs on Chinese Idiom Translation mT5: A massively multilingual pre-trained text-to-text transformer

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:13.796060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:30:13.796060Z digest=sha256:3b231210e3c51636fa1f3bc2a439a730627104059ecd6b244c7bbcbffb856187

Observation b35d30f8-26c7-45d1-8471-6fe81b10b6dd · inbound

On the Fitness Landscape in the $NK$ Model cites this paper.

On the Fitness Landscape in the $NK$ Model mT5: A massively multilingual pre-trained text-to-text transformer

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:56.406142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:56.406142Z digest=sha256:4d94743ae2db528a34b5e4a33ab1bebcc2693e90f01be3c68a2c95bb0373c282

Observation 96b37668-2db2-4ca7-b74d-1866b18fa4d5 · inbound

Evaluating the Impact of Verbal Multiword Expressions on Machine Translation cites this paper.

Evaluating the Impact of Verbal Multiword Expressions on Machine Translation mT5: A massively multilingual pre-trained text-to-text transformer

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T20:51:50.889253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T20:48:48.443837Z digest=sha256:1bd032e68dfdbd450946315641c04570afbe30c531a0ec05b185eeccee7e049d

Observation e03f44c5-ce01-4b24-bed8-34cfb1afdc60 · inbound

MultimodalHugs: Enabling Sign Language Processing in Hugging Face cites this paper.

MultimodalHugs: Enabling Sign Language Processing in Hugging Face mT5: A massively multilingual pre-trained text-to-text transformer

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T20:35:49.861730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:35:49.861730Z digest=sha256:d7dd5e381c0fef1b1db3d0067228fc324add9be8b204446914e8348ae2a4c4fd

Observation e38c37ac-6877-4f39-b748-bcf09aeb0517 · inbound

Effective vocabulary expansion of multilingual language models for extremely low-resource languages cites this paper.

Effective vocabulary expansion of multilingual language models for extremely low-resource languages mT5: A massively multilingual pre-trained text-to-text transformer

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T02:57:46.473548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:57:46.473548Z digest=sha256:ee04f73b4781728cc1c1c12313e1d005aa75431daee75a7b665950401f5d0e91

Observation 0e035655-ebaa-4ec6-9f06-add5ca997a6d · inbound

LASQ: A Low-resource Aspect-based Sentiment Quadruple Extraction Dataset cites this paper.

LASQ: A Low-resource Aspect-based Sentiment Quadruple Extraction Dataset mT5: A massively multilingual pre-trained text-to-text transformer

Reference 68

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:40:58.967716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:34:01.296047Z digest=sha256:a43dfb01c933a772e63eaa125a13a326a27b18a104c83e341ed2dceb6aaf408e

Observation d5ccd539-58af-4a7f-9387-39acadcf1736 · inbound

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance cites this paper.

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance mT5: A massively multilingual pre-trained text-to-text transformer

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:16:09.179547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:02:54.374120Z digest=sha256:099fad381a9da6c503be34f5fc21f13f62e99c0d8091524468c4d279c7a81daa

Observation 03de33f2-8404-4675-986a-e5f5178d444c · inbound

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language cites this paper.

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language mT5: A massively multilingual pre-trained text-to-text transformer

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:16:05.908729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T02:06:18.801108Z digest=sha256:c1ee30335a4803ad364d5c5b95c5afa6c07fb50317d63666100a18cb5ada0c0d

Observation 36686b3e-b607-4161-9716-de6de0cad4fe · inbound

Enhancing ASR Performance in the Medical Domain for Dravidian Languages cites this paper.

Enhancing ASR Performance in the Medical Domain for Dravidian Languages mT5: A massively multilingual pre-trained text-to-text transformer

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:15:59.633309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:15:32.992695Z digest=sha256:1f73f94756006fbc58ff855319a4b7b69d041562678ac9d1257a1ce27ffa66f1

Observation 3a1c18ec-2bed-4157-8bce-2339e327250e · inbound

GLIER: Generative Legal Inference and Evidence Ranking for Legal Case Retrieval cites this paper.

GLIER: Generative Legal Inference and Evidence Ranking for Legal Case Retrieval mT5: A massively multilingual pre-trained text-to-text transformer

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:13.285738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T05:25:43.489962Z digest=sha256:6c9690c486af6e88da85d3875f48ed91d3b15eb55f85100123e9659b19ae664f

Observation 3931dae7-d4a8-438a-ab92-17f67bc853a1 · inbound

Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus cites this paper.

Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus mT5: A massively multilingual pre-trained text-to-text transformer

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:19.932378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T19:29:17.346870Z digest=sha256:1ad397747240d0ad8996ae41d5b453790a67b064c6481eac225e748468852616

Observation 1daeba2e-1445-444c-bd41-35f7e7f7fdae · inbound

Rethinking Data Curation in LLM Training: Online Reweighting Offers Better Generalization than Offline Methods cites this paper.

Rethinking Data Curation in LLM Training: Online Reweighting Offers Better Generalization than Offline Methods mT5: A massively multilingual pre-trained text-to-text transformer

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:11:53.364814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:09:21.652035Z digest=sha256:ef45fd901cab479f86758483b224502c495ef9e559c9c50d06da5a24883bd60e

Observation d5d0ce95-4540-4403-be98-43e05ff8160f · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages mT5: A massively multilingual pre-trained text-to-text transformer

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:38:21.844553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:6d3c56125438f6c2652eec646bfae7711f3058230ada97eed255ba80e334dc1a

Observation 48eefc56-0a8a-44eb-ae32-8e2562fa245a · inbound

CRAFT: Cost-aware Refinement And Front-aware Tuning of Prompts cites this paper.

CRAFT: Cost-aware Refinement And Front-aware Tuning of Prompts mT5: A massively multilingual pre-trained text-to-text transformer

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:46:46.415939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T06:39:17.268337Z digest=sha256:82188313217939f92241ea065fd2b0ae75df9495011fc56e75962a3c19a15d73

Observation f86f1e86-80ee-45e1-95e9-682e21463514 · inbound

Modular Monolingual Adaptation using Pretrained Language Models cites this paper.

Modular Monolingual Adaptation using Pretrained Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 63

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:36:59.100196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T01:11:42.032901Z digest=sha256:8cf942c4ad14a40957dc4d8f4e4aab50b5793f577e2faf5b17f954aa3f3de77a

Observation ae38762e-8508-4394-bf0c-1c6d050dbab7 · inbound

Data Synthesis and Parameter-Efficient Fine-Tuning for Low-Resource NMT: A Case Study on Q'eqchi' Mayan cites this paper.

Data Synthesis and Parameter-Efficient Fine-Tuning for Low-Resource NMT: A Case Study on Q'eqchi' Mayan mT5: A massively multilingual pre-trained text-to-text transformer

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:47:31.224479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T16:20:46.642433Z digest=sha256:7d70a3cc0f0718384660858d2b57985710fa13783f761c5a1f2a9d88cad43407

Observation 87dcf9ae-a28b-4aa8-b81f-9329e72cd4a5 · inbound

Phonemes to the Rescue: Multilingual Tokenization Based on International Phonetic Alphabet cites this paper.

Phonemes to the Rescue: Multilingual Tokenization Based on International Phonetic Alphabet mT5: A massively multilingual pre-trained text-to-text transformer

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:39:34.615803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T16:53:25.556222Z digest=sha256:8ce5119f953c43b34ef33f569c3ca1d2ece8609b74ea60196047a82e9535baff

Observation 7221deb0-67ad-4415-ac92-0ec541a09599 · inbound

Rethinking Indic AI from a Lens of Cultural Heritage Preservation cites this paper.

Rethinking Indic AI from a Lens of Cultural Heritage Preservation mT5: A massively multilingual pre-trained text-to-text transformer

Reference 240

Resolution
verified exact
local_arxiv, observed 2026-07-08T02:24:27.919647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-08T02:22:31.134521Z digest=sha256:34fcfec7e38a4cf9f4a402c240e6fe42c20b576952beb685b5cce0f80706669b

Observation 4f9568b7-67dd-4ed2-9150-9a81991d3ef0 · inbound

DIRECT: Direct Decoding for Efficient and Aligned Sequence Labeling with Large Language Models cites this paper.

DIRECT: Direct Decoding for Efficient and Aligned Sequence Labeling with Large Language Models mT5: A massively multilingual pre-trained text-to-text transformer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-30T17:59:32.553795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T17:59:32.553795Z digest=sha256:64d1521a377fe2d66949f04a522693b5d8617e3886307ad57e4c3fb6aa724d65