Pith. sign in

Paper Citation Record · LEDGER

Unsupervised Cross-lingual Representation Learning for Speech Recognition

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2006.13979.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2006.13979 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:27:19.302520Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:09:41.969238Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 53867186-9827-4bb7-a66c-f5990b2fb264 · inbound

Self-Supervised Convolutional Audio Models are Flexible Acoustic Feature Learners: A Domain Specificity and Transfer-Learning Study cites this paper.

Self-Supervised Convolutional Audio Models are Flexible Acoustic Feature Learners: A Domain Specificity and Transfer-Learning Study Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T12:27:19.302520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:27:19.302520Z digest=sha256:ce80510b25d30d3d549f9fe896fd3125360b9f632ed97ef6d4178a3923774f73

Observation 31249799-37a2-4892-b4d0-c234593b4b53 · inbound

GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling cites this paper.

GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T10:37:03.570902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:37:03.570902Z digest=sha256:8df42b75f6b0da106ad1a71e0d5dea2e4e98c6b0902d761c9c371502e347eda5

Observation 0e429902-49c8-44c6-97b5-1fe27a9c0ce7 · inbound

ZIPA: A family of efficient models for multilingual phone recognition cites this paper.

ZIPA: A family of efficient models for multilingual phone recognition Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T12:55:40.024238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:55:40.024238Z digest=sha256:ec02541f2c368897c03868506e3e286d69d957b9e2ea5d615212300507c88b34

Observation a85d6171-4281-455f-9264-4deca39ca57f · inbound

SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset cites this paper.

SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:37:59.266363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:37:59.266363Z digest=sha256:8131f8f88d478f52a5591c479d512825155456f10d45d4b7f19e551b27727d50

Observation 98ab8b0e-39c6-44e4-9cf5-c318e8c5a46f · inbound

Vedavani: A Benchmark Corpus for ASR on Vedic Sanskrit Poetry cites this paper.

Vedavani: A Benchmark Corpus for ASR on Vedic Sanskrit Poetry Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T12:14:42.285623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:14:42.285623Z digest=sha256:94dc75f74657097a97d73101125385106fdc5499433628bfd541ec14745e1f78

Observation e52cc9c2-3249-4c6f-96fe-1ffed8a1af08 · inbound

UD-KSL Treebank v1.3: A semi-automated framework for aligning XPOS-extracted units with UPOS tags cites this paper.

UD-KSL Treebank v1.3: A semi-automated framework for aligning XPOS-extracted units with UPOS tags Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:46.820340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:01:46.820340Z digest=sha256:3445199d81ed3465b4dfb1bf6b3eccff3ca2c3a828f7d9c27f85cf9e6c2911ea

Observation f1b31344-c420-416c-9a96-ebf90e51f097 · inbound

Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning cites this paper.

Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:49.820015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:34:49.820015Z digest=sha256:fe55c1e27a8ef77caa79058b69b63ff13dbf12af2511cd0d0bff62c957eeb12d

Observation 88cb9a4f-b88e-4145-8c66-caeb5886c12d · inbound

Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection cites this paper.

Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.129554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.129554Z digest=sha256:e83294ce87095b554c63638f71859a0feafc231446962ff23ad912c9c83d1012

Observation c0d15c12-3187-4410-894e-48984a8915a1 · inbound

Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models cites this paper.

Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:39:22.322031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:39:22.322031Z digest=sha256:a550c5478d51ec85d0a189d63382466657519de3a67c3358e73123e19a435b83

Observation 8830ce8b-09bc-4675-af92-97d446d65ef9 · inbound

TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation cites this paper.

TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:27:29.308951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:27:29.308951Z digest=sha256:3dd3195b1a0455e85ff090767140331350e9bf5bf07f23035fb42e2845b446eb

Observation 41b657f4-015d-434b-b217-61c7c4443fe8 · inbound

Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems cites this paper.

Bona fide Cross Testing Reveals Weak Spot in Audio Deepfake Detection Systems Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T19:37:20.032674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:37:20.032674Z digest=sha256:413349d4d8e8b95e1840dcf5924d303bd5bdf365f0b770dbd88b3f899f170c82

Observation 28d5c376-de19-43ae-80aa-0fbd64522cbc · inbound

Prominence-aware automatic speech recognition for conversational speech cites this paper.

Prominence-aware automatic speech recognition for conversational speech Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T18:11:15.997955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:11:15.997955Z digest=sha256:2263ef8c594422adb6be7f7a0bc2f2a655f1b9cb6856548e3cf1ff61f2532097

Observation 721520c4-c719-475d-9685-3c36740ee176 · inbound

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models cites this paper.

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:29:31.132461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:29:31.132461Z digest=sha256:f0e39615e6716f01e5f778038387aa3e2271d854550ee9a2d00dce408310f1a5

Observation 37726469-f05b-4102-ad28-7c08f9ba63ff · inbound

FAC-FACodec: Controllable Zero-Shot Foreign Accent Conversion with Factorized Speech Codec cites this paper.

FAC-FACodec: Controllable Zero-Shot Foreign Accent Conversion with Factorized Speech Codec Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:01:06.824442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T07:56:46.302918Z digest=sha256:f1835bf0d1a39e80c7828a16d39981dc0d83175be23c6d0a2dd61bf60770a913

Observation 56f4377f-3a06-45be-ac84-90a15a47fd14 · inbound

ENEC: A Lossless AI Model Compression Method Enabling Fast Inference on Ascend NPUs cites this paper.

ENEC: A Lossless AI Model Compression Method Enabling Fast Inference on Ascend NPUs Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:58:03.011495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:56:38.684104Z digest=sha256:cff50ae67538f415b65df4444758de6c9a15820d16d10579681182f52faa7a0d

Observation ea73a65e-fc5b-49f6-9e0d-fcaea6b83d8e · inbound

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy cites this paper.

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:47:24.686187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T19:11:40.161390Z digest=sha256:c9c87695fe3c9e65eb3589f8691c4387bfe1ad5544aac26b4cbebc96d07d5792

Observation 3d33b2a8-1c9f-4932-af46-b461b42c97c4 · inbound

A study on the impact of region specific data on the performance of Indic ASR cites this paper.

A study on the impact of region specific data on the performance of Indic ASR Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:35.604071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T15:06:19.198758Z digest=sha256:3f02b268e7c229fecf6c91616d72a80d4bdc98e33189203443dbce04a38596ad

Observation 8c183895-9eba-4a4d-a3e4-56483739f3b4 · inbound

Overcoming Decoder Inconsistencies in Whisper for Dravidian and Low-Resource Languages cites this paper.

Overcoming Decoder Inconsistencies in Whisper for Dravidian and Low-Resource Languages Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:47:31.183833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T16:21:39.960715Z digest=sha256:848a26ae9f96d4facab138dc11e68ed8ff21ce17878bb0bd9903fe91b52d4e67

Observation 1ea0e690-d1bd-42d6-a64e-af5e414ccedf · inbound

Pretrained self-supervised speech models can recognize unseen consonants cites this paper.

Pretrained self-supervised speech models can recognize unseen consonants Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:07:56.057986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:14:47.932613Z digest=sha256:3c7de02df8a51e144a20194f7677e58dc28702decc55fa6dc6231a99b8998e72

Observation decfd35a-bd0a-432d-9585-356787c27d38 · inbound

Adding Robust Code-Switching Capabilities to High Performance Multilingual ASR cites this paper.

Adding Robust Code-Switching Capabilities to High Performance Multilingual ASR Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:41.971092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:01:57.319491Z digest=sha256:0ff1ef40f5def929f146f7277a4b1aa8485fea9cb1e5442623ca4cdf6342ba1f

Observation a6500f5d-adf0-422f-b698-46425134df1c · inbound

Which Languages Transfer Best to Warlpiri? A Similarity-Based Study for Low-Resource ASR cites this paper.

Which Languages Transfer Best to Warlpiri? A Similarity-Based Study for Low-Resource ASR Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-14T13:08:11.026186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T13:08:11.026186Z digest=sha256:156bad261f2b448655e3dbae2bc7687fb0fd1e583e9ff0e01442485473bca333

Observation c0e44047-2705-449b-9c3c-749894f32257 · inbound

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection cites this paper.

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection Unsupervised Cross-lingual Representation Learning for Speech Recognition

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T02:09:49.283745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:09:49.283745Z digest=sha256:1aaf255b3c0f87cc28550cf2f31d019e1b3948355defad3dd78a3449caa3794a