Pith. sign in

Paper Citation Record · LEDGER

vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:1910.05453.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1910.05453 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:42:56.928928Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:27:34.881381Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3d1b2d55-e44f-4080-9d8a-81bdb4d79a59 · inbound

Multimodal Medical Code Tokenizer cites this paper.

Multimodal Medical Code Tokenizer vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T00:42:56.928928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:42:56.928928Z digest=sha256:e4b41207418ead2bad8ab25fbbd89fcb77b4c79420a554bda7b75da803d225b4

Observation f358c909-a9a3-40c7-8130-9d129917742f · inbound

Bitrate-Controlled Diffusion for Disentangling Motion and Content in Video cites this paper.

Bitrate-Controlled Diffusion for Disentangling Motion and Content in Video vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:44:59.929335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:44:59.929335Z digest=sha256:733694173274904bb57e5b0f0b0d420795f1cb59a82657e7c7de4955a25b69a1

Observation 25d505d1-4f49-424e-9ff5-4b9809c63378 · inbound

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs cites this paper.

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:01:24.421342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T12:57:04.450462Z digest=sha256:b68aa5bd9ba31f3d95c276a14315764da867728f969f6fbff9415110e578918d

Observation cf743998-8796-4234-9694-97d682b945bc · inbound

Incomplete Multi-View Multi-Label Classification via Shared Codebook and Fused-Teacher Self-Distillation cites this paper.

Incomplete Multi-View Multi-Label Classification via Shared Codebook and Fused-Teacher Self-Distillation vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T16:53:00.174441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T16:50:19.940458Z digest=sha256:a55e7d920dc8016804d8653e33e7c6c532e1a6a0c9c6ddd1a714e41f40e1f2f4

Observation 4857b5f0-a324-44d1-9439-9d915dabba4f · inbound

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization cites this paper.

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.887715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T12:25:52.847432Z digest=sha256:227fc665e68e3c3f7aba56030fdb7f3aa756485633b7367fe32e3e3fb3cd3fd8

Observation a62efb3b-79da-4838-af2e-43f700c455bc · inbound

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization cites this paper.

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:15:07.879196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T23:14:32.494076Z digest=sha256:cf5bfc06e319e453df0bd5dd7463d8e1fac2b2800f8b3b899df4b53bc5d2b9a4

Observation 2082989e-05b8-4318-81fd-1a4cd4ee0a7a · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 235

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.749867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:ffe55e24957894e13283f0147f64699674965de3eccfaf2c6fa717ae50aceb6d

Observation 76f3b6db-b6af-4849-8009-e46213ca13ca · inbound

How Optimality Structures Sparse Dictionaries: A Theory for Understanding SAE Representations cites this paper.

How Optimality Structures Sparse Dictionaries: A Theory for Understanding SAE Representations vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 272

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:46:26.206836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T11:36:45.020411Z digest=sha256:abd5bf32fa279baac367ad27357c5e9cd520e07fadeb3162437c673d23b17d9e

Observation db3eab54-383c-4847-9ac9-28123ace8316 · inbound

End-to-End Training for Discrete Token LLM based TTS System cites this paper.

End-to-End Training for Discrete Token LLM based TTS System vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T03:27:34.882863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T15:22:06.893507Z digest=sha256:163c43ec02d337642d1f13eb14be363a2923aa25f5363e4d00351167c73e3401