Pith. sign in

Paper Citation Record · LEDGER

High-Fidelity Audio Compression with Improved RVQGAN

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2306.06546.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.06546 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:14:45.767421Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

24
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e270137-8315-4fb3-9ec5-c3805b530c73 · inbound

AI-Generated Song Detection via Lyrics Transcripts cites this paper.

AI-Generated Song Detection via Lyrics Transcripts High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:21:30.071947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:21:30.071947Z digest=sha256:ec7856f90e6b29f746f253ad700a7257f87227107e014c331de0116d75ad4706

Observation 72e08b84-24df-4c1c-8c9d-fa09d208c4c9 · inbound

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine cites this paper.

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine High-Fidelity Audio Compression with Improved RVQGAN

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:34.455589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:47:34.455589Z digest=sha256:af6b61e25641c63acb0e898b0336dcacd1267c50fb011c81c2517426bb839b59

Observation 2798e571-ad33-4e5a-ada7-a1d5ba7f446d · inbound

Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning cites this paper.

Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning High-Fidelity Audio Compression with Improved RVQGAN

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:14:46.240127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:14:46.240127Z digest=sha256:512e90c2ffbc1a29add7b5694015b06fb1f1a6add28d065e717745a50ff9d8b2

Observation 0ea89150-798c-40b1-9ff6-63b4d87fc4ef · inbound

Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission cites this paper.

Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission High-Fidelity Audio Compression with Improved RVQGAN

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T11:29:01.971972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:29:01.971972Z digest=sha256:f26fcc4c5c08f42eb526279550e6f9df63ca012803e97caa09f416f8ec595e44

Observation bc504376-b5c5-4872-aca3-1de37ec2c946 · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding High-Fidelity Audio Compression with Improved RVQGAN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.163714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:be71d3022f5049b56e5d5a12ac9a21a7f7d38a0306a3decdcb3b8c076321b472

Observation c97c7d44-f0d5-45ca-8fd4-c117f5a372e6 · inbound

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns cites this paper.

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns High-Fidelity Audio Compression with Improved RVQGAN

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-14T22:43:13.611383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:43:13.611383Z digest=sha256:0302b10653c44ffc164ac578ff31419a42f4fec1844b75ee57c0dd05eb972fa8

Observation 8da1a194-4f45-40a8-91f2-6cc16c018a92 · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model High-Fidelity Audio Compression with Improved RVQGAN

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.792883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:494fe13707e36cc5fd3b763905e9dd8447533cdeba10063756e437a481ae658e

Observation 2e20c943-ac0c-42ee-bc6b-e527083ac1ff · inbound

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs cites this paper.

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.398296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:13:47.971429Z digest=sha256:3e7d8fd0680e8fb60f191b29dd4b59ae20b9e279d76fd36fb3c37f53ceddb0a0

Observation d476b8f2-8481-4bbe-8e4b-97e984948d92 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:44:00.668460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T06:43:52.735211Z digest=sha256:fdab1b63bed24b4f954c68351de40cff6c888ab3edf3accfb0b7572dcbe61378

Observation 5f42bc6e-c314-4a57-9a10-33625285b2f3 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:23.802002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T05:44:46.831360Z digest=sha256:ba7e6dd42978834254f8c4fea1e40f1a73dcecf09022acf5e5f54f26c8be4d03

Observation 672edff9-c431-48c5-a3ec-704f628101a4 · inbound

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models cites this paper.

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:36.067248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T14:58:27.176375Z digest=sha256:fe1fb05d832b35620c808dbec6639703c58d41d576abc83ec9f63a421dbdcd91

Observation ad2efb5b-5feb-4db4-b3bc-8923f82e9f78 · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents High-Fidelity Audio Compression with Improved RVQGAN

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.521886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:ea7b00940b64a17dbd756bc3f9ff03e96eb44722eaa38d34040aab62265db972

Observation c0923909-5374-4f44-94ef-df9eedf159f7 · inbound

NAC: Neural Action Codec for Vision-Language-Action Models cites this paper.

NAC: Neural Action Codec for Vision-Language-Action Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:59:37.904034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T13:59:53.484305Z digest=sha256:0d5245dd5a90cede3758c0c06371153447ff913de12204aac4e421d4a8c56244

Observation 5d2969a1-f330-4ccb-8171-eb5121f4d6cd · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.247716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:666e1e1f272b52e73c4753bdb234e58159c08ab3c89f6109cf4d3510ad18db4d

Observation b520bed0-bfc9-4c40-a6f7-2a2ee91cd0bf · inbound

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts cites this paper.

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts High-Fidelity Audio Compression with Improved RVQGAN

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:30.515762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:23:30.515762Z digest=sha256:bae096dcaf47c17794a2ebca0f44f1453fcba351bfe7a65fb311a5dadb834d93

Observation 753f55db-093b-4210-a661-34c9eb1e7920 · inbound

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness cites this paper.

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness High-Fidelity Audio Compression with Improved RVQGAN

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T08:25:28.089701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:25:28.089701Z digest=sha256:ca16bdd8372f7f8de57d750065282d115b4bb05fcda430eb15d7b6a17a308528

Observation ffac6167-9e67-498e-ba85-c20864d82cf3 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T10:35:00.630413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T10:35:00.630413Z digest=sha256:f98c4aef74c20595f5e358110bec0bd1b873a1337e2b113260771993e49e62a0

Observation b693b372-f65d-4222-9903-f5161639ba97 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T01:50:59.050808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:50:59.050808Z digest=sha256:478e178975c967a09e64f55382019a404e348a0e39e09c352afc4a93637e445a

Observation 6998a798-10b3-4f4b-b2bb-868dab788b76 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks High-Fidelity Audio Compression with Improved RVQGAN

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T16:29:24.694740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:29:24.694740Z digest=sha256:6eafda0b9748b664915f56f6784da4242641f2d44172530f491150b4c525aace

Observation 59066311-494b-4bd7-b822-fd5f4f5b1025 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks High-Fidelity Audio Compression with Improved RVQGAN

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:45.767421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:45.767421Z digest=sha256:a38facc72bd2cd044332b4d286b8b781ad8107e01534cb65f2fe863237ac1450

Observation 613f7c29-7e12-4080-b576-24f4866498b2 · inbound

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs cites this paper.

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T14:08:33.591322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:08:33.591322Z digest=sha256:16ec2b9301d39a743c37e8d3fc2c869d49e3119c3c4b8242f9d35247931f9aa2