Pith. sign in

Paper Citation Record · LEDGER

High-Fidelity Audio Compression with Improved RVQGAN

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2306.06546.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.06546 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:14:45.767421Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

24
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e270137-8315-4fb3-9ec5-c3805b530c73 · inbound

AI-Generated Song Detection via Lyrics Transcripts cites this paper.

AI-Generated Song Detection via Lyrics Transcripts High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:21:30.071947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:21:30.071947Z digest=sha256:8bdbc17ca9d5818a7d1141d242bebd7f50943cabb26b535708483662ac8df272

Observation 72e08b84-24df-4c1c-8c9d-fa09d208c4c9 · inbound

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine cites this paper.

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine High-Fidelity Audio Compression with Improved RVQGAN

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:34.455589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:47:34.455589Z digest=sha256:af6b61e25641c63acb0e898b0336dcacd1267c50fb011c81c2517426bb839b59

Observation 2798e571-ad33-4e5a-ada7-a1d5ba7f446d · inbound

Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning cites this paper.

Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning High-Fidelity Audio Compression with Improved RVQGAN

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:14:46.240127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:14:46.240127Z digest=sha256:512e90c2ffbc1a29add7b5694015b06fb1f1a6add28d065e717745a50ff9d8b2

Observation 0ea89150-798c-40b1-9ff6-63b4d87fc4ef · inbound

Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission cites this paper.

Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission High-Fidelity Audio Compression with Improved RVQGAN

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T11:29:01.971972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:29:01.971972Z digest=sha256:a912080e79c187fc80fb58b0ea4332fedc8bbec747509e9a98633d8eab3435ac

Observation bc504376-b5c5-4872-aca3-1de37ec2c946 · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding High-Fidelity Audio Compression with Improved RVQGAN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.163714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:a0021b84643aa719872b749dc4da1033e57f54be625102b8248b032dd9aec294

Observation c97c7d44-f0d5-45ca-8fd4-c117f5a372e6 · inbound

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns cites this paper.

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns High-Fidelity Audio Compression with Improved RVQGAN

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-14T22:43:13.611383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:43:13.611383Z digest=sha256:0302b10653c44ffc164ac578ff31419a42f4fec1844b75ee57c0dd05eb972fa8

Observation 8da1a194-4f45-40a8-91f2-6cc16c018a92 · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model High-Fidelity Audio Compression with Improved RVQGAN

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.792883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:930bfdd81c9c8fab1f803d72826006043945bca09974f80dd7ced9194fb61bd2

Observation 2e20c943-ac0c-42ee-bc6b-e527083ac1ff · inbound

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs cites this paper.

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.398296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:13:47.971429Z digest=sha256:74e08084212bea7f9985bec4e1b76cce1e8e86171df2b042981cecbef82611bb

Observation d476b8f2-8481-4bbe-8e4b-97e984948d92 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:44:00.668460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T06:43:52.735211Z digest=sha256:4c547396c73f4e06a8935d8a997d8fa59e1e7f276aad51b23b4a8043ee021055

Observation 5f42bc6e-c314-4a57-9a10-33625285b2f3 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:23.802002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T05:44:46.831360Z digest=sha256:94298115cb7847660cd4b0a10916edad0b943fe95bf294e1cc80bf5fd3b6b967

Observation 672edff9-c431-48c5-a3ec-704f628101a4 · inbound

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models cites this paper.

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:36.067248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T14:58:27.176375Z digest=sha256:7e0afd8e1ff83d0937d053c89a56eccd848d58e63ef2bda0a821fd337a11b40e

Observation ad2efb5b-5feb-4db4-b3bc-8923f82e9f78 · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents High-Fidelity Audio Compression with Improved RVQGAN

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.521886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:adeeb11fc7e129bf6830552c03d35bbc3fe986c102ed003c9fa5bb46fb43641b

Observation c0923909-5374-4f44-94ef-df9eedf159f7 · inbound

NAC: Neural Action Codec for Vision-Language-Action Models cites this paper.

NAC: Neural Action Codec for Vision-Language-Action Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:59:37.904034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T13:59:53.484305Z digest=sha256:acf596f7fa9f7ac03078c3ccb7ec3266375babf5af3e97e05d83354839e831aa

Observation 5d2969a1-f330-4ccb-8171-eb5121f4d6cd · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.247716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:ed1fc09540c65372fdb62ed214a7800ae6ea1276174a68fe4715f3b8db63a98f

Observation b520bed0-bfc9-4c40-a6f7-2a2ee91cd0bf · inbound

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts cites this paper.

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts High-Fidelity Audio Compression with Improved RVQGAN

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:30.515762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:23:30.515762Z digest=sha256:bae096dcaf47c17794a2ebca0f44f1453fcba351bfe7a65fb311a5dadb834d93

Observation 753f55db-093b-4210-a661-34c9eb1e7920 · inbound

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness cites this paper.

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness High-Fidelity Audio Compression with Improved RVQGAN

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T08:25:28.089701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:25:28.089701Z digest=sha256:ca16bdd8372f7f8de57d750065282d115b4bb05fcda430eb15d7b6a17a308528

Observation ffac6167-9e67-498e-ba85-c20864d82cf3 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T10:35:00.630413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T10:35:00.630413Z digest=sha256:f98c4aef74c20595f5e358110bec0bd1b873a1337e2b113260771993e49e62a0

Observation b693b372-f65d-4222-9903-f5161639ba97 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T01:50:59.050808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:50:59.050808Z digest=sha256:478e178975c967a09e64f55382019a404e348a0e39e09c352afc4a93637e445a

Observation 6998a798-10b3-4f4b-b2bb-868dab788b76 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks High-Fidelity Audio Compression with Improved RVQGAN

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T16:29:24.694740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:29:24.694740Z digest=sha256:152d9dd995e28b06d390b925018d192d855ba6355f0403d240e610d9334f728d

Observation 59066311-494b-4bd7-b822-fd5f4f5b1025 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks High-Fidelity Audio Compression with Improved RVQGAN

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:45.767421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:45.767421Z digest=sha256:7f077769e3ecab6ec9a741caca4e19e644ad90f3f5d3225bfee6b1618bd18807

Observation 613f7c29-7e12-4080-b576-24f4866498b2 · inbound

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs cites this paper.

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T14:08:33.591322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:08:33.591322Z digest=sha256:6946c8a4ad1112217322c27f74a68ae926849bfcc5405f2940d081b3754d54e8