Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:00:02.820379Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2412.11449.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:00:02.820379Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:00:02.477825Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T11:29:04.296162Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 362bc031-fc62-45fd-8ce4-a426307fae6b · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bc0db306-bc06-4066-b8ba-52afd7addff8 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music For the case of speech, we use the LibriSpeech TTS dataset [27]
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2e5c4b4d-1b49-4c65-9fac-592d68c1abba · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7729d3d9-6cc1-4807-8946-76d1527f2000 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music We are only interested in quantifying the likelihood scores for the gener- ative architecture pre-training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 82ae8649-7b3c-48fc-9f31-ec5132a82eae · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music The proposed architecture outperforms a purely token-based model by combining continuous audio and dis- crete token-based representations
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e785a6ed-4fb8-4a30-a4ae-af7a8410fb0d · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music We thank both Google and Stanford HAI for this initiative
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f5cda316-aa4f-4aa5-852c-433e36aeb251 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Audiolm: a language modeling approach to audio generation,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f665c2dd-1a20-43f4-8462-a91236f5a276 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music VideoGPT: Video Generation using VQ-VAE and Transformers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7658836d-6cc0-4610-a070-d02b1bfd9d85 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Attention is all you need,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 63d715bd-9b42-419d-b9a0-c1defdcb6030 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Language Models are Few-Shot Learners
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6b9e469-c7c9-4de3-8093-7581663fd5c5 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41a5132d-32c1-4dce-aa0c-dbcdc6375438 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music A generative model for raw audio using transformer architectures,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1948d389-e4ef-4ffa-b2d3-bdf9fdae0a7d · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music A Language Model With Million Context Length For Raw Audio
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 852147a6-0030-4b57-934d-0f67c8d6511e · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Music transformer: Generating mu- sic with long-term structure,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9cea3149-d154-4584-a557-582ba3c8dfde · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music A Framework for Generative and Contrastive Learning of Audio Representations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d2bfc67e-447f-46ff-a2fb-684b61a9c752 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music WaveNet: A Generative Model for Raw Audio
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f88346-8bc6-4d2c-94ee-27060dc810a6 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Audio Transformers
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8384468e-a393-4ee6-aca8-843f12b7a0ba · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Psla: Improving audio tagging with pretraining, sampling, labeling, and aggregation,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bc88afda-ab7d-4059-81e1-54b6faafef48 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 94d345ac-4436-4f86-95d8-d32f6e75e212 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Gemini: A Family of Highly Capable Multimodal Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94c9d3b6-9c77-4c60-988c-e384a9c0773a · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8544ca89-6801-4570-89f0-1bfd1474ca8c · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Neural Discrete Representation Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d1a1d4-5e7f-4036-af3a-d60c1ad548f5 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8aae6c97-b2d8-4cbd-a690-ec6025554102 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Jukebox: A Generative Model for Music
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef4894c0-7435-482a-88f3-22bf0f6ec3cc · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Soundstream: An end-to-end neural audio codec,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a95da40-1dc8-4f35-be2e-7e394ededdab · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music High Fidelity Neural Audio Compression
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5319be5-4530-4a66-ad75-9e56b1497feb · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music On generative spoken language model- ing from raw audio,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 86f30933-fecc-4be8-b275-a4d90c83ee36 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music textless-lib: a Library for Textless Spoken Language Processing
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d436e966-8d7c-4ba0-a6c8-6cc1af5cefea · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music MusicLM: Generating Music From Text
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29b3b38f-fa64-4225-bc32-ffba9023eec6 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Simple and controllable music generation,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e5f36c6a-aafe-4c43-b936-92cdbf36fdc6 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a96e566-92fa-4a33-85b3-7a1633467b05 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music A deep learning approach for low-latency packet loss concealment of audio sig- nals in networked music performance applications,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f46b34ee-6b5c-4c4e-a38b-7c4fc7a83079 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Multi-Format Contrastive Learning of Audio Representations
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5095b69-910a-4388-afa3-2496334042d6 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Robust speech recognition via large-scale weak supervision,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa684bb-75b4-4b82-becf-3dc6af2a1aeb · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86d3b56c-09a3-419a-9982-f206de098694 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Distilling the Knowledge in a Neural Network
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abbe882d-ecfd-4aac-b51a-9a87a5514d40 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music MiniLLM: Knowledge distillation of large language models,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 109c07dd-92fa-4297-a3bd-cefecc9a5f05 · outbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music A simple and effective pruning approach for large lan- guage models,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b6b9e469-c7c9-4de3-8093-7581663fd5c5 · inbound
Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ec0687-a5c1-4e87-a590-60a3940f4dff · inbound
DebateBench: A Challenging Long Context Reasoning Benchmark For Large Language Models Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0262db34-b7b9-408d-b110-e776b3ad75c3 · inbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.