Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:54:58.755219Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 1 inbound Pith citation observation for arXiv:2508.11598.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:54:58.755219Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:54:58.525812Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T19:54:59.403728Z
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c87c66bf-5d6a-4851-91b1-d73aff9c5ccf · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Representing Speech Through Autoregressive Prediction of Cochlear Tokens
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e70517cc-c426-43e4-b519-7ec9eba9316d · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens water” and “river
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1dbc760f-0472-4349-bb5c-34037b0f7723 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens er” was often confused with “r
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8418ae8c-cc49-47e6-8870-8f8d089400a8 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Transformation Imitation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87b8727c-cc16-4faa-a61c-702bb755647c · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens acknowledges support from The K
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d35525ca-2a30-43d0-9329-ae571f573311 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens SUPERB: Speech processing Universal PERformance Benchmark
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7257bd88-eb8c-483a-84c1-1faded21bf78 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba6f1da7-8ec5-4af4-a722-1c40c19dfc80 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens A Review of Deep Learning Techniques for Speech Processing
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86bff55e-0e93-4447-aa6c-b3938d7f5878 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Soundstream: An end-to-end neural audio codec,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 587e6140-14f7-4998-a625-0c5aa2904316 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens High Fidelity Neural Audio Compression
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f225f353-cfd7-4e63-b6bc-70213e002bb5 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens HiFi-Codec: Group-residual Vector quantization for High Fidelity Audio Codec
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4075519f-d176-4b8a-bce3-882d355bb389 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens High-fidelity audio compression with improved rvqgan,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b43b3d0a-1d76-4908-9c67-cf1874151514 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Language-Codec: Bridging Discrete Codec Representations and Speech Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f67f68b-b935-425c-86b6-f24317671d3e · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Speechtok- enizer: Unified speech tokenizer for speech language models,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 61224e10-6883-4e89-a553-dd944a8b051c · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1886f8f4-88fc-4818-9a5b-9a573165cfbb · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f87b3e5-0f62-484c-a2bf-6ee9e158b49b · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7744f64-4ab0-4588-a3cc-55690244c93e · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Hierarchical organization of hu- man auditory cortex: evidence from acoustic invariance in the re- sponse to intelligible speech,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d4cc089-2728-4852-b9a9-0427f40f18a8 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Intonational speech prosody encoding in the human auditory cortex,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8464c6ed-1479-49e3-b937-6b7d393d173a · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Hubert: Self-supervised speech rep- resentation learning by masked prediction of hidden units,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9516fb69-478a-4798-bfa4-95e9d6cda6bf · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43f76c7d-ebb6-4dc9-9aa6-e9fbefa6023f · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Wavlm: Large-scale self- supervised pre-training for full stack speech processing,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed71dc35-6842-4fef-9545-5fdea44e45e5 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Blind phoneme segmentation with temporal prediction errors
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b177be0e-b3b8-4219-b791-7006268441cc · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens An Unsupervised Autoregressive Model for Speech Representation Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d39a96-2ea6-41e1-881d-0a986ab7417b · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Acquiring language from speech by learning to remember and predict,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 04458adf-6e74-495f-ae62-26106e634722 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Vector-Quantized Autoregressive Predictive Coding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4125ea9c-5623-4cdf-97da-333987beef75 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Audio albert: A lite bert for self-supervised learning of audio representation,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6774d475-e580-4647-9742-715858ad99a6 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens On generative spoken language modeling from raw au- dio,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c35d794d-6ba2-43a0-a662-565838525c7c · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Audiolm: a language modeling approach to audio gener- ation,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a55b06c-d015-495f-a306-39f267076dbf · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Improving Textless Spoken Language Understanding with Discrete Units as Intermediate Target
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4eb29282-33de-4d8d-b2ac-40aa568d4df9 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens BERT: Pre- training of Deep Bidirectional Transformers for Language Under- standing,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 54f7547d-408e-47fd-b58f-3c2612a5e04e · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Representation Learning with Contrastive Predictive Coding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05670626-1df8-47d1-a877-7bccc75d583d · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Wav2vec 2.0: a framework for self-supervised learning of speech representa- tions,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0ea69db-a483-47df-a0f9-960dc3ab132b · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Contrastive learning of general-purpose audio representations,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c036ce17-2c66-4ac3-a5f0-bec0c25cf112 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Unsupervised speech segmentation and variable rate repre- sentation learning using segmental contrastive predictive coding,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d8f38926-fe18-4bef-b40f-325195474c9d · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Self-normalization and noise- robustness in early auditory representations,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5178f83-656d-412d-a88b-be14c90d6f35 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 764197d2-608a-47ea-a155-1fae0d4a3019 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Multiresolution spectrotem- poral analysis of complex sounds,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7dfd53ae-2f9a-4016-9a9a-a5fa14fbd665 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Model metamers reveal divergent invariances between biological and ar- tificial neural networks,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a985c03-b404-43b8-903e-93fc23763d3c · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Derivation of auditory filter shapes from notched-noise data,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71303311-fc97-4f31-aa77-67021373af4d · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Sound texture perception via statistics of the auditory periphery: evidence from sound syn- thesis,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0098db9-3dc5-4551-b07b-e52f6010aaf4 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b59ce30-34e4-413c-8261-07bcdeafbaef · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Lib- rispeech: an asr corpus based on public domain audio books,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 653e79ba-084b-461e-be72-dbbf8515a5e7 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Improving language understanding by generative pre-training,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5858c51-8ec3-42f6-899c-28e3d76aa871 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Gaussian Error Linear Units (GELUs)
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e70758c-5b26-475c-ab84-35265a244a88 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Root Mean Square Layer Normalization
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d165a1c-5b8b-4e06-a373-0f6b2f87e3a8 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Libri-light: A benchmark for asr with limited or no supervision,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e2c0c792-4af0-45a9-a891-87aa239c8340 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Timit acoustic phonetic continuous speech cor- pus,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 98d1da4f-be71-4a68-852d-3bc94bbb395d · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Speaker-independent phone recogni- tion using hidden markov models,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eb7042ec-1ac7-468a-b583-3db817e346b2 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Scikit-learn: Machine Learning in Python,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2d0db29-082a-4ab4-8bf5-da843b7777f7 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens The Zero Resource Speech Benchmark 2021: Metrics and baselines for unsupervised spoken language modeling
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38b04a89-0a8f-4ecd-b6ce-a817b2096755 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Over-reliance on english hinders cognitive science,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 099f2fd0-5e9a-45d3-8cef-f1ff9189d8b0 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens To- wards inclusive automatic speech recognition,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3aaeb462-9d71-4e5c-ba65-82b91e492f2a · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Say- cam: A large, longitudinal audiovisual dataset recorded from the infant’s perspective,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 071b6908-1925-49bf-b58e-757b99b7b746 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Call for Papers -- The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19b671da-5252-4497-9d4b-a0caffead11c · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens A task-optimized neural network repli- cates human auditory behavior, predicts brain responses, and re- veals a cortical processing hierarchy,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d7f67bdb-df30-48c7-bf43-d124bd131507 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Toward a realistic model of speech processing in the brain with self-supervised learning,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 02a3ce6b-c9d5-4b31-b65d-c42660187c23 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Dissecting neural computations of the human auditory pathway using deep neural networks for speech,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aee78f11-cfe0-443e-8813-2da3653d8bce · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Many but not all deep neural network audio models capture brain responses and exhibit correspondence between model stages and brain regions,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 91f4b4e3-66e6-4218-ac8c-0418f1508ccb · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Speech taskonomy: Which speech tasks are the most predictive of fmri brain activity?
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3aff2a49-132d-4df5-803b-27c0a1a5079a · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Language in brains, minds, and machines,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0264a1d7-45ec-4d1d-a2b8-29e6e9f2387a · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Brain-tuned Speech Models Better Reflect Speech Processing Stages in the Brain
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3849522-5e98-42e5-b262-13da0475a8e1 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens An algorithm for the machine calculation of complex fourier series,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13352801-f597-4a72-8ac0-fc4301aedecb · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Lib- rispeech: An ASR corpus based on public domain audio books,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec17cb52-6a2a-46f9-a014-6dc96e011ce3 · outbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens codebook usage
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c87c66bf-5d6a-4851-91b1-d73aff9c5ccf · inbound
Representing Speech Through Autoregressive Prediction of Cochlear Tokens Representing Speech Through Autoregressive Prediction of Cochlear Tokens
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.