Pith. sign in

Paper Citation Record · LEDGER

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

As of 19 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 3 inbound Pith citation observations for arXiv:2508.17863.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.17863 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:46:49.282451Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T00:56:43.991936Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T07:27:44.920929Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9b52bf9-b6e4-40c0-b305-81f589a83a6f · outbound

This paper cites On The Landscape of Spoken Language Models: A Comprehensive Survey.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs On The Landscape of Spoken Language Models: A Comprehensive Survey

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.064133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.064133Z digest=sha256:07c9a91b864a3cbc6f24207d4b8392165bc02088213d81f54cbca4a3fa790e19

Observation 3aa58d23-2c8b-4667-9f07-2639666c394f · outbound

This paper cites Qwen Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.124721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.124721Z digest=sha256:1d3ca4ed5434b5362cbd7b6cb4aaaf451efeac8787f088f7199b0ff7b614ba26

Observation 555c79fd-572b-4e21-916f-edbf18b4c574 · outbound

This paper cites SLURP: A Spoken Language Understanding Resource Package.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SLURP: A Spoken Language Understanding Resource Package

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.223367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.223367Z digest=sha256:4e4752920e7d704a988beae35dcf165b6e5b26905a7eb62c5f60f3280d03fb2d

Observation ef17f421-9602-4c3a-9911-b47eb4b02ad4 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.957548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.328362Z digest=sha256:c874186f88fd67ac5748971bb23c5e30b12d2e5531a7d5df73d507c43f6064f8

Observation 21e7f332-5c39-4404-8cd5-2d104bc89f32 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.949765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.390216Z digest=sha256:f1f5c2d378855b7a221a89d4456a87940bdbe43efe8e6022718af9737c81ba46

Observation aeff3fc1-8532-4454-a968-25da03344a4f · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.487403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.487403Z digest=sha256:e1905a3d47228b6c3db6c74cfb75e5fb59ded40307878ace5bd3ab34ac71aad5

Observation 63cde880-590b-4f8a-a196-9b77100c7de0 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.941858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.555373Z digest=sha256:60bb570441430fd921196e4d334c3ccebf174a4b8be41b95fddef44de58d103a

Observation 59d58ead-f7e1-4f16-a889-aa6095d78d70 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.691984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.691984Z digest=sha256:9c1b1b411bfab9d31bec853b944c0f86a4bbb8b7d3e8aa7aacf5b730b6cae0f8

Observation 8c6a3ded-de9b-4d0d-8255-efc9e2389adb · outbound

This paper cites LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.765382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.765382Z digest=sha256:0c0bb44a0ccfdb5bedaa48029a22558febbb906cdb8aa996457df7d1469aec64

Observation 52db9ce8-41ec-454d-8e0a-076ed64207bc · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.934533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.860324Z digest=sha256:a91e8057cc0cc8900de6cde7fbc936960d609b8022752693f0b4311570120d45

Observation 273e77fb-f566-4c2a-bd53-c60e976523b4 · outbound

This paper cites Qwen2-Audio Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen2-Audio Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.924933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.924933Z digest=sha256:e49b2a08fbf3d2ac5bfbe264b66afdd8c3e16ea257eaa573337cad7b68e4390f

Observation 251c40b2-4f2c-478d-b659-c45cef6a6c6c · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.021893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.021893Z digest=sha256:3e1a20cb775ea80a2d1684918de0e4c2b3ea9dab06cfdee53be37f7e9f485894

Observation b90e7ae5-32b0-4904-aa3f-5e5b23cb3264 · outbound

This paper cites Recent Advances in Speech Language Models: A Survey.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.070568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.070568Z digest=sha256:310ad3456716287673cb100e63931ce69f998f8baaa0d3061bfe44db25412381

Observation 25d5a620-79c0-4275-94f5-f82068149967 · outbound

This paper cites High Fidelity Neural Audio Compression.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs High Fidelity Neural Audio Compression

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.093307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.093307Z digest=sha256:ac1823b5ecea7593b02f9e58a8fc18d760fa30e1e20c2828a4fbec002721320b

Observation 53e548ca-4d3d-48ad-8491-64055f7569f8 · outbound

This paper cites Exploring the Benefits of Tokenization of Discrete Acoustic Units.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Exploring the Benefits of Tokenization of Discrete Acoustic Units

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.652664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.202777Z digest=sha256:eb598e6a8d510dbc5569239eace6811415a4094188cde05a353736fbc0208a19

Observation f513f72a-db5f-4106-bc4c-05b4d09d7087 · outbound

This paper cites The Llama 3 Herd of Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs The Llama 3 Herd of Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.356620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.356620Z digest=sha256:3920cb610345684ecaa943ddf05d89d05871fc3db3f3899027e11d89c1cfa079

Observation 2bba73a3-5426-4bf6-bedb-0330b9655176 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.926577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.486899Z digest=sha256:e59d78e4ec8f9094c5f5a2799bc87eefc508e05be682c963f0b2301a87479e76

Observation 60e763a5-eb26-4df5-868b-81e21fab4b52 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.918359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.635372Z digest=sha256:3057189131b5446999e40d3fe1ca8872bc0054a0102128efdec0b8bc200ac094

Observation 003dd5a5-a870-433e-8397-f892f8ba0cb1 · outbound

This paper cites Listen, Think, and Understand.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Listen, Think, and Understand

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.796047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.796047Z digest=sha256:854e01bce3198fc317edbe887efcc9698b7bedd06d438d4ccca364585c8ef184

Observation 9ea639e1-8de5-4d71-86c4-f09baeca2449 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.909209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.963009Z digest=sha256:802b861c48fff498ebc7c5ccc326fe37fc98879a1b669c8d79785a3fe2548cf3

Observation 8e34093a-3e8c-481c-93d0-767033c9d995 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LoRA: Low-Rank Adaptation of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.182080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.182080Z digest=sha256:3197c9906851e3f1e02f4b3e014839bddcf76b3578115fcffb535940f1cec74e

Observation 5288e3c4-2220-4f74-bd8a-ccf2e787a69d · outbound

This paper cites WavChat: A Survey of Spoken Dialogue Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs WavChat: A Survey of Spoken Dialogue Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.184979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.184979Z digest=sha256:51bc211a6160f23359c5a779ad8f3198151c70d004522fc3b8a19d1b3e2b256f

Observation 127383d8-3888-406f-bf51-3a37cc4d4d52 · outbound

This paper cites WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.187588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.187588Z digest=sha256:4ae7d5b22a870fc4badc7ad8e34a29b3db492dac5ca4a81e5c26a76dd15ab846

Observation 950e6ffd-d212-4326-9577-37c7267beefe · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.190729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.190729Z digest=sha256:896addec1e829f4d2a7c3a988f47fc9fc490cbcdeaed0f9e5a0aa54cf3378939

Observation 8a3261b7-ad91-47e5-9be0-886189330964 · outbound

This paper cites Continuous Speech Tokenizer in Text To Speech.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Continuous Speech Tokenizer in Text To Speech

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.602331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.193445Z digest=sha256:5e9fc5ccadc27d2be31c13b2e0cca6cdbfcef209c400db90d889cc8690c80af0

Observation ca7456a1-cfec-42dc-8ed9-fd7c83588c02 · outbound

This paper cites An Embarrassingly Simple Approach for LLM with Strong ASR Capacity.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.196106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.196106Z digest=sha256:1244cbdb1a41ede9ea86278fb020cd1df222229826398ababa02fae697cc7aab

Observation 288f89b5-1b4f-4b77-a6be-f36784bd7c75 · outbound

This paper cites PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.581631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.199048Z digest=sha256:6cf6b2a83bde3f1a50cb87fcdf5a690d612de7510a95c474dcd3c741ca347101

Observation 7dcb80e2-5104-49ea-bc20-9dc7a316d8d6 · outbound

This paper cites How Should We Extract Discrete Audio Tokens from Self-Supervised Models?.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs How Should We Extract Discrete Audio Tokens from Self-Supervised Models?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.201803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.201803Z digest=sha256:e391448b05ce2b757181b4056bf6e04b3952e0da24aaabf47143d965d3db3add

Observation 36df5c38-9ff3-478f-b9d6-01f7a273e5bc · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.900264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.204517Z digest=sha256:4188b91efe5a9923470b7adef66e1172ffd3b4f78458f7c0caeda82232164546

Observation bd0e828b-c896-4ba8-9bdb-1f81344a04c7 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.892093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.206910Z digest=sha256:15a9bcb973802f1ab52f4bd14c7f2759e83d2f86bc09167b3d82ee2aeda6e13b

Observation 6296b8dd-7b47-4047-8c23-5c862a9ecd22 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.883742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.209410Z digest=sha256:c7cef3f44a851460d48e1ec95c114293b3194adcfa427d9a3e09000c7c2c0add

Observation 49a1d482-f3a4-417d-8dde-03d944f1e11b · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.211813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.211813Z digest=sha256:faba5ebe3e507cd821b2af7129ec6bf4eb5f0ef5dad3df62deeb24491b83d924

Observation e55f726f-0bc6-463b-b818-abeb078d964f · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.875522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.214337Z digest=sha256:c9ba90c7e2d1da8df4de1cdd74bea6c628ad4cd132c27db16fe4f0a414a77ba9

Observation 6b3b7ae4-e59d-4345-a001-4097b9b36a86 · outbound

This paper cites AudioPaLM: A Large Language Model That Can Speak and Listen.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs AudioPaLM: A Large Language Model That Can Speak and Listen

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.216718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.216718Z digest=sha256:56ea0c9267c327dd5310149bfe4c1b8d2d7b0693c39dd1c64e926c07f10a0d5d

Observation 871b63f0-a24a-474b-8214-bd978d1f0d83 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.867337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.220533Z digest=sha256:b44f06df59cf57df1b432c2c4fac54dbe6d243d8117110720b8e9b4d776f8133

Observation c582b0e4-44dc-4def-87c9-84d8c2620ae5 · outbound

This paper cites DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.441586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.223732Z digest=sha256:db8aaa45a50c8a92f3f4e28c72493d834b3e384241794388113ea4b9a9459b83

Observation 4073ece4-78b8-40ac-baa2-e52fee4a7071 · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.226321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.226321Z digest=sha256:4cff84eec4ed316a4026f817e14c7e2dd8be678e4bc1fe88d40ad07d1bd4f343

Observation 7a2eb8ed-93bf-4b58-9b78-14df767b8381 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.858172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.229228Z digest=sha256:537edfb8fbabf6647e5c8a27ce9d9ca2fda68729641f81971c18d37186a5a480

Observation 58daee93-3456-4c1e-a19a-d8a94f8dfe07 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.849882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.231950Z digest=sha256:16631b79caf835064ec2a6f882d90e224031e7d575a05451fa2aedfd4826bb39

Observation c278616e-00d3-432b-aeb3-9c587e7b02bb · outbound

This paper cites BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.234700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.234700Z digest=sha256:bf59b964f18556931b6f7cf29bf2870f2bb96f179f2850a6ec5297a528727d0d

Observation b3502b97-07a2-4492-a9f7-d06e1864cb13 · outbound

This paper cites InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.412149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.237800Z digest=sha256:28550769aa2bdd3e186119c1e22bbbbdf2de791923c61278310485c287406e40

Observation 82e7c989-1115-41fd-9d0b-10bca00a2b5c · outbound

This paper cites VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.240587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.240587Z digest=sha256:0233b4281e9445771cd87252eddab6762145db53e1f69518015a5d79055aa14a

Observation 31db02af-61fd-466c-b863-13297a7e3d11 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.841549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.243540Z digest=sha256:46f703be47ffaf3aec5aed13b8ba3cdb559607b7155920b478d3f98a1e6719bd

Observation 1e47ca38-e878-4336-b0eb-ce153a8f1336 · outbound

This paper cites Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.246073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.246073Z digest=sha256:2254bf2a25cf24e448bda1788db0922953f6acf17572192f93be8b5724c67257

Observation 794ca939-5f0f-460b-8cb1-be4930fd03a9 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.832557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.248919Z digest=sha256:a550f8cc37cdf4d525ff60837ac2bc723420dae90be0f233acda725ebf100eb8

Observation 60f263ac-2bba-43eb-8519-23d2f46f4e95 · outbound

This paper cites Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.251449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.251449Z digest=sha256:eac06965e88def163cf38db7a6119883e6e774b6f0fa1bc75daaf1d4a59ac78f

Observation 5b315595-faf4-4481-9962-1b2c55c9451b · outbound

This paper cites Qwen2.5-Omni Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen2.5-Omni Technical Report

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.254645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.254645Z digest=sha256:381abad951e165001b1d9f631ffd1a8260d36cdff1b26422bc30cb01aea9b2e4

Observation b1a419f0-0c28-45fa-8f38-e6035255882b · outbound

This paper cites Comparing Discrete and Continuous Space LLMs for Speech Recognition.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Comparing Discrete and Continuous Space LLMs for Speech Recognition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.257437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.257437Z digest=sha256:68cd8fbd6e5bbfe3fbf18d089108c118dae31616e865851a8920dc3cfa2a8612

Observation edfe98b9-dc87-4e50-9f77-c65c579c2dbf · outbound

This paper cites ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.260250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.260250Z digest=sha256:230bb3f10f6849640ae7e8dd5c9425e423e7c0bc359a1e6f780a8bc020bd38e8

Observation 144c10e5-3bbb-4014-b9ec-f8b7e4a9d11e · outbound

This paper cites GigaST: A 10,000-hour Pseudo Speech Translation Corpus.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GigaST: A 10,000-hour Pseudo Speech Translation Corpus

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.262707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.262707Z digest=sha256:a8e6270677726a05fe1bae3e940a8ff3665bfe2e9b504a8baf31fc2aa2345959

Observation a7fb8689-7e29-4934-9057-0f440c292082 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.265779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.265779Z digest=sha256:6753cf1211208ec70066f6ed71ab81870d6e5f3e6bd81400aecb157e87a787fd

Observation e0daf9c9-2cad-4e82-9a27-23774e80b77b · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.823915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.268492Z digest=sha256:68cc7fa082e9ee48466dd61b3e3952d5fbb6f0b45c0766d6ab9e31ae4a6b0d8d

Observation 494656bc-0d91-41ad-b14b-d49cb69a3022 · outbound

This paper cites GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.270919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.270919Z digest=sha256:89803a32ec007f60fb0e8c1fbf20aff63388715ee038941bcae9c51a145d0702

Observation fcec29af-ec7c-4da5-a927-e41fc3fa48d8 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.273789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.273789Z digest=sha256:6f7a6ddfcbcaa50e45a33d95d9c0c776689335c502b2dd50bf77024039374cca

Observation e89c1e01-398a-4a9b-9a5f-14c3a259b39d · outbound

This paper cites SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.276513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.276513Z digest=sha256:d4b5fbee1b494f1a07eaea86602e0847b7b15d1d84fa9df482e93bd3665c59bc

Observation 5d45cb9e-8915-4788-a5b7-956688efa298 · outbound

This paper cites online" 'onlinestring :=.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs online" 'onlinestring :=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.279342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.279342Z digest=sha256:8cd82261e24d8e8056beca43bf156f449735240365e1291dc1620a483ab19ced

Observation 446a4bcb-cf18-4417-b33f-cc99e74a81ab · outbound

This paper cites write newline.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.282451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.282451Z digest=sha256:0dd9ccbfaaa33230b6df1361bddaf9791d83d4564fa8eb941b6e92a148851a61

Pith citing papers

Observation 5e0eb325-70d6-4597-a064-93b3bef5b1ae · inbound

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages cites this paper.

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.222863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:27:37.596817Z digest=sha256:4714ef39fe08e1e888b41de669b47b56b166c03bfa1647208ea3d390b27b91f1

Observation 3667a07c-94e2-4be2-8936-8432bcc774f9 · inbound

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation cites this paper.

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:27:44.922554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T12:04:50.483329Z digest=sha256:8166821da9188c0e5170e73052d597e5855fa16c36556a8004da5dce1ec55deb

Observation 89a5f7bf-cfe4-4d3e-be39-55c2f5ca66a6 · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.248185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:a5070bbd01209df4ea659bf400eceddb42b59f5c87031abbbad69dff281d6825