Pith. sign in

Paper Citation Record · LEDGER

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 3 inbound Pith citation observations for arXiv:2508.17863.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.17863 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:46:49.282451Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T00:56:43.991936Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T07:27:44.920929Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9b52bf9-b6e4-40c0-b305-81f589a83a6f · outbound

This paper cites On The Landscape of Spoken Language Models: A Comprehensive Survey.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs On The Landscape of Spoken Language Models: A Comprehensive Survey

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.064133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.064133Z digest=sha256:dd02578909b0a2b3d7db75ed476330539aa960fed8200531e061e9d0ab088bb4

Observation 3aa58d23-2c8b-4667-9f07-2639666c394f · outbound

This paper cites Qwen Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.124721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.124721Z digest=sha256:df5dd58d2e5d0669659ea7f328c7577b560a8d0ca07a89b369e36465f51cb12e

Observation 555c79fd-572b-4e21-916f-edbf18b4c574 · outbound

This paper cites SLURP: A Spoken Language Understanding Resource Package.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SLURP: A Spoken Language Understanding Resource Package

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.223367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.223367Z digest=sha256:16f961a766b9f65865e4ee41c9638b55fcdd4257e9e65b509ed00985b563ec3b

Observation ef17f421-9602-4c3a-9911-b47eb4b02ad4 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.957548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.328362Z digest=sha256:cdc915e69fa236829e38a484168918759ae152624b62b465af751888f3dfe8f9

Observation 21e7f332-5c39-4404-8cd5-2d104bc89f32 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.949765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.390216Z digest=sha256:eff5900c3d839cc462ea6b26e9ec8bd5a1e9d3fb4de7e61491a4ca6d169eaca7

Observation aeff3fc1-8532-4454-a968-25da03344a4f · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.487403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.487403Z digest=sha256:ce0b4eb3f470013959796411cd868a2b31c5afd28c1fc75fca3d27e590584c5a

Observation 63cde880-590b-4f8a-a196-9b77100c7de0 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.941858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.555373Z digest=sha256:3c642de054c491a2ee48643610a007fd469faf6b6fab4740bf7a73ddeab82149

Observation 59d58ead-f7e1-4f16-a889-aa6095d78d70 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.691984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.691984Z digest=sha256:799cd7ac0f93d3564b53d3b9226f38dafb1f7f8f54836c83e5a848c8f94df539

Observation 8c6a3ded-de9b-4d0d-8255-efc9e2389adb · outbound

This paper cites LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.765382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.765382Z digest=sha256:4ae19d93cff082f901815efdc18392b69cee2982956f1803711e8cc866babfd9

Observation 52db9ce8-41ec-454d-8e0a-076ed64207bc · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.934533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:47.860324Z digest=sha256:bc1a8a4431963689579e5467ab501a5064a880294fc1acb0d04f439fa9eeebc2

Observation 273e77fb-f566-4c2a-bd53-c60e976523b4 · outbound

This paper cites Qwen2-Audio Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen2-Audio Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:47.924933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:47.924933Z digest=sha256:f6f28ba6ce822e9617517b8caaa728db1b6a13a4b6a29b0503172b3afda84c25

Observation 251c40b2-4f2c-478d-b659-c45cef6a6c6c · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.021893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.021893Z digest=sha256:283f15c6de68cd3dbf33d437e10802c5a96ec143b6851fcdac61b4be5c9cc64e

Observation b90e7ae5-32b0-4904-aa3f-5e5b23cb3264 · outbound

This paper cites Recent Advances in Speech Language Models: A Survey.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.070568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.070568Z digest=sha256:11257f349b6690965486d65ee3fd7a171bac536720a8029f493e3af78d32e212

Observation 25d5a620-79c0-4275-94f5-f82068149967 · outbound

This paper cites High Fidelity Neural Audio Compression.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs High Fidelity Neural Audio Compression

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.093307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.093307Z digest=sha256:f21785b68e706dd80df30c08d6a3e72e1145907e824e7a4ee75ee764129884fd

Observation 53e548ca-4d3d-48ad-8491-64055f7569f8 · outbound

This paper cites Exploring the Benefits of Tokenization of Discrete Acoustic Units.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Exploring the Benefits of Tokenization of Discrete Acoustic Units

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.652664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.202777Z digest=sha256:69b11b799c2343dff4a59c6c05a19ade3f492d6218a2cc9fc96f87e42170f804

Observation f513f72a-db5f-4106-bc4c-05b4d09d7087 · outbound

This paper cites The Llama 3 Herd of Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs The Llama 3 Herd of Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.356620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.356620Z digest=sha256:ff973980e1bf185740dbfc00ac8b58a3b1b2821c8ca9957041c485a344778370

Observation 2bba73a3-5426-4bf6-bedb-0330b9655176 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.926577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.486899Z digest=sha256:bd145870c055e0e95df738b679f0182f09979eb0c75e0f9766e662b2526773df

Observation 60e763a5-eb26-4df5-868b-81e21fab4b52 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.918359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.635372Z digest=sha256:2b55ee1b653f5079488ac3eebef105006addb91241971a279c050ec3c802c88f

Observation 003dd5a5-a870-433e-8397-f892f8ba0cb1 · outbound

This paper cites Listen, Think, and Understand.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Listen, Think, and Understand

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.796047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.796047Z digest=sha256:16d4d341e80bc926bf528a2f484e03c7143ad116ba721840d5937d658c371fd0

Observation 9ea639e1-8de5-4d71-86c4-f09baeca2449 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.909209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:48.963009Z digest=sha256:a02de023344ab6eb4f341535af3af1155ca6de902718ae0c78456467bf5d7cbb

Observation 8e34093a-3e8c-481c-93d0-767033c9d995 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LoRA: Low-Rank Adaptation of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.182080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.182080Z digest=sha256:ca014ee16320d3c66fe4065e7c9026029cfd9a2b702a77db4653b4a3d4f103e7

Observation 5288e3c4-2220-4f74-bd8a-ccf2e787a69d · outbound

This paper cites WavChat: A Survey of Spoken Dialogue Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs WavChat: A Survey of Spoken Dialogue Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.184979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.184979Z digest=sha256:7483fe7ddc27b8a9012018424f5e01985977ca5fe06db9b6bd9a33ce799887af

Observation 127383d8-3888-406f-bf51-3a37cc4d4d52 · outbound

This paper cites WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.187588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.187588Z digest=sha256:10ac9798a453c55bd00333094bfaa748d6c7edbaa7e761cd5c334c99b7a598f5

Observation 950e6ffd-d212-4326-9577-37c7267beefe · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.190729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.190729Z digest=sha256:5355cf44635298dd5ce11020af8c19a94b04122ed76c9255bbee0ecde767a2ec

Observation 8a3261b7-ad91-47e5-9be0-886189330964 · outbound

This paper cites Continuous Speech Tokenizer in Text To Speech.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Continuous Speech Tokenizer in Text To Speech

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.602331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.193445Z digest=sha256:f684133426877581624f501b5e7b99f96203a9fd34dc5cca03ec8150e8537aaf

Observation ca7456a1-cfec-42dc-8ed9-fd7c83588c02 · outbound

This paper cites An Embarrassingly Simple Approach for LLM with Strong ASR Capacity.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.196106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.196106Z digest=sha256:df40c2c29a2809f3f7f25095d0fd783e6f29be4cd35577087946645cbcc9a995

Observation 288f89b5-1b4f-4b77-a6be-f36784bd7c75 · outbound

This paper cites PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.581631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.199048Z digest=sha256:c9c28266f14b5e168ee6ce7d4d6eadedd9d25b968bade6d5ff686894145634fa

Observation 7dcb80e2-5104-49ea-bc20-9dc7a316d8d6 · outbound

This paper cites How Should We Extract Discrete Audio Tokens from Self-Supervised Models?.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs How Should We Extract Discrete Audio Tokens from Self-Supervised Models?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.201803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.201803Z digest=sha256:0dcbed4aca81006ff13952482ef05c5c3fa006111cfb49e6c445e944a201bdc9

Observation 36df5c38-9ff3-478f-b9d6-01f7a273e5bc · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.900264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.204517Z digest=sha256:919b3579cc448eff2454a59488713111b012481da51e9f12c0c0692208996bf6

Observation bd0e828b-c896-4ba8-9bdb-1f81344a04c7 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.892093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.206910Z digest=sha256:b73113c5f51d66bba9e3fa4b9f12fe7279ee26f0bce65aa9df8487c6a25ad967

Observation 6296b8dd-7b47-4047-8c23-5c862a9ecd22 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.883742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.209410Z digest=sha256:d9343f04786ec59ad5eb88506b22b9b5b4306bcbafab0a160d963955d950338e

Observation 49a1d482-f3a4-417d-8dde-03d944f1e11b · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.211813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.211813Z digest=sha256:6288a087aa6608f329c34b3614f581c197451f0133cafececf8aceff13ae22e7

Observation e55f726f-0bc6-463b-b818-abeb078d964f · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.875522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.214337Z digest=sha256:06f1c20e490b93f5da5e6643c61c80e243b7fbf1e2ab6fdd10d9ab1adffcc1d5

Observation 6b3b7ae4-e59d-4345-a001-4097b9b36a86 · outbound

This paper cites AudioPaLM: A Large Language Model That Can Speak and Listen.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs AudioPaLM: A Large Language Model That Can Speak and Listen

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.216718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.216718Z digest=sha256:faa0fbced9c26ed3137bb1f08df15ca64b64e6d6312c2a2b439e5e75e697643b

Observation 871b63f0-a24a-474b-8214-bd978d1f0d83 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.867337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.220533Z digest=sha256:257c285133008da157d644ab4925cf05f27e090b7e131f44caeb59aa48744f21

Observation c582b0e4-44dc-4def-87c9-84d8c2620ae5 · outbound

This paper cites DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.441586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.223732Z digest=sha256:88414ad13efd2dd3255426eed1877229674b64592d7275f1e8bcfccbc5799eac

Observation 4073ece4-78b8-40ac-baa2-e52fee4a7071 · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.226321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.226321Z digest=sha256:9d5a0f1e1d161064e90b607f0a81d158419555f675fdb3f615f3daab09885b66

Observation 7a2eb8ed-93bf-4b58-9b78-14df767b8381 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.858172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.229228Z digest=sha256:251f607e87843753d34298160a2beef4a54f31f0ad2398230d22d784d7a15320

Observation 58daee93-3456-4c1e-a19a-d8a94f8dfe07 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.849882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.231950Z digest=sha256:91296877b1a2c035fc4ee62e009a13081646215da77aa4201aeda4be153d5794

Observation c278616e-00d3-432b-aeb3-9c587e7b02bb · outbound

This paper cites BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.234700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.234700Z digest=sha256:2888025728ab0f7dee7898d66666d248137cd9bc2ec99048cd832ee123cd668e

Observation b3502b97-07a2-4492-a9f7-d06e1864cb13 · outbound

This paper cites InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:46:49.412149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.237800Z digest=sha256:57a45f40dc2a46c0b2512bddcf6acc0f0ff641f7750fd28073e017e043b1c71e

Observation 82e7c989-1115-41fd-9d0b-10bca00a2b5c · outbound

This paper cites VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.240587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.240587Z digest=sha256:b0c91383e9b730c6609754a9bbbb1484b473df8e2a1f03a9e521db1fa059afca

Observation 31db02af-61fd-466c-b863-13297a7e3d11 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.841549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.243540Z digest=sha256:b16b88b31260bbd31c4f9d82c6ce2f157c377421d3319ead2369364de75844ef

Observation 1e47ca38-e878-4336-b0eb-ce153a8f1336 · outbound

This paper cites Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.246073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.246073Z digest=sha256:69372c711beac7ec2c7dbbca29a0ef79499a73a40ec36220ad6ec8725a6e3515

Observation 794ca939-5f0f-460b-8cb1-be4930fd03a9 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.832557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.248919Z digest=sha256:fe7e8c7a6d04e58796758c933044084d35b4fbde86c27062802879e8e0e9524b

Observation 60f263ac-2bba-43eb-8519-23d2f46f4e95 · outbound

This paper cites Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.251449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.251449Z digest=sha256:42e583843390069d84d5d14bd7ed3e6b15bccb5668e13d396a4e9d8da6c00a08

Observation 5b315595-faf4-4481-9962-1b2c55c9451b · outbound

This paper cites Qwen2.5-Omni Technical Report.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Qwen2.5-Omni Technical Report

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.254645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.254645Z digest=sha256:af03e0655ec9497efe52eeead4fb60c3169586e11bbe2b3149a23388ecedabbb

Observation b1a419f0-0c28-45fa-8f38-e6035255882b · outbound

This paper cites Comparing Discrete and Continuous Space LLMs for Speech Recognition.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Comparing Discrete and Continuous Space LLMs for Speech Recognition

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.257437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.257437Z digest=sha256:637535f78145c666a17a9e0a501b913c654979497dcc3d7d018845f7aa273cc3

Observation edfe98b9-dc87-4e50-9f77-c65c579c2dbf · outbound

This paper cites ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs ALMTokenizer: A Low-bitrate and Semantic-rich Audio Codec Tokenizer for Audio Language Modeling

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.260250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.260250Z digest=sha256:241f02301b99849fadbae41034c37aa140ff89281437a49ff896331c47201e87

Observation 144c10e5-3bbb-4014-b9ec-f8b7e4a9d11e · outbound

This paper cites GigaST: A 10,000-hour Pseudo Speech Translation Corpus.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GigaST: A 10,000-hour Pseudo Speech Translation Corpus

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.262707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.262707Z digest=sha256:e5da80135325e0f7edb5b533121957031a8bcdfe9620fc7f8a210930903adc3b

Observation a7fb8689-7e29-4934-9057-0f440c292082 · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.265779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.265779Z digest=sha256:c6d94ebe0a69a81c5704ecae6f36a8e422ba1ba967903e41f9e1c47251130df9

Observation e0daf9c9-2cad-4e82-9a27-23774e80b77b · outbound

This paper cites an unresolved cited work.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:46:49.823915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T16:46:49.268492Z digest=sha256:c4c66fccdc4753ac7ff23b42d43fb1a160ff2f84875fb22292134391da57583c

Observation 494656bc-0d91-41ad-b14b-d49cb69a3022 · outbound

This paper cites GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.270919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.270919Z digest=sha256:0724a683cd54ebdc9d3f8e122ae2767d9ec16e66681e06800261ae26a93312ff

Observation fcec29af-ec7c-4da5-a927-e41fc3fa48d8 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.273789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.273789Z digest=sha256:6784993f75b8bc4f6298dad1397edd140685f32aa0a295bb5c3172853c14aaed

Observation e89c1e01-398a-4a9b-9a5f-14c3a259b39d · outbound

This paper cites SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.276513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.276513Z digest=sha256:079a7c7a060d90e39c9b250bbac84b22770e2b9b100ba8f78a6badb145460843

Observation 5d45cb9e-8915-4788-a5b7-956688efa298 · outbound

This paper cites online" 'onlinestring :=.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs online" 'onlinestring :=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.279342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.279342Z digest=sha256:c516da6a8956e0eb057f0dc37ac2bd63fbe30c24b230ea11967e6653b218d17c

Observation 446a4bcb-cf18-4417-b33f-cc99e74a81ab · outbound

This paper cites write newline.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.282451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.282451Z digest=sha256:ac06facc173203ab1acc3544aac67c45224ab9e1ab5fefa8365b77392cd9e793

Pith citing papers

Observation 5e0eb325-70d6-4597-a064-93b3bef5b1ae · inbound

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages cites this paper.

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.222863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T16:27:37.596817Z digest=sha256:5d8ee17d3d98286bba885be17ec526d830b48935d483433517158ae68adff496

Observation 3667a07c-94e2-4be2-8936-8432bcc774f9 · inbound

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation cites this paper.

Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:27:44.922554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T12:04:50.483329Z digest=sha256:dfe8786c58dfca3c1d09ef56cc52dff35a79dcb45116f41fdf48e6eabf884921

Observation 89a5f7bf-cfe4-4d3e-be39-55c2f5ca66a6 · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.248185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:dda3651d675a5de237f502b88375f371895e8d0de355d5c676efedbb8c4af3c1