Pith. sign in

Paper Citation Record · LEDGER

Recent Advances in Speech Language Models: A Survey

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2410.03751.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.03751 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:06.241559Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ded06926-e606-4b56-a2ba-23eb7a8141fe · inbound

On The Landscape of Spoken Language Models: A Comprehensive Survey cites this paper.

On The Landscape of Spoken Language Models: A Comprehensive Survey Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:45:08.060933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T20:44:57.476464Z digest=sha256:4108b14fd1dca0c527bc6997fd00e39ad80438e1c2569115aaf3a3b8fc375396

Observation 776258d4-e59f-4fe8-a55e-cb813b7dfa3a · inbound

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems cites this paper.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.241559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.241559Z digest=sha256:54676cf755973cf537f31c96b90ab9caf7f57ca63ab7ac1c178073d0ca6485e5

Observation 8b62c3c8-bdc0-46e4-bea0-2e93f641eeb5 · inbound

Speechless: Speech Instruction Training Without Speech for Low Resource Languages cites this paper.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.984529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:15.984529Z digest=sha256:0844660c1d8a0d5e3719aee61d093ab37e86b9e68da46a2d1f2b89be8cd0e405

Observation 361d5b18-6658-4d8a-b318-2349c31d9196 · inbound

Voice of a Continent: Mapping Africa's Speech Technology Frontier cites this paper.

Voice of a Continent: Mapping Africa's Speech Technology Frontier Recent Advances in Speech Language Models: A Survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:18.428449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:34:18.428449Z digest=sha256:4d9849055927fb39179e3ebec772a4401845949c7a5abcd65670c94ab83d4aca

Observation 63c5ade6-2bbe-4a47-bec2-2bbe12f9f813 · inbound

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games cites this paper.

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games Recent Advances in Speech Language Models: A Survey

Reference 94

Resolution
malformed identifier
arxiv_id, observed 2026-05-19T12:02:16.643857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T12:01:42.681135Z digest=sha256:6bf3eedce3e7b1e5812dbb7143ec9369feaf6342f388cda87b91ff1c30dcc2d8

Observation 41695169-4cb5-48d0-8092-c55a769324ee · inbound

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model cites this paper.

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model Recent Advances in Speech Language Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:11.869766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:11.869766Z digest=sha256:2407919a83d8b1a3278da203ab09db33d014d51c6682d10d9d216c3d149707ba

Observation 038a4d0f-3d1f-4eb3-8280-74bcfe506cac · inbound

Intelligibility of Text-to-Speech Systems for Mathematical Expressions cites this paper.

Intelligibility of Text-to-Speech Systems for Mathematical Expressions Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:39:47.340916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:39:47.340916Z digest=sha256:805aaf559843d966a918a6858bec89a49b27d85851083a63dea64827c2d2e131

Observation bb72e72e-eb31-46a0-8cd5-6748b6f876b0 · inbound

OpusLM: A Family of Open Unified Speech Language Models cites this paper.

OpusLM: A Family of Open Unified Speech Language Models Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:36.004978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:36.004978Z digest=sha256:c418cb01a6524a7e12de5b94af0fab0850e512f9412bc29876ff01e2424ce52e

Observation 9037fce0-afe2-4dc2-b1df-ea9ddf694be3 · inbound

GR-LLMs: Recent Advances in Generative Recommendation Based on Large Language Models cites this paper.

GR-LLMs: Recent Advances in Generative Recommendation Based on Large Language Models Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:00.999223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:00.999223Z digest=sha256:3528b7985a7a88f4d90ad2b0571a1e89cac6c906baf045e7bc6f2ed817e2f5f6

Observation 1020d1a2-ec69-4a76-8428-208b2c7a54f9 · inbound

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine cites this paper.

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine Recent Advances in Speech Language Models: A Survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:34.319918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:47:34.319918Z digest=sha256:6b9483ffee239d0bfcb99efd1fed332b2df7924aeec54c51b45f258f530e6a7b

Observation b90e7ae5-32b0-4904-aa3f-5e5b23cb3264 · inbound

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs cites this paper.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.070568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.070568Z digest=sha256:f22737cfae3d6a22d4584b73d7440a033d4bc7f6a192a02c1fc3ed5c5f2e0c3c

Observation 7082b871-1105-4515-aeb9-7af8f2c182b1 · inbound

Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts cites this paper.

Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T12:14:52.274618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:14:52.274618Z digest=sha256:2955070e0526225536f187e53ff63e422f129e755d7bac4568ad5e17a86048c2

Observation bb42877f-23b2-493b-9ed0-d2a644d69623 · inbound

Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models cites this paper.

Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T11:25:29.124574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:25:29.124574Z digest=sha256:c654369635f7e9a3db8ee9ffa3016a8c9ef4eea1c5a4c06b87f041e60ec3cded

Observation 098c0c90-6348-4246-8634-a424d1abd924 · inbound

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs cites this paper.

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs Recent Advances in Speech Language Models: A Survey

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:11:43.169892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T18:07:01.257126Z digest=sha256:d4d800294b6f7aff8a7837e0514e636a3964880eee9129fec8027005e5875098

Observation f4569ce4-5a72-4073-a5ad-9e98464af008 · inbound

TokenChain: A Discrete Speech Chain via Semantic Token Modeling cites this paper.

TokenChain: A Discrete Speech Chain via Semantic Token Modeling Recent Advances in Speech Language Models: A Survey

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:16:09.482195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T09:14:58.542628Z digest=sha256:f0351d612f6c4a3da040496df5750f1d804f638ff732f54ac4466998bf65b339

Observation 70a23ab9-d3d9-4e3b-a5a1-5a7e2e7923d0 · inbound

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models cites this paper.

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models Recent Advances in Speech Language Models: A Survey

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T07:46:03.599452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T07:43:23.913399Z digest=sha256:f436859cf393bc7b9242a0fcbc91d6afe4e018aeee290a39dff77d111323d20d

Observation a159f64d-02d7-402d-9b29-77da7e49d8cd · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Recent Advances in Speech Language Models: A Survey

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.166907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:f39e42a0c1bb2bd7a3511c82ecefda553164228bfe3e41363c78d55961a741b5

Observation f096af3e-4743-424b-9274-f767bba495c4 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:18.429430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T08:34:56.898815Z digest=sha256:8be6fe017541569f75d51005a41117733b5f847bb90de0bba05fd512108af2a8

Observation 0134ecc2-d108-4441-9734-5a020a491d06 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:44:07.790794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T10:43:27.176535Z digest=sha256:bf0ce301553e7739318488111b053133d2a328fd783d2a060b750505e98b32cc

Observation 3434c926-3507-48f9-9abf-af20e547d63a · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T22:57:11.059962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:57:11.059962Z digest=sha256:2cb611d9b0349284027cc98f3cefb3c4b7de705af23682c2aab9bfb6aeb6c0c6

Observation 94d926ba-c1c5-4cd3-8ec3-617393cf85f5 · inbound

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction cites this paper.

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction Recent Advances in Speech Language Models: A Survey

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:59:50.641103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T07:59:22.416259Z digest=sha256:8a0b6e306876f9c6201c21a8f996800d95f95e0ea11dffb1f2aa089abfb12ddf

Observation c94f2c74-304e-43c0-a0a7-0e9a07fef51c · inbound

Audio-Mind: An Auditable Agentic Framework for Audio Understanding cites this paper.

Audio-Mind: An Auditable Agentic Framework for Audio Understanding Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:13:17.598343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T10:03:54.653164Z digest=sha256:789d3687994029de01ac1d68f45b119661e416c87bb1d95f5522fc592118607a

Observation eadf26e4-da33-4572-955a-55ef64290ebb · inbound

MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond cites this paper.

MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond Recent Advances in Speech Language Models: A Survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T05:50:28.242656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:50:28.242656Z digest=sha256:73cef2f000fdd3e8d1489ad055174435bd7549a92e053396632de040c9bb6bf1