Pith. sign in

Paper Citation Record · LEDGER

Recent Advances in Speech Language Models: A Survey

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2410.03751.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.03751 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:06.241559Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ded06926-e606-4b56-a2ba-23eb7a8141fe · inbound

On The Landscape of Spoken Language Models: A Comprehensive Survey cites this paper.

On The Landscape of Spoken Language Models: A Comprehensive Survey Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:45:08.060933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T20:44:57.476464Z digest=sha256:e94c8939151a2760e334490e3365b86b2534914ae40805882f12e3ddca6f695b

Observation 776258d4-e59f-4fe8-a55e-cb813b7dfa3a · inbound

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems cites this paper.

Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.241559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:06.241559Z digest=sha256:54676cf755973cf537f31c96b90ab9caf7f57ca63ab7ac1c178073d0ca6485e5

Observation 8b62c3c8-bdc0-46e4-bea0-2e93f641eeb5 · inbound

Speechless: Speech Instruction Training Without Speech for Low Resource Languages cites this paper.

Speechless: Speech Instruction Training Without Speech for Low Resource Languages Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.984529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:15.984529Z digest=sha256:0844660c1d8a0d5e3719aee61d093ab37e86b9e68da46a2d1f2b89be8cd0e405

Observation 361d5b18-6658-4d8a-b318-2349c31d9196 · inbound

Voice of a Continent: Mapping Africa's Speech Technology Frontier cites this paper.

Voice of a Continent: Mapping Africa's Speech Technology Frontier Recent Advances in Speech Language Models: A Survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:18.428449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:34:18.428449Z digest=sha256:4d9849055927fb39179e3ebec772a4401845949c7a5abcd65670c94ab83d4aca

Observation 63c5ade6-2bbe-4a47-bec2-2bbe12f9f813 · inbound

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games cites this paper.

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games Recent Advances in Speech Language Models: A Survey

Reference 94

Resolution
malformed identifier
arxiv_id, observed 2026-05-19T12:02:16.643857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T12:01:42.681135Z digest=sha256:fb8fec8dda78b9abf01e8d58c4d9af9cf13a2a1f9e5b0b535f0bade64ad974b3

Observation 41695169-4cb5-48d0-8092-c55a769324ee · inbound

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model cites this paper.

Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model Recent Advances in Speech Language Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:11.869766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:11.869766Z digest=sha256:2407919a83d8b1a3278da203ab09db33d014d51c6682d10d9d216c3d149707ba

Observation 038a4d0f-3d1f-4eb3-8280-74bcfe506cac · inbound

Intelligibility of Text-to-Speech Systems for Mathematical Expressions cites this paper.

Intelligibility of Text-to-Speech Systems for Mathematical Expressions Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:39:47.340916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:39:47.340916Z digest=sha256:805aaf559843d966a918a6858bec89a49b27d85851083a63dea64827c2d2e131

Observation bb72e72e-eb31-46a0-8cd5-6748b6f876b0 · inbound

OpusLM: A Family of Open Unified Speech Language Models cites this paper.

OpusLM: A Family of Open Unified Speech Language Models Recent Advances in Speech Language Models: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:36.004978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:36.004978Z digest=sha256:0718ffcc27d24e1979ffd0f7de8a0ff690f3278bec984aff6f76829733ab5648

Observation 9037fce0-afe2-4dc2-b1df-ea9ddf694be3 · inbound

GR-LLMs: Recent Advances in Generative Recommendation Based on Large Language Models cites this paper.

GR-LLMs: Recent Advances in Generative Recommendation Based on Large Language Models Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:00.999223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:00.999223Z digest=sha256:3528b7985a7a88f4d90ad2b0571a1e89cac6c906baf045e7bc6f2ed817e2f5f6

Observation 1020d1a2-ec69-4a76-8428-208b2c7a54f9 · inbound

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine cites this paper.

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine Recent Advances in Speech Language Models: A Survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:34.319918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:47:34.319918Z digest=sha256:6b9483ffee239d0bfcb99efd1fed332b2df7924aeec54c51b45f258f530e6a7b

Observation b90e7ae5-32b0-4904-aa3f-5e5b23cb3264 · inbound

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs cites this paper.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:48.070568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:48.070568Z digest=sha256:f22737cfae3d6a22d4584b73d7440a033d4bc7f6a192a02c1fc3ed5c5f2e0c3c

Observation 7082b871-1105-4515-aeb9-7af8f2c182b1 · inbound

Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts cites this paper.

Mic Drop or Data Flop? Evaluating the Fitness for Purpose of AI Voice Interviewers for Data Collection within Quantitative & Qualitative Research Contexts Recent Advances in Speech Language Models: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T12:14:52.274618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:14:52.274618Z digest=sha256:2955070e0526225536f187e53ff63e422f129e755d7bac4568ad5e17a86048c2

Observation bb42877f-23b2-493b-9ed0-d2a644d69623 · inbound

Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models cites this paper.

Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models Recent Advances in Speech Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T11:25:29.124574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:25:29.124574Z digest=sha256:c654369635f7e9a3db8ee9ffa3016a8c9ef4eea1c5a4c06b87f041e60ec3cded

Observation 098c0c90-6348-4246-8634-a424d1abd924 · inbound

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs cites this paper.

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs Recent Advances in Speech Language Models: A Survey

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:11:43.169892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T18:07:01.257126Z digest=sha256:1128f882bd1e8eee691e04c947108d8cca963b585cbe1b126d94a29512569af2

Observation f4569ce4-5a72-4073-a5ad-9e98464af008 · inbound

TokenChain: A Discrete Speech Chain via Semantic Token Modeling cites this paper.

TokenChain: A Discrete Speech Chain via Semantic Token Modeling Recent Advances in Speech Language Models: A Survey

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:16:09.482195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T09:14:58.542628Z digest=sha256:1cecdd12879714c31d3207963d0f731d8aa63c7d80a9deac986147285806ed07

Observation 70a23ab9-d3d9-4e3b-a5a1-5a7e2e7923d0 · inbound

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models cites this paper.

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models Recent Advances in Speech Language Models: A Survey

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T07:46:03.599452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T07:43:23.913399Z digest=sha256:03367ce2431502ac75db3c9f0f5b3349464d48707723bb47afcf3e81a29dc121

Observation a159f64d-02d7-402d-9b29-77da7e49d8cd · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Recent Advances in Speech Language Models: A Survey

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.166907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:dc06a3878ac540b949249a5d200c939eb72fe938418228d78ca3bb35f7b62d0c

Observation f096af3e-4743-424b-9274-f767bba495c4 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:18.429430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T08:34:56.898815Z digest=sha256:a8b013e4574ce15174885e393e39eee1e7a9d83761658d533ad8cef87d68a48d

Observation 0134ecc2-d108-4441-9734-5a020a491d06 · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:44:07.790794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T10:43:27.176535Z digest=sha256:ebff5c7d3d3c12958ff9571189a9a47fd5d9af7a962c0330d498e80bfc2c6b76

Observation 3434c926-3507-48f9-9abf-af20e547d63a · inbound

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning cites this paper.

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T22:57:11.059962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:57:11.059962Z digest=sha256:2cb611d9b0349284027cc98f3cefb3c4b7de705af23682c2aab9bfb6aeb6c0c6

Observation 94d926ba-c1c5-4cd3-8ec3-617393cf85f5 · inbound

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction cites this paper.

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction Recent Advances in Speech Language Models: A Survey

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:59:50.641103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:59:22.416259Z digest=sha256:7791b5f7169bd74d17151c227699585360a8fc01baffda6d3ef6885bf0d43392

Observation c94f2c74-304e-43c0-a0a7-0e9a07fef51c · inbound

Audio-Mind: An Auditable Agentic Framework for Audio Understanding cites this paper.

Audio-Mind: An Auditable Agentic Framework for Audio Understanding Recent Advances in Speech Language Models: A Survey

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:13:17.598343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-29T10:03:54.653164Z digest=sha256:ab2d26f5f309afe22a82d1e7fc1c3ef01c69ad4f0debc86f822cd5a74441712a

Observation eadf26e4-da33-4572-955a-55ef64290ebb · inbound

MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond cites this paper.

MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond Recent Advances in Speech Language Models: A Survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T05:50:28.242656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:50:28.242656Z digest=sha256:73cef2f000fdd3e8d1489ad055174435bd7549a92e053396632de040c9bb6bf1