Pith. sign in

Paper Citation Record · LEDGER

LLaSM: Large Language and Speech Model

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2308.15930.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.15930 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:41:21.974244Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:37:30.579102Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2bcc33b7-1fb3-4b57-ad7c-82fe00b902f9 · inbound

Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models cites this paper.

Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models LLaSM: Large Language and Speech Model

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:57:28.748529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T18:57:28.666194Z digest=sha256:785860af2e06213c1922e8c0f46bf99f873af996983b00b1ca6508e0aceaab43

Observation 13641b6c-2ae1-4619-b66e-0e9b6d6c7119 · inbound

Survey of different Large Language Model Architectures: Trends, Benchmarks, and Challenges cites this paper.

Survey of different Large Language Model Architectures: Trends, Benchmarks, and Challenges LLaSM: Large Language and Speech Model

Reference 248

Resolution
unresolved
no resolver link, observed 2026-08-11T22:41:21.974244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:41:21.974244Z digest=sha256:35a8f8644069513e5c902456df0cf538a64c8106866dbcb7bab7e28349d6770b

Observation 4989efec-6c3f-4609-bfa4-00641759fab6 · inbound

Leveraging Chain of Thought towards Empathetic Spoken Dialogue without Corresponding Question-Answering Data cites this paper.

Leveraging Chain of Thought towards Empathetic Spoken Dialogue without Corresponding Question-Answering Data LLaSM: Large Language and Speech Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T18:52:42.762925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:52:42.762925Z digest=sha256:187b36d1196e8b62009f3a2064bef4ecf2388c042cc74f89b7d716496e003076

Observation 13bc8628-fe21-4399-95c0-379f3913b227 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey LLaSM: Large Language and Speech Model

Reference 200

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.367960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:1328d81e45c67aea6faef0cd34d3380e527f771d97e4727a648b50c6e9fcb7de

Observation 5dbef67c-0ebb-4356-b9c2-becaa5d10622 · inbound

On The Landscape of Spoken Language Models: A Comprehensive Survey cites this paper.

On The Landscape of Spoken Language Models: A Comprehensive Survey LLaSM: Large Language and Speech Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:45:07.915989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T20:44:57.476464Z digest=sha256:f54b007562137d04f76141345fb5af1c760b58124191967a817fae12a3c5fc4d

Observation f8e0976e-a24b-468f-853e-bd45b02588c7 · inbound

Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving cites this paper.

Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving LLaSM: Large Language and Speech Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.680423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:30:31.680423Z digest=sha256:920e628331c9b851878edff3aec85fd7432cfd11985f188bb16c56f57c6a54d0

Observation a5b6eea1-bc57-44ca-9cab-ee12f7df98f1 · inbound

EdgeLoRA: An Efficient Multi-Tenant LLM Serving System on Edge Devices cites this paper.

EdgeLoRA: An Efficient Multi-Tenant LLM Serving System on Edge Devices LLaSM: Large Language and Speech Model

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T20:59:01.315222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:59:01.315222Z digest=sha256:695e83b8182bf9afc1b8faceeacea757f3d1b5f6c8d87307e5c04931e2675ba4

Observation a5992823-3346-4bbb-9f37-b589bb962408 · inbound

Unlocking Speech Instruction Data Potential with Query Rewriting cites this paper.

Unlocking Speech Instruction Data Potential with Query Rewriting LLaSM: Large Language and Speech Model

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:34.713583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:21:34.713583Z digest=sha256:de76b8fb8730beea01e855eb38ff0ed4f8086ccc925fe5e416f09bca34b8f45b

Observation 3d004f2d-1099-4593-ae43-eff200e1d2a0 · inbound

MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks cites this paper.

MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks LLaSM: Large Language and Speech Model

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:41:59.657702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-19T02:41:52.996457Z digest=sha256:402c078d5885f77a5b63c22330ef0482edd9d8cbd6bfa94dd436ba9e59d27f82

Observation a119b833-5f6d-4ced-a0db-4db488a71afd · inbound

MeshLLM: Empowering Large Language Models to Progressively Understand and Generate 3D Mesh cites this paper.

MeshLLM: Empowering Large Language Models to Progressively Understand and Generate 3D Mesh LLaSM: Large Language and Speech Model

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T05:46:03.056127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:46:03.056127Z digest=sha256:74e1818797335e107a383f942b236dc1cdec89a93b12ee811554145421f09131

Observation a6b5faf1-ede3-4ecf-9000-e7f17fb39398 · inbound

Enhancing Speech Large Language Models through Reinforced Behavior Alignment cites this paper.

Enhancing Speech Large Language Models through Reinforced Behavior Alignment LLaSM: Large Language and Speech Model

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:24:23.481271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-21T22:23:52.392075Z digest=sha256:2be3ee64c0b5191fc353e5a5095fc5617a4cef35f70e5036d4654a53d480fba6

Observation bcd77563-fa62-4f61-8ae6-2bf308d2e71d · inbound

A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff cites this paper.

A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff LLaSM: Large Language and Speech Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T20:09:57.922788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:09:57.922788Z digest=sha256:884de56fdd862bb6c57feeeef35fa17dc6788d06a6c574eec9aa2247d420fb13

Observation b56a350d-6d71-441f-a73c-5883d90de0f8 · inbound

Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection cites this paper.

Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection LLaSM: Large Language and Speech Model

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:35:18.863265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T11:32:10.126062Z digest=sha256:1d999e0aad3e82b5a062b828f4df0f742b34d6d61f31c2a2c61c61a4da7aebdd

Observation c55f5e85-e411-4c6e-b2e9-10faaf5f9493 · inbound

Is Text All You Need? Text as a Universal Information Bottleneck for Speech LLMs cites this paper.

Is Text All You Need? Text as a Universal Information Bottleneck for Speech LLMs LLaSM: Large Language and Speech Model

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.580557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T16:26:32.594986Z digest=sha256:3fb3ee579827812c5e607d8c8f5ed3162e9fd885cc559f07109e76f90d3c0ddb

Observation 8a0cae49-a0c4-4d5b-b531-85e456ba98c8 · inbound

Continuous Audio Thinking for Large Audio Language Models cites this paper.

Continuous Audio Thinking for Large Audio Language Models LLaSM: Large Language and Speech Model

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:47:17.971674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T21:52:49.839901Z digest=sha256:78639ddb3524fac1cf719ff9b9c7874c8cc44d11596de556f6e9192808b2664f