Pith. sign in

Paper Citation Record · LEDGER

DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2409.20007.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.20007 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:05.211406Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T20:45:07.957840Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e7f87eb2-8520-4742-a5e5-6c7565ae4172 · inbound

On The Landscape of Spoken Language Models: A Comprehensive Survey cites this paper.

On The Landscape of Spoken Language Models: A Comprehensive Survey DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:45:07.965137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T20:44:57.476464Z digest=sha256:6bece86fb25c256a204fd747bd102a64cef117d4544956fc0a85d63bc8f3f19b

Observation e32b1938-11de-4f94-be17-2b46538f74ee · inbound

Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples cites this paper.

Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:05.211406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:05.211406Z digest=sha256:b3eaf397729e5896b5d364834b5b29b113c904503c262800ab6f7358eb5b6b99

Observation 0279a263-cbd1-4f53-baf0-7984b7ea1c83 · inbound

Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models cites this paper.

Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:29.352657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:29.352657Z digest=sha256:96beb99257763709ad5a6471398367174dd0b381d78b81d972da9f1425270cf0

Observation 0593cf21-cea0-43b3-aead-54ce8edbed39 · inbound

Speech-IFEval: Evaluating Instruction-Following and Quantifying Catastrophic Forgetting in Speech-Aware Language Models cites this paper.

Speech-IFEval: Evaluating Instruction-Following and Quantifying Catastrophic Forgetting in Speech-Aware Language Models DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:33.707065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:33.707065Z digest=sha256:34e613f4b4e2b5fd5536b8d2a157d580905a409be4e896561ff47a82042c2472

Observation 60fba7c7-1529-45b8-83a6-bd55603d68ff · inbound

Towards Reliable Large Audio Language Model cites this paper.

Towards Reliable Large Audio Language Model DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:44.535221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:44.535221Z digest=sha256:1775b57af0595d1c8dfd6317def4fc51bd967f57129e31906513cc542dc2b066

Observation b06639b3-16e3-4229-8b83-9b470c429822 · inbound

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering cites this paper.

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T19:38:24.048548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T19:38:24.048548Z digest=sha256:2702420cee3ac0cafcafb97d90c80598b128f2884664682b2c0b42fc41bd9df6

Observation 784b4ec4-77fc-4888-923c-2e56cdc2d81e · inbound

Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs cites this paper.

Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:51:17.771428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T21:49:21.785096Z digest=sha256:b7befd587fac69a4c716436f52a1cf5f041929cdcfd6326721aadae42bab844d

Observation e3bb098a-53bf-4cb5-8000-ce0788cb06f2 · inbound

Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities cites this paper.

Can Large Audio Language Models Ignore Multilingual Distractors? An Evaluation of Their Selective Auditory Attention Capabilities DeSTA2: Developing Instruction-Following Speech Language Model Without Speech Instruction-Tuning Data

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T23:17:57.448378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T23:17:08.124240Z digest=sha256:42d492bb8d4740a416ecb8932216733016d53e7458dfca9e190c58411c72f2e7