Pith. sign in

Paper Citation Record · LEDGER

Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2501.07246.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.07246 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:55:54.386404Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T07:57:44.602800Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5380246a-a48d-41ff-8600-a81a1a2b2424 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.165169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:a89d04e65337bc4d3b94a8301b69ddd031f443f90f07a05a2820e656a8faa7ea

Observation 13a23776-4513-41db-ab68-e23694a70cd5 · inbound

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks cites this paper.

AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:54.386404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:54.386404Z digest=sha256:1c828326ae50fb5555e202a4e0375ab83a819e754c6be37689a89fd6437770e0

Observation aeccd722-b14f-4a26-9f9d-afbb627ffbb9 · inbound

MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing cites this paper.

MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T19:12:03.971484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:12:03.971484Z digest=sha256:3891ab26debed4c35346c7a18092364b779e05347afef9fd2f836fded0d7a6f5

Observation a50a7049-c8dd-4a2b-a12a-59b3bb6e311c · inbound

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models cites this paper.

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:42:44.780651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T03:42:44.523919Z digest=sha256:b37c685f323b1ab722c6fe45288fc31bc08be6263ba4cf7473e86b9d0e702cc7

Observation 55f84e21-e92a-4fcf-9637-01eb02f816c3 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 165

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:59.038715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:59.038715Z digest=sha256:4e4b11e3728c23cda7b85c93d439e76fe1c8f25cad4ccb1b6f8d735cf7e1c99a

Observation cde40ad5-62e2-4b8c-9419-fc4ae9ffae6f · inbound

WoW-Bench: Evaluating Fine-Grained Acoustic Perception in Audio-Language Models via Marine Mammal Vocalizations cites this paper.

WoW-Bench: Evaluating Fine-Grained Acoustic Perception in Audio-Language Models via Marine Mammal Vocalizations Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T14:43:29.201863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:43:29.201863Z digest=sha256:47026a31f3ae133b98c27f725e1b1c0073951247b26718cd2289b54b46e5be56

Observation 1719fbd3-f860-4a3a-8498-d9b655d6e714 · inbound

CapTalk: Unified Voice Design for Single-Utterance and Dialogue Speech Generation cites this paper.

CapTalk: Unified Voice Design for Single-Utterance and Dialogue Speech Generation Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:16:00.091350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:15:28.918204Z digest=sha256:897b02ab80e71348f671340d9e3d83a5a92a38737c10616edc70adebb59d58b5

Observation 573a377e-7948-4806-93bf-82d7c680b46e · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.284460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T14:10:03.707886Z digest=sha256:046e955442598d5196d7a5d0526615dc3d3a8147cb0fcac8a042ec035e40948c

Observation c52941fc-617e-4149-83df-0a3feeb17730 · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T21:18:46.566338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:18:46.566338Z digest=sha256:4a98ba00f84f3411e0448b6e408fbaa518ef2ea572415805259520f02ad60da0

Observation 175ca406-d890-4c58-a09e-05a6b7acb062 · inbound

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models cites this paper.

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:41:26.451182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T14:29:18.348031Z digest=sha256:0dcbdf3a730f6ad910dab717194776eb4344f9b51f656423965d497541cdfba7

Observation 76084fc5-e8de-4826-8f2d-079b026f4194 · inbound

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models cites this paper.

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:51:09.757516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T16:59:25.973854Z digest=sha256:7b070c307aec7bb565bd2e83148c4bbb716ae1f479cc7cbe9a8ffc89bba9abbc

Observation de39710d-62e0-40fb-b156-2400a34e14d7 · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.059663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:a13a53b96260a866733312d62a50df68dd7c7cf2446584b9d5ba4d814269d2be

Observation 9b2aa08d-227c-4b95-bd7a-197abbc39417 · inbound

Learning When to Think While Listening in Large Audio-Language Models cites this paper.

Learning When to Think While Listening in Large Audio-Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:43:50.751241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T18:37:34.409802Z digest=sha256:b50234db29528e1695fd630d06865a0e9ba5a3361a2f6c036f14db5fbb71ba4a

Observation ab455650-dc21-44a6-b3ac-a30ed14f413d · inbound

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track cites this paper.

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:07:21.325490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T21:02:19.300441Z digest=sha256:ef5282729fb6b1644b47506980ca2b859efc8cced0dbee65774eea88a1b46ab7

Observation 3d9f9e8d-5f9e-4426-a25f-0c2a226b3bf4 · inbound

TinyGiantALM: A Compact Audio-Language Model for Intent-Aware Reasoning under Resource Constraints cites this paper.

TinyGiantALM: A Compact Audio-Language Model for Intent-Aware Reasoning under Resource Constraints Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:17:29.817376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T18:18:41.816342Z digest=sha256:102a3298da1fe47dddea6765b79156496069e5f8db941f16accae97bc8dbd527

Observation 2017c47c-82d9-43f0-a6f4-89138b205358 · inbound

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains cites this paper.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:57:44.604366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:5c6935565b688d9dc255cd4784edc7e0438d029a883d17f4764a0a33022715aa

Observation f3641d3d-397c-4875-ac20-768f76134745 · inbound

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models cites this paper.

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T11:20:17.838169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:20:17.838169Z digest=sha256:1e97a5b0d0a1f6b9921f6877a476582e3b544b30b86dbbb80945c006e02d1edc

Observation 4ed1485f-57ef-4921-baab-20702ce030ac · inbound

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment cites this paper.

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-01T07:12:17.582095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:12:17.582095Z digest=sha256:5a6aeb15f4ea1a90734d887f2113fa30c036a9c42064a4fb957d542f9fedc077

Observation d596c087-c165-4388-85fd-298506c7bd68 · inbound

Weak-to-Strong On-Policy Distillation cites this paper.

Weak-to-Strong On-Policy Distillation Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-01T00:26:29.186928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:26:29.186928Z digest=sha256:d49c72219a8517ce529304365c243b36ac6a65931b5dc98e38865000b9e6f98e