Pith. sign in

Paper Citation Record · LEDGER

State-Space Large Audio Language Models

As of 13 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2411.15685.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15685 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:04:44.626476Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:23:41.773642Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:23:47.775279Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 272374f9-250f-4829-9815-c453ed1d6647 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

State-Space Large Audio Language Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.450611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.450611Z digest=sha256:079222df082cce5807dd02a52d28b77ce550f794ea859a3abc8a5c81f2c476f2

Observation 1d5f0bb7-7496-45b7-be5a-7523f573eaad · outbound

This paper cites Explor- ing the limits of transfer learning with a unified text-to-text transformer,.

State-Space Large Audio Language Models Explor- ing the limits of transfer learning with a unified text-to-text transformer,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.456158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.456158Z digest=sha256:dd13eab8ece7bfffbbe69818a94b62ed6fade7c779bd7d983c32c94290130c1a

Observation 864a61d4-b592-49bb-b4b2-2299d9643a2c · outbound

This paper cites Language Models are Few-Shot Learners.

State-Space Large Audio Language Models Language Models are Few-Shot Learners

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.460807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.460807Z digest=sha256:f0d3eba6b71b19fd704acf61f384941be1073a656daf0c08e6189df39d02cff6

Observation 9f87d954-3b97-4147-9c20-6e53c97a56b6 · outbound

This paper cites Training language models to follow instructions with human feedback,.

State-Space Large Audio Language Models Training language models to follow instructions with human feedback,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.319699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.465663Z digest=sha256:591eaa5c3fd00b0a51d306db481d9ca7902931d539249e6056c5097b51c27ed1

Observation adb16fc8-8405-42da-84ad-213454a4f5f2 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

State-Space Large Audio Language Models OPT: Open Pre-trained Transformer Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.471082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.471082Z digest=sha256:a956d6d7d85050c1cac872d2f5163295364de634f005147b18f5d601e8de0f3c

Observation 7c472a50-9126-40e1-9686-67e1a22ab773 · outbound

This paper cites A Survey of Large Language Models.

State-Space Large Audio Language Models A Survey of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.476229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.476229Z digest=sha256:bd10563210034928cf479172902c16e2241997503f294f891f18c4dec0fa7d67

Observation 1bf8d7be-1b4c-40ca-9bac-75bff6459c25 · outbound

This paper cites Pengi: An audio language model for audio tasks,.

State-Space Large Audio Language Models Pengi: An audio language model for audio tasks,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.303731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.481729Z digest=sha256:808bffca14854180104ec2a781ed0516548d045fd5148977dc9892bb77fb28fb

Observation 575f365d-278e-45c3-bf71-ebeef7723e4d · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

State-Space Large Audio Language Models SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.486403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.486403Z digest=sha256:288c0b2662229f4d313b0da9908c86b4fe47a69e8787d4ec7190150ef5b04723

Observation 2782f4f9-4361-4c18-841b-6148e704c862 · outbound

This paper cites Listen, Think, and Understand.

State-Space Large Audio Language Models Listen, Think, and Understand

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.491203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.491203Z digest=sha256:077b306d6833560d23d198d14a16770267731649e3d1149d09fee0c5c49a8a81

Observation ab25b993-9f24-45eb-8ad2-071d055f12f7 · outbound

This paper cites Joint audio and speech understanding,.

State-Space Large Audio Language Models Joint audio and speech understanding,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.287787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.496219Z digest=sha256:142c3186971c6eefe3ebfc985adf6b445a8ecf845ad10177017d799238131267

Observation 63878ce0-6cab-4f08-8cd3-b2f90b15d257 · outbound

This paper cites Audiogpt: Understanding and generating speech, music, sound, and talking head,.

State-Space Large Audio Language Models Audiogpt: Understanding and generating speech, music, sound, and talking head,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.272449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.500735Z digest=sha256:fdebcbbed9dd2b2c2ed56a5fd5ea01bef77818e5c598a2a0743caf82938ddec8

Observation c4a2494b-f4ab-415f-a79a-70998b6a3c6c · outbound

This paper cites GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities.

State-Space Large Audio Language Models GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.505485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.505485Z digest=sha256:c426bb42036e95c84bd7f38bf32ecef62dbf48362659cf9578f1b2f3cfd30935

Observation 3be49039-91f5-44ff-a143-203dfa4262f9 · outbound

This paper cites Attention is all you need,.

State-Space Large Audio Language Models Attention is all you need,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.510899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.510899Z digest=sha256:af46fadc863c663c673395335f33ab69f4fe6f11befe14fcaba4739bb4bd76b5

Observation fe57f6c0-6794-462a-a70f-84cefef2c630 · outbound

This paper cites Linformer: Self-Attention with Linear Complexity.

State-Space Large Audio Language Models Linformer: Self-Attention with Linear Complexity

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.515472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.515472Z digest=sha256:36719bccebb305660d44c25580d36d0c3e81e532cf2ab2499bcb45f5f9c2b6c1

Observation da4603c1-f7ac-42b9-bfd3-90272c465648 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

State-Space Large Audio Language Models Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.520214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.520214Z digest=sha256:301e6dae2e7e85b29a7066517265639e1e6a1c0653bdeb2a3eada65cf4ac6499

Observation 12a3a627-b817-4887-bc50-3e3cf5fb3fa1 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

State-Space Large Audio Language Models Efficiently Modeling Long Sequences with Structured State Spaces

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.524512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.524512Z digest=sha256:a1d2e3a8d047799ca2a3d8a82dee9da13e51ba44be0f8e1b0ba1f5b7e1df1794

Observation 15d6765f-ea8b-4f66-9863-202b4e879905 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

State-Space Large Audio Language Models Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.529498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.529498Z digest=sha256:0d0442b2b9dc0f15224128f83e9ba488cbd2f8e2726037932462e87f927b1297

Observation f6f88bef-02c5-4700-bcdb-670359c23866 · outbound

This paper cites VMamba: Visual State Space Model.

State-Space Large Audio Language Models VMamba: Visual State Space Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.534507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.534507Z digest=sha256:85a4f48c90b19ed0205d6d70056af46f831985ae30c861668640cdd293d32c9c

Observation 19792666-4b0c-4e0b-8a24-13f6be1bcc6b · outbound

This paper cites DASS: Distilled Audio State Space Models Are Stronger and More Duration-Scalable Learners.

State-Space Large Audio Language Models DASS: Distilled Audio State Space Models Are Stronger and More Duration-Scalable Learners

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.539022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.539022Z digest=sha256:3172eaa6588d3bdc7372833a7ee3ccb5cf585367b5e9af19dd7b32bf5c054ea7

Observation f32d957d-4fcd-49e5-9f4a-290ad292baa6 · outbound

This paper cites Long Range Language Modeling via Gated State Spaces.

State-Space Large Audio Language Models Long Range Language Modeling via Gated State Spaces

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.543697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.543697Z digest=sha256:8f8dcdfca2cda14b7d11fe2d75c9b45a4376b1bb5b51a6cc8f1f25acdbccaf2a

Observation 9d9d7d63-f346-40e3-88e0-b3be26b8a6f8 · outbound

This paper cites Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model.

State-Space Large Audio Language Models Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.548796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.548796Z digest=sha256:c71321925ca8bbd238d378730e18f02aa475f3a57cdeb30efbf4f3b0741edf76

Observation 22042d7d-d742-40b4-aa5e-a3139f004b8e · outbound

This paper cites Audio mamba: Bidirectional state space model for audio representation learning,.

State-Space Large Audio Language Models Audio mamba: Bidirectional state space model for audio representation learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.236188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.553574Z digest=sha256:01014a3bc38c22eeddf76c5f9094e51439bf9ba08cf8c3fed3b4e77289771e41

Observation 145a8cfe-1e78-430d-adc2-0ef047f59f1e · outbound

This paper cites Audio Mamba: Pretrained Audio State Space Model For Audio Tagging.

State-Space Large Audio Language Models Audio Mamba: Pretrained Audio State Space Model For Audio Tagging

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.558184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.558184Z digest=sha256:24d57d5685f8c9744653e1c98cb6ed7e30f0e67bcb953190560293ed2ebe4867

Observation eff6762c-db0a-441f-bf51-e294705e2604 · outbound

This paper cites SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model.

State-Space Large Audio Language Models SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.563143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.563143Z digest=sha256:a19361aa5e4416c44feb5992cacf3644254e9a323309bb4285cc799719024e90

Observation dc80da73-6746-44d4-87b5-ac17c579e499 · outbound

This paper cites Instruction tuning for large language models: A survey,.

State-Space Large Audio Language Models Instruction tuning for large language models: A survey,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.567928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.567928Z digest=sha256:e8dfb2fac3494ffc1069d2b31361302ee327dea4871e50158fb0e5cf4a225327

Observation 432ba048-6687-4f1c-8b2c-c24609d87cd2 · outbound

This paper cites Hts-at: A hierarchical token-semantic audio transformer for sound classification and detection,.

State-Space Large Audio Language Models Hts-at: A hierarchical token-semantic audio transformer for sound classification and detection,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.219679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.572428Z digest=sha256:8d8c0fe8ae1c90d741063eb2b60680b14dfada02b87e777af613515b7d1f78b9

Observation 623af4e4-d30b-4153-8d37-2fd7dc0db7d5 · outbound

This paper cites Language models are unsupervised multitask learners,.

State-Space Large Audio Language Models Language models are unsupervised multitask learners,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.576825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.576825Z digest=sha256:2d82e5fe55e5e47d1e96ce02c1c1e4a99eb2d60f6f9b254ca614062c4a480b60

Observation 75116d44-7e9a-4b3b-862a-5c7cf73ea329 · outbound

This paper cites Robust speech recognition via large- scale weak supervision,.

State-Space Large Audio Language Models Robust speech recognition via large- scale weak supervision,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.581407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.581407Z digest=sha256:f6f302698d2fcb551df057324858b974db46288ad573af71b05371187d3080eb

Observation 7c20f9bb-4d80-486e-b72f-37a54b7a28be · outbound

This paper cites Conv-tasnet: Surpassing ideal time– frequency magnitude masking for speech separation,.

State-Space Large Audio Language Models Conv-tasnet: Surpassing ideal time– frequency magnitude masking for speech separation,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.586120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.586120Z digest=sha256:bb08012587c86b65aafec1add1d83517ec4928065920ae3cccaf7b041785a0d8

Observation c82c3ff4-178a-45ab-a3cd-dad617685b13 · outbound

This paper cites BEATs: Audio Pre-Training with Acoustic Tokenizers.

State-Space Large Audio Language Models BEATs: Audio Pre-Training with Acoustic Tokenizers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.590645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.590645Z digest=sha256:1d0b9176735c5a4c6eb265c5d5689533155c1a3140d4a90c63829f7b89576826

Observation d096bba1-b40e-4e6f-9f69-2b1d370c861e · outbound

This paper cites AST: Audio Spectrogram Transformer.

State-Space Large Audio Language Models AST: Audio Spectrogram Transformer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.595070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.595070Z digest=sha256:c914e64e2145e35abf0cfbe7023f3df5f8b8d93e3eb21e0f83e574dcfe0cf2d0

Observation f6c6ccf2-faea-4832-b5e1-2754b0ed7c20 · outbound

This paper cites Au- dioclip: Extending clip to image, text and audio,.

State-Space Large Audio Language Models Au- dioclip: Extending clip to image, text and audio,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.174553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.599726Z digest=sha256:17eeb2a49e5ace4023c6a2f945cb999d2d6d55f12ef74d68f5e947f7f90f5111

Observation 2b3af54f-4da7-43b9-a855-324b50c73b6f · outbound

This paper cites Large-scale contrastive language- audio pretraining with feature fusion and keyword-to-caption augmen- tation,.

State-Space Large Audio Language Models Large-scale contrastive language- audio pretraining with feature fusion and keyword-to-caption augmen- tation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.603950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.603950Z digest=sha256:c3ba4d46bd7c148f74f043c29f4cefecf03f325eea3a0472d418a9e91079372b

Observation 8ea84b7d-1e8a-43bc-acf4-2c35414e883b · outbound

This paper cites Clap learning audio concepts from natural language supervision,.

State-Space Large Audio Language Models Clap learning audio concepts from natural language supervision,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.608462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.608462Z digest=sha256:11b877431cf469e035d77167bed670e2765ffe7a382ad681f7358559089603a4

Observation 9cfac192-3e15-495a-aa66-26643f4f1669 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

State-Space Large Audio Language Models Audio set: An ontology and human-labeled dataset for audio events,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.612706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.612706Z digest=sha256:0513ff9a2b79cc1eb60d3fdbddbdcc3d59304f3eae13def38b52b33d1f635182

Observation 16d6cc27-6e1d-49a1-959d-c279a508e147 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

State-Space Large Audio Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.617190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.617190Z digest=sha256:3367f8ecba4c56a29d371c3eec843a7236cfcc84c746a215533d5bc60b258205

Observation 0aa2be76-7806-450d-bf01-bce034983a57 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,.

State-Space Large Audio Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:04:45.127438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:04:44.622107Z digest=sha256:1b4a3f75ded344964f3afb0c678452c0607387d618217885389dc9ebb4b3caeb

Observation fca4f650-c47d-414d-8871-93ec79217e4d · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

State-Space Large Audio Language Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T14:04:44.626476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:04:44.626476Z digest=sha256:ab4d4d4943abfa6c7ae88652a3e9e526997ae91282017b18cbe4a18b951df580

Pith citing papers

Observation 32074b85-7582-4e86-b335-79e93b4ede4e · inbound

Towards Reliable Large Audio Language Model cites this paper.

Towards Reliable Large Audio Language Model State-Space Large Audio Language Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:23:47.897076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:23:41.773642Z digest=sha256:a926fae02eca888156b0cadc9e25f117fa5a6e56188c7bb0e71652064c9108bf