Pith. sign in

Paper Citation Record · LEDGER

Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2411.05361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.05361 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:05:27.345407Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T05:51:09.778662Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 11c44623-f80e-45ca-b817-daca864c1833 · inbound

VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models cites this paper.

VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:24:18.189539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:24:18.189539Z digest=sha256:e12c312ed9a99e332a760c0056e6d7314d4f303cceadf286e2481b017b734f0b

Observation effb8c2d-a8ee-4f4b-ad31-ee6a915cf8df · inbound

A Preliminary Exploration with GPT-4o Voice Mode cites this paper.

A Preliminary Exploration with GPT-4o Voice Mode Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T20:02:51.333001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:02:51.333001Z digest=sha256:43a16e90896d7e64eb57717cb9e98ad66485666485edf7c858463b0cd5a1e6d0

Observation 4de67696-d25a-4949-8830-bf817564285e · inbound

BLAB: Brutally Long Audio Bench cites this paper.

BLAB: Brutally Long Audio Bench Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:05:27.345407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:05:27.345407Z digest=sha256:7e820f747e3cdbff6763acc90b3c896d4dc017cf5b38ed87c5bd799a19898927

Observation c1daf752-92ab-424b-9087-ce9f84192261 · inbound

Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples cites this paper.

Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:05.282635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:05.282635Z digest=sha256:0989b1a59e5a5f6d94bfe4a163de6bb26dc284325f51109b0fc45b33dae96a1d

Observation 6e3323cd-c372-4bd9-8862-80d0d81e1159 · inbound

Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models cites this paper.

Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:28.620577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:28.620577Z digest=sha256:154ae0c0ba2865febac4f4e72ae48f8ca8529212c637eaeca2ba7d3b329cf3b4

Observation 5aad3ced-54b1-4a76-8309-61beb49207e2 · inbound

SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant cites this paper.

SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:23.197926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:23.197926Z digest=sha256:ccb73e60c57525a1bb3be8335432f024640a38a87d89e762921694d853d9021d

Observation 05b77574-85b5-4e7f-8c00-7224fed46ca4 · inbound

Multi-Distillation from Speech and Music Representation Models cites this paper.

Multi-Distillation from Speech and Music Representation Models Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:35.914407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:35.914407Z digest=sha256:aa7902d4cfd1c97f385c56e3678ca54f8bbf77db1114187d5626c0a05ba06349

Observation 2020553c-8add-41be-a6f6-88fe1e2b4c62 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.611415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.611415Z digest=sha256:c8dadee7823a51de3dd69e2b3a959e55d30b1b52ad5b1253615b953d4443130e

Observation 2e7e6682-6c59-46bc-a742-1b2da70970a0 · inbound

AV-EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Omni-modal LLMS with Audio-visual Cues cites this paper.

AV-EMO-Reasoning: Benchmarking Emotional Reasoning Capabilities in Omni-modal LLMS with Audio-visual Cues Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:06:45.452664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:06:45.452664Z digest=sha256:d857999508fa40297fd232d0feaa6dc35886fca06590bf4ccaeec9fedeff243e

Observation 2c4d4b4c-7559-43a4-8d3d-dd5bbd1163e0 · inbound

ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining cites this paper.

ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T16:09:50.207908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:09:50.207908Z digest=sha256:71c95de78b7e14be2ed0156baf35585a96918f84496f397af387d250ad207ebb

Observation 93844cbf-f821-4643-85a8-7b8266add351 · inbound

Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment cites this paper.

Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:09.779986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T05:50:05.920842Z digest=sha256:6234eccf1fa23da47e116c16ee302a306bc739cb91abcb226d52b20d6ecefd50