Pith. sign in

Paper Citation Record · LEDGER

SUPERB: Speech processing Universal PERformance Benchmark

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2105.01051.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2105.01051 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:29:07.279462Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:53:55.761527Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5beaa773-cc2b-43ff-8413-2812414a0ebd · inbound

WavChat: A Survey of Spoken Dialogue Models cites this paper.

WavChat: A Survey of Spoken Dialogue Models SUPERB: Speech processing Universal PERformance Benchmark

Reference 236

Resolution
unresolved
no resolver link, observed 2026-08-12T20:13:58.234238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:13:58.234238Z digest=sha256:0552a10f4cd2781c7e03470890bd4d021340189cb5839cfa2b610e35d758c81b

Observation 7f08a44d-91ff-4e09-8a10-aea7e0730d3b · inbound

CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing cites this paper.

CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing SUPERB: Speech processing Universal PERformance Benchmark

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T21:30:10.997692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:30:10.997692Z digest=sha256:08a3ad2dac0ce3ae242c1b5a626fa67b4be8123b61957560cedc3e4bd967a2b8

Observation aa365929-7583-4da4-924a-e36fdb2f6d70 · inbound

Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis cites this paper.

Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis SUPERB: Speech processing Universal PERformance Benchmark

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T17:24:22.432903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:24:22.432903Z digest=sha256:a5e9d5f9a21e2b936c8e15c139ae0421554b0f9aa5955a62d0ad62ec4ae519c0

Observation 482a3266-c427-437a-8a90-2b17a9a8bf9d · inbound

Noise-Agnostic Multitask Whisper Training for Reducing False Alarm Errors in Call-for-Help Detection cites this paper.

Noise-Agnostic Multitask Whisper Training for Reducing False Alarm Errors in Call-for-Help Detection SUPERB: Speech processing Universal PERformance Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T18:06:18.040748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:06:18.040748Z digest=sha256:927ed6f2acaf039bfc286b6c9ed6a1d5628c2a31673e8c31dcb8c7478b388ccf

Observation 0d6ff449-43b6-4c1b-8bc7-29e7b25f4afd · inbound

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey cites this paper.

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey SUPERB: Speech processing Universal PERformance Benchmark

Reference 153

Resolution
unresolved
no resolver link, observed 2026-08-10T14:36:19.889484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:36:19.889484Z digest=sha256:0980d4d7656ac6dd07bd4a65e0210f2f0234b0df976bb6447ebb92147d8a1ed6

Observation 36a1a052-2da7-43d5-a94e-3fa582a15f61 · inbound

When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation cites this paper.

When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T19:16:37.865512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:16:37.865512Z digest=sha256:619e348c0200e6c01dfa8418080a682a893a656f193a15a8ac68c9277ec2cc5a

Observation 784e7f01-6faf-43e2-a3e3-16adc5fa465a · inbound

Comprehensive Layer-wise Analysis of SSL Models for Audio Deepfake Detection cites this paper.

Comprehensive Layer-wise Analysis of SSL Models for Audio Deepfake Detection SUPERB: Speech processing Universal PERformance Benchmark

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T04:35:12.644162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:35:12.644162Z digest=sha256:81012ef31e0c06e3a69beb491afbb1649f48f03c7ca5a364c29f786ac987420a

Observation 91d0c985-1dcc-433f-9059-aaa8909fc1e0 · inbound

Large Language Models based ASR Error Correction for Child Conversations cites this paper.

Large Language Models based ASR Error Correction for Child Conversations SUPERB: Speech processing Universal PERformance Benchmark

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:32.131218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:32.131218Z digest=sha256:66ac69fb83faffbe2f6d83e60ba72dc0ff91f9277d6feb6e9f99e60c76306d66

Observation c822dc6b-1ea6-4b31-a86a-eea374679c94 · inbound

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation cites this paper.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation SUPERB: Speech processing Universal PERformance Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.829556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.829556Z digest=sha256:10d9efaf878b9a4b8f2f884a7be5eb432d0b60cad8bc10e2750d292c7cbd3171

Observation 23de2174-1071-4659-8a82-a16a72715415 · inbound

StressTest: Can YOUR Speech LM Handle the Stress? cites this paper.

StressTest: Can YOUR Speech LM Handle the Stress? SUPERB: Speech processing Universal PERformance Benchmark

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T13:02:18.245043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T13:00:23.002962Z digest=sha256:e64d8073bd4c309e0f7213c2f3a7a311bd02194d653278d77182282a1ce63f64

Observation b7fa7fea-a8de-4bd5-b118-5815cf44f159 · inbound

Continual Speech Learning with Fused Speech Features cites this paper.

Continual Speech Learning with Fused Speech Features SUPERB: Speech processing Universal PERformance Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:59.069395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:59.069395Z digest=sha256:27b1d6b8f8ef7deea025bd53c325123aaa4c368e14c5638401626d800c459bd7

Observation d8de2a0e-6168-4d88-ab94-a6a625aaacae · inbound

Joint ASR and Speaker Role Tagging with Serialized Output Training cites this paper.

Joint ASR and Speaker Role Tagging with Serialized Output Training SUPERB: Speech processing Universal PERformance Benchmark

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:32:30.437423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:32:30.437423Z digest=sha256:a5b463a8c53c13aab5899b82f0605eef0cbf6499108e3a42cadc78df5ea608c3

Observation f4a97ade-ec12-49e6-9a14-2af222fe9663 · inbound

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases cites this paper.

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases SUPERB: Speech processing Universal PERformance Benchmark

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T20:58:31.169287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:58:31.169287Z digest=sha256:30a0d5da3c9449f03d9b6a0d176a637eed3fa7d0964b97aedd5593e4f3b696be

Observation 724aaa79-c9b0-4550-bf4b-3bd285639038 · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation SUPERB: Speech processing Universal PERformance Benchmark

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:59.287586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:59.287586Z digest=sha256:9fdaf1a611a4370b7c38aa251edbee8364bf51782118c80ad82cb01c07ccfe4f

Observation 07b7f3c5-28e0-411e-80db-09449c0dffa8 · inbound

Towards Robust Speech Recognition for Jamaican Patois Music Transcription cites this paper.

Towards Robust Speech Recognition for Jamaican Patois Music Transcription SUPERB: Speech processing Universal PERformance Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:27.022505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:27.022505Z digest=sha256:f1d0621758991467ce6c03cd5d4eeebfbdfe532fab02f44e5933afb9b6e6fe14

Observation 4660fb3d-fdf7-40e6-89ae-fe74d8854c43 · inbound

Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space cites this paper.

Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space SUPERB: Speech processing Universal PERformance Benchmark

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T11:07:20.126818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:07:20.126818Z digest=sha256:1965c8889e793f549a15d17e599e157ddae6ae3847e060a7405ff0743b396a91

Observation d35525ca-2a30-43d0-9329-ae571f573311 · inbound

Representing Speech Through Autoregressive Prediction of Cochlear Tokens cites this paper.

Representing Speech Through Autoregressive Prediction of Cochlear Tokens SUPERB: Speech processing Universal PERformance Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:54:58.546041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:54:58.546041Z digest=sha256:9ebece583c4bd5e6d1f1e16320815e8f8e27c0957ae088b1df4177285d2a6825

Observation b381d03e-da15-48d3-bd19-0adf626e619d · inbound

Deformation Driven Suction Cups: A Mechanics-Based Approach to Wearable Electronics cites this paper.

Deformation Driven Suction Cups: A Mechanics-Based Approach to Wearable Electronics SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T19:46:48.209823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:46:48.209823Z digest=sha256:a86ce5c8a1fb201bdeb860167021b3eb5582b117f16544cde88cd93988594eea

Observation 925214f4-052a-4fc9-9a96-b3d59c39ffe5 · inbound

AVEX: What Matters for Animal Vocalization Encoding cites this paper.

AVEX: What Matters for Animal Vocalization Encoding SUPERB: Speech processing Universal PERformance Benchmark

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T22:16:51.709697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T22:15:48.885339Z digest=sha256:1ce390d364b98a8577325c296b3115d1f932585ccd34f745b037d4f50c87473c

Observation 5f4738f6-a9f9-4743-b03f-76a552e53c59 · inbound

Multiple-Noise-Resilient Nonadiabatic Geometric Quantum Control of Solid-State Spins in Diamond cites this paper.

Multiple-Noise-Resilient Nonadiabatic Geometric Quantum Control of Solid-State Spins in Diamond SUPERB: Speech processing Universal PERformance Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T19:39:22.296512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:39:22.296512Z digest=sha256:53fd30243a7073bb74c4731a90277667c6d606aa264df35612e4b078bf71f3cf

Observation 995b18a4-2d47-4976-ae2b-68ece4f38635 · inbound

Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection cites this paper.

Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection SUPERB: Speech processing Universal PERformance Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:07.279462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:07.279462Z digest=sha256:f27cfe9e98c01d55bc3a888141f93ec89a0fdd7a56b0a0a2c6b32b5a8aef4b5b

Observation 5d9a2068-4691-4b0e-a7a7-c09ff5e2cc0c · inbound

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation cites this paper.

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation SUPERB: Speech processing Universal PERformance Benchmark

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T11:41:02.892997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:41:02.892997Z digest=sha256:3d1bb4c21fb96cab3e71d00f3b490ac92e971e184d1c6150e01a29d76c968304

Observation a697c031-e64b-4621-b4ed-b4c604146834 · inbound

Joint Learning using Mixture-of-Expert-Based Representation for Speech Enhancement and Robust Emotion Recognition cites this paper.

Joint Learning using Mixture-of-Expert-Based Representation for Speech Enhancement and Robust Emotion Recognition SUPERB: Speech processing Universal PERformance Benchmark

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T18:11:42.900862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T18:07:34.965356Z digest=sha256:a484fe77ecfd79c284cb130d1519c400efcba72a4f3f355eb5bb1cb5baabb304

Observation f984d490-1f1f-4575-8b54-3c91fb536ed3 · inbound

ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining cites this paper.

ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining SUPERB: Speech processing Universal PERformance Benchmark

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-13T16:09:50.207908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:09:50.207908Z digest=sha256:5dee9bf5bda2a368e5ffdd89ce9838f9ab5eb9c4ee3bf8fcfd00f0630d81c629

Observation 9102b529-8e2c-43ce-90e9-6f0d6c407315 · inbound

ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals cites this paper.

ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals SUPERB: Speech processing Universal PERformance Benchmark

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:25:54.281562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:30:30.431777Z digest=sha256:45c5c91f9f29afcfc48a67360ea7e43b247208891eac3fb43181828eaf6e90c3

Observation 1d174128-92b1-4b5d-90be-656fa86a94bc · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes SUPERB: Speech processing Universal PERformance Benchmark

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:33.505238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:5969908155abf3c7e0cd4fea16e7aa214819be0f3946b88b66c6c767e4b0d8c8

Observation 2e8d27b2-891d-4f3f-bca1-e5eff96c860f · inbound

Multi-layer attentive probing improves transfer of audio representations for bioacoustics cites this paper.

Multi-layer attentive probing improves transfer of audio representations for bioacoustics SUPERB: Speech processing Universal PERformance Benchmark

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:28.421913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T04:11:17.994180Z digest=sha256:5f780ffa757bebf8ae7fa727027d1ffeeb850f69f177d4c08103108bdf359d57

Observation 4bcd3ac3-e213-4e36-8687-991e32b4b297 · inbound

AudioMosaic: Contrastive Masked Audio Representation Learning cites this paper.

AudioMosaic: Contrastive Masked Audio Representation Learning SUPERB: Speech processing Universal PERformance Benchmark

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T01:53:28.854680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T01:52:01.164694Z digest=sha256:ed29f4c16845e33c736cda693d2638cd5fef20c098035e972baee78c15ab8624

Observation 458ee8da-b048-4de5-9728-65bcad93a99e · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages SUPERB: Speech processing Universal PERformance Benchmark

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.758423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:416481bedf20926a38eeea048bd202ec3010277b75878a30e937372911dcccf1

Observation 26975453-eb90-4d85-87ab-71f476c7aee6 · inbound

A Unified and Reproducible Experimentation Framework for Speech Understanding cites this paper.

A Unified and Reproducible Experimentation Framework for Speech Understanding SUPERB: Speech processing Universal PERformance Benchmark

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:16:12.193979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:885517e4bb1af6d314aaff43761cd7e113417ab1d4f294642fbbfa04b43fa74e

Observation 82c14792-db0e-40de-9be4-3c637d99cb18 · inbound

MOSS-Audio Technical Report cites this paper.

MOSS-Audio Technical Report SUPERB: Speech processing Universal PERformance Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T00:56:25.070239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T13:05:29.813707Z digest=sha256:a00842adf96f9726cb847a10f78222277800678c8cb5bdfc0910e2f38337a334

Observation 3342e72b-077b-4df4-ae0a-95680c353790 · inbound

SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails cites this paper.

SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:47:19.562187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T21:19:56.932689Z digest=sha256:00e7d79362e74d77c7d69b9c5ffed7c0fec2a292add52e1f3c14cad31ce848c5

Observation 2906d0af-dab0-419d-b7fb-6f3e014dc63a · inbound

S-JEPA : Soft Clustering Anchors for Self-Supervised Speech Representation Learning cites this paper.

S-JEPA : Soft Clustering Anchors for Self-Supervised Speech Representation Learning SUPERB: Speech processing Universal PERformance Benchmark

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:19:24.216627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T19:46:47.653439Z digest=sha256:b2b0b495c4e4639dc3198ceb68b2bf95b463ba905939579b4cf58ecb74fb7a4e

Observation fdcd89da-18a7-46e6-8099-6ac99c19f2e2 · inbound

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack cites this paper.

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack SUPERB: Speech processing Universal PERformance Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.328046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T14:19:59.573391Z digest=sha256:c83e72ce6a60385a89d88bb71b76a816533c25e8c66fc95007e521974cf41081

Observation c6760762-9e33-4456-8c11-2dda8ff85c86 · inbound

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack cites this paper.

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack SUPERB: Speech processing Universal PERformance Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:33:54.408623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T04:44:52.537405Z digest=sha256:7ae1ce8c02e0d80d7714a35f9f20bf6db9303dd8082e0cf6428ba869641ee095

Observation e04976fb-eddd-48d7-817e-3474b7469bb9 · inbound

MSU-Bench: Towards Speaker-Centric Understanding in Conversational Multi-Speaker Scenarios cites this paper.

MSU-Bench: Towards Speaker-Centric Understanding in Conversational Multi-Speaker Scenarios SUPERB: Speech processing Universal PERformance Benchmark

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T11:49:50.798715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T07:36:34.307652Z digest=sha256:5485850a78acda37a43369de6ac9476dad2b281f3106bd698225ad0a1c18d1de

Observation cd021a1a-d2c2-4ef3-a5e2-6faea7688acd · inbound

End-to-End Voice Intent Recognition for Spontaneous Human-Drone Interaction with Naive Users cites this paper.

End-to-End Voice Intent Recognition for Spontaneous Human-Drone Interaction with Naive Users SUPERB: Speech processing Universal PERformance Benchmark

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:29:38.212507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T13:30:12.101045Z digest=sha256:d20f31a10da67332c7e52757ccef010dbb216694117d1098fb71e282f1961010

Observation 15289bd6-f251-4997-93be-5a9d1e8d1798 · inbound

SIGMA: Saliency-Guided Sparse Mask Attacks for Speech Emotion Recognition cites this paper.

SIGMA: Saliency-Guided Sparse Mask Attacks for Speech Emotion Recognition SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:24:57.677906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T04:44:59.279562Z digest=sha256:8e0109e7d8998947b60ed28c150bf522dc4a69674cbe148aae71c97b68036578

Observation f5bf67c4-e13c-4e07-ba87-265be4b59ea9 · inbound

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning cites this paper.

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning SUPERB: Speech processing Universal PERformance Benchmark

Reference 111

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T05:56:39.882231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-02T05:52:55.818877Z digest=sha256:45752c7760f193594feef3f3eab0b8ed2b5c6480e2f2d02f13954ffa63eee741

Observation 7963645b-9cd7-4068-b5d1-df5bcd427f8c · inbound

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models cites this paper.

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models SUPERB: Speech processing Universal PERformance Benchmark

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T14:53:55.763035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-07T14:53:07.512543Z digest=sha256:f9337ecc4a263f892d84691eed447056955ebf3c87d37b495dd03d683c8b31d1

Observation b307710e-4c92-4ad4-8562-05cadfad0ae0 · inbound

Multi-Phonation Graph Learning with Self-Supervised Speech Embeddings for ALS Detection and Progression Prediction cites this paper.

Multi-Phonation Graph Learning with Self-Supervised Speech Embeddings for ALS Detection and Progression Prediction SUPERB: Speech processing Universal PERformance Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:44.725464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:44.725464Z digest=sha256:3fd1ec5da08af6a522b88ae9a933869f86d6f45f7d27fb96f61a6611c674cfc9

Observation 00876789-57bd-4e7b-8726-22dd09ef36fa · inbound

Speaker Verification Under Real Classroom Conditions for English Speech cites this paper.

Speaker Verification Under Real Classroom Conditions for English Speech SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:52:34.565904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:52:34.565904Z digest=sha256:cdaec07621ad0f3cc817f0734e6b93b78344380cb608b07047712ace1a47dd60

Observation b503bbd8-287a-4316-95cf-01a9b0185bb6 · inbound

Speaker Verification Under Real Classroom Conditions for English Speech cites this paper.

Speaker Verification Under Real Classroom Conditions for English Speech SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T14:50:30.359241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:50:30.359241Z digest=sha256:5fc9ec8d67645c63f1b492ae25e3f66579d42c478c6dad08130ee5d7eb680d96