Pith. sign in

Paper Citation Record · LEDGER

FSD50K: An Open Dataset of Human-Labeled Sound Events

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2010.00475.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.00475 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:09.913604Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:37:36.034511Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 67de3525-a519-43b2-88af-a073a98efd70 · inbound

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds cites this paper.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.913604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.913604Z digest=sha256:6af201b0abce6b21551ca1f4da8489825325bb8de8cfd1927b9708d0045fc913

Observation 064eb756-a393-4b7e-a412-0dad2ce5e181 · inbound

From Large-scale Audio Tagging to Real-Time Explainable Emergency Vehicle Sirens Detection cites this paper.

From Large-scale Audio Tagging to Real-Time Explainable Emergency Vehicle Sirens Detection FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:49:06.634596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:49:06.634596Z digest=sha256:7c7050c082e654114c97ed86767d5a975fe300d57af185ad016ccb6e343a62ea

Observation 4b3b02fd-e68b-480c-b3a2-1847fad0a99a · inbound

Real-Time Emergency Vehicle Siren Detection with Efficient CNNs on Embedded Hardware cites this paper.

Real-Time Emergency Vehicle Siren Detection with Efficient CNNs on Embedded Hardware FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:52:00.071485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:52:00.071485Z digest=sha256:aeca9a68d913595a8b190de31b75e2026f47418c069df4bb412afbe256d0af8f

Observation bee0be84-1bd7-4e1e-a8d3-fcb2c2b3ce93 · inbound

Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning cites this paper.

Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:56:46.317985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:56:46.317985Z digest=sha256:650100980cf34573b414b8f628a585cc87d93c391d1d71afec818a1740303fc2

Observation b858d7b1-faf9-42f1-89aa-1be051d1e7f8 · inbound

Low-latency Assistive Audio Enhancement for Neurodivergent People cites this paper.

Low-latency Assistive Audio Enhancement for Neurodivergent People FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T18:09:02.390461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:09:02.390461Z digest=sha256:fa1befc54bbc3fb7aa3e04b90c3690f9a7decd0a49d7cf8fe22bc851349d88a6

Observation d5b1cbd7-acff-4994-8927-6db5f8958285 · inbound

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents cites this paper.

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:46:37.596710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T16:44:32.104114Z digest=sha256:323cc7e932628d0789041ba656e3e613f17ebf86b6aad4c8b56d03f4ada85d66

Observation 9852ab7e-bf0b-43d6-8374-91d4be3778eb · inbound

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents cites this paper.

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T16:48:32.050707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:48:32.050707Z digest=sha256:9e90ef37c3e59229865d2492f9556a22fbb6cd96bd5a2cda4ff31cf5a681dd39

Observation 483a9fdf-7bb7-4c01-96f0-73ed0886bf1b · inbound

Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification cites this paper.

Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:44.964946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:44.964946Z digest=sha256:1b2758a56e63af46817ed3336301242954b18c3561b91043e61629741ee9f047

Observation 062c6ea6-8664-4f9e-8876-df52ca2b672b · inbound

DPDFNet: Boosting DeepFilterNet2 via Dual-Path RNN cites this paper.

DPDFNet: Boosting DeepFilterNet2 via Dual-Path RNN FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T15:34:29.194197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:34:29.194197Z digest=sha256:7e84ab139873094a125a68c93ddc55da686980fb8a35768330a73d1b49843286

Observation 9a2cea04-ae17-4add-8682-7e3a8f575ac7 · inbound

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models cites this paper.

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T17:01:07.806530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T16:58:42.970074Z digest=sha256:a9b3ecd94c66aec088c009496afd8a6dbbe1ceebe734f022f7eef5b22721c727

Observation d992015e-071c-4b5e-bc53-911ccd136b16 · inbound

Fine-grained Soundscape Control for Augmented Hearing cites this paper.

Fine-grained Soundscape Control for Augmented Hearing FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T20:00:00.358353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:00:00.358353Z digest=sha256:769b742751a558152e1f2e99bb93aab35b35d6c280752935f508d33a55c005c7

Observation 4ea5f234-c3e1-4b20-b44b-243c99124b5c · inbound

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models cites this paper.

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:36.035958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T14:58:27.176375Z digest=sha256:45ff683531cefd9bc46032299c7f16a6d7223f70a03a558b75aa3e07e4725594

Observation 7fb4a74b-5154-47f9-9d28-fbcc3f8f59a8 · inbound

Fusion Embedding: A Unified Embedding Space for Text, Image, Video, and Audio cites this paper.

Fusion Embedding: A Unified Embedding Space for Text, Image, Video, and Audio FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T14:46:35.435587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:46:35.435587Z digest=sha256:ee7e6c0597bc3327c5cc27bee68996aa4d88657c6363e2e710f9a5f87d50916b

Observation 77fb404d-c569-4f15-a722-967f86973136 · inbound

Hear, Invoke, and Understand: A Skill-Calling Multimodal Agent for Large Audio Language Models cites this paper.

Hear, Invoke, and Understand: A Skill-Calling Multimodal Agent for Large Audio Language Models FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T19:02:40.771164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:02:40.771164Z digest=sha256:e5d93fd07ad24a8b7af8f6268a6d529ef907e81baf9cfcc767673e2cbd53f68e