Pith. sign in

Paper Citation Record · LEDGER

Auditory Intelligence: Understanding the World Through Sound

As of 22 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2508.07829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07829 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:55:00.437900Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact9
  • verified fuzzy38
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d799a26c-be9e-4a90-bdc0-004f4d199ce9 · outbound

This paper cites GPT-4 Technical Report.

Auditory Intelligence: Understanding the World Through Sound GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:55.018861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:55.018861Z digest=sha256:47b1270d758d26ee868f9bbc6d08c65170b1c487786424e55bad6bfeb41e4d89

Observation 4e1cdf58-475e-45f1-b332-78498679b8c4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Auditory Intelligence: Understanding the World Through Sound Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:55.093314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:55.093314Z digest=sha256:fd4163b325816da10af0e6b3b4edebc774343f97c570900a24a67963f5df8ced

Observation 24a23040-184e-4408-8fb7-4eb7dd8f6171 · outbound

This paper cites DeepSeek-V3 Technical Report.

Auditory Intelligence: Understanding the World Through Sound DeepSeek-V3 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:55.207940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:55.207940Z digest=sha256:9d0040dc48d750ec0c910931693058a0886059d08995861a0a0294d20d987779

Observation 706dea90-ca0a-4a26-8799-cf7c8e25b31a · outbound

This paper cites Virtanen, M.

Auditory Intelligence: Understanding the World Through Sound Virtanen, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:09.763230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.290539Z digest=sha256:6a5a2febd16b72b162f5f067d7b07effcbaae63dba9a1872ecf82986296c22e1

Observation fef540cf-73ea-4132-a74f-d3f7504bc6b3 · outbound

This paper cites Sound event detection in domestic environments with weakly labeled data and soundscape synthesis,.

Auditory Intelligence: Understanding the World Through Sound Sound event detection in domestic environments with weakly labeled data and soundscape synthesis,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:09.449372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.368526Z digest=sha256:c40513d7dc0895d2759ef3c821c021469bad821672fa68687e33c6d3be4fe1e7

Observation 5bd1999d-f935-467f-9f97-9e972676f921 · outbound

This paper cites Convolutional recurrent neural networks for polyphonic sound event detection,.

Auditory Intelligence: Understanding the World Through Sound Convolutional recurrent neural networks for polyphonic sound event detection,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:09.138402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.459643Z digest=sha256:714364f484f2a05b1915accfcb36cfcd8904395761bdb975f75e2de486f4bec9

Observation eb209ae2-e210-4867-93c9-415fbddc20b0 · outbound

This paper cites Metrics for polyphonic sound event detection,.

Auditory Intelligence: Understanding the World Through Sound Metrics for polyphonic sound event detection,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:08.824581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.545487Z digest=sha256:8df5fc740cc0b6ec1ea0828e671a3eafa085f7a167db2a3f851a10d6abe8c655

Observation 230cba28-708a-4c03-b05e-054c3d08235c · outbound

This paper cites A framework for the robust evaluation of sound event detection,.

Auditory Intelligence: Understanding the World Through Sound A framework for the robust evaluation of sound event detection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:08.447221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.616813Z digest=sha256:8d039f4c459aa1e816eaf0d322018cf14ceeb8fec786a5413cec5dd1d413f199

Observation 9aa04b05-3827-4cbe-a996-4e467a4ee220 · outbound

This paper cites Towards Understanding of Frequency Dependence on Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Towards Understanding of Frequency Dependence on Sound Event Detection

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:02.185813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.618843Z digest=sha256:77185ff67958285d0e9118ec0cddea36dc968b6af6879a53c878ef848abc8613

Observation b63d5103-2042-45f4-a322-ae8232c20e11 · outbound

This paper cites JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:02.001185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.621227Z digest=sha256:94449723879ebe7bb9cd82d7d2bb231d2f1d1ee8680ef8cc1ad838fac34da2f7

Observation 5af3869f-a693-4e89-8801-49d2dc06c548 · outbound

This paper cites SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,.

Auditory Intelligence: Understanding the World Through Sound SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:08.124515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.623704Z digest=sha256:d3fe3f5cf2186070f0c196ffb8163e849db1bc135130e2289fb17d731c2b473f

Observation cd7fd502-ae51-4374-af28-ea291568350d · outbound

This paper cites Conformer: Convolution- augmented Transformer for Speech Recognition,.

Auditory Intelligence: Understanding the World Through Sound Conformer: Convolution- augmented Transformer for Speech Recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.854378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.696331Z digest=sha256:5fac05241aa920eff275c2cad99d45dcce6fbda75414d33152a816a0ae3a94ed

Observation e6ce8a0c-c5a9-4e18-8add-d94a0a1e06a2 · outbound

This paper cites Coherence-based phonemic analysis on the ef- fect of reverberation to practical automatic speech recognition,.

Auditory Intelligence: Understanding the World Through Sound Coherence-based phonemic analysis on the ef- fect of reverberation to practical automatic speech recognition,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.597433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.792917Z digest=sha256:92c177fcc5956ec7e1b4eae83823c05c4c2296ad454b3157717709d30e236da6

Observation 9e209df9-6509-4644-808f-04218b52ba55 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

Auditory Intelligence: Understanding the World Through Sound wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.258853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.856904Z digest=sha256:accbb674f129c19be113f6b4e9f83cce3133b4fe2337c5279b5f658f13679d0c

Observation 8f743f9e-7593-4972-969e-e9173af92efa · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Auditory Intelligence: Understanding the World Through Sound Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:07.011330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:55.890183Z digest=sha256:04c523d7234139ae91afa107eca37775fd6911347690d6872330aaad6ec8c5c2

Observation 60b67069-b919-4e57-8af0-b7ef5baa3df6 · outbound

This paper cites Attentive statistics pooling for deep speaker embedding,.

Auditory Intelligence: Understanding the World Through Sound Attentive statistics pooling for deep speaker embedding,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:56.043026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:56.043026Z digest=sha256:4f75da868b14048b11ecbd9fb2637a7d163239f58d110d974a7648eede7cde0d

Observation 68704b73-a938-4901-9684-5e008bac89d6 · outbound

This paper cites Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,.

Auditory Intelligence: Understanding the World Through Sound Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.780807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:56.245420Z digest=sha256:caab03d64be974e7fb4dbac466d3ae456d9a3d9632522a9166bff77dbd52b025

Observation 188f8834-4afc-4ca9-bfd0-63c267134a61 · outbound

This paper cites Analysis-based optimization of temporal dynamic convolutional neural network for text-independent speaker verification,.

Auditory Intelligence: Understanding the World Through Sound Analysis-based optimization of temporal dynamic convolutional neural network for text-independent speaker verification,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.713473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:56.417343Z digest=sha256:214f04dfc38f19a51b3e41fe50887a3c50e585c890b1deb4cfe3166b1e6a0a16

Observation e88911b6-715e-4cb2-abf6-f55cbece2c99 · outbound

This paper cites Integrating fre- quency translational invariance in tdnns and frequency positional in- formation in 2d resnets to enhance speaker verification,.

Auditory Intelligence: Understanding the World Through Sound Integrating fre- quency translational invariance in tdnns and frequency positional in- formation in 2d resnets to enhance speaker verification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.589708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:56.581826Z digest=sha256:7618c0731587f42da438a969a44cb9caf755c5c8ecd85027e5c31066efaf9b35

Observation f6828fbc-3eea-4ec9-858f-b650999147e0 · outbound

This paper cites Convolution-based channel-frequency atten- tion for text-independent speaker verification,.

Auditory Intelligence: Understanding the World Through Sound Convolution-based channel-frequency atten- tion for text-independent speaker verification,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.462490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:56.697963Z digest=sha256:123a205a4c7e53a7e07bda450f1b457356c42c15458b731208b760728a688035

Observation b0868592-5570-465e-a0d1-f6f08c9e6e85 · outbound

This paper cites Panns: Large-scale pretrained audio neural networks for audio pattern recognition,.

Auditory Intelligence: Understanding the World Through Sound Panns: Large-scale pretrained audio neural networks for audio pattern recognition,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.313111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:56.862622Z digest=sha256:8bdc78a80c802432df0a36c20e5d43ef6dc4d56cfb8db7ab8723f717d9c0692c

Observation 2de542e1-234c-4a89-9208-af40a6aa3584 · outbound

This paper cites Deep learning based cough detection camera using enhanced features,.

Auditory Intelligence: Understanding the World Through Sound Deep learning based cough detection camera using enhanced features,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:06.135508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.030581Z digest=sha256:57b0f02f368e8d5bed2860dd5e293da8865d93d1170d0b46ba2924490181ee54

Observation d64d0547-8f1c-40a6-8018-c7acb8450811 · outbound

This paper cites Real-time sound recognition system for human care robot considering custom sound events,.

Auditory Intelligence: Understanding the World Through Sound Real-time sound recognition system for human care robot considering custom sound events,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.927205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.193849Z digest=sha256:97cc42cc6b5f3f7ec08c032aaa86b2def5f88673dfb32eb97781b15a5017c38c

Observation f3480612-f4b5-44cb-815d-fbe705c16e93 · outbound

This paper cites Ast: Audio spectrogram trans- former,.

Auditory Intelligence: Understanding the World Through Sound Ast: Audio spectrogram trans- former,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.744457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.356267Z digest=sha256:88de57e57e4a2ed12e2ab24584bf7492f35c10035e2f70cc323334977d23edf5

Observation 049bd74c-97f3-445e-8e5c-83cf7d744dfe · outbound

This paper cites Beats: Audio pre-training with acoustic tokenizers,.

Auditory Intelligence: Understanding the World Through Sound Beats: Audio pre-training with acoustic tokenizers,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.563083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.480436Z digest=sha256:65cf9f181d9357984cf8b53308f3236a7048618a47025c95c4f558436dba8dce

Observation a325f9c7-30a8-497f-9c02-cd21b79416e7 · outbound

This paper cites Overview and evaluation of sound event localization and detection in dcase 2019,.

Auditory Intelligence: Understanding the World Through Sound Overview and evaluation of sound event localization and detection in dcase 2019,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.397189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.627088Z digest=sha256:ab1f3bd12eedabb62d83afa1027182eb2a1ff57509fd0cea7a48d3d59aac424d

Observation 02947e24-3dad-417e-89fe-c00cdcf9eba2 · outbound

This paper cites STARSS22: A dataset of spatial recordings of real scenes with spa- tiotemporal annotations of sound events,.

Auditory Intelligence: Understanding the World Through Sound STARSS22: A dataset of spatial recordings of real scenes with spa- tiotemporal annotations of sound events,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.229253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.738938Z digest=sha256:31513ce458006e9a310d66dcffd17064fe1386fbe22beae6d07a52c54757deff

Observation 9e0b3c59-c1e2-4daa-bfd4-278b4d3cd8c5 · outbound

This paper cites Data augmentation and squeeze-and-excitation network on multiple dimension for sound event localization and detection in real scenes,.

Auditory Intelligence: Understanding the World Through Sound Data augmentation and squeeze-and-excitation network on multiple dimension for sound event localization and detection in real scenes,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:05.050564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:57.899374Z digest=sha256:09ad08161b19c86f02b3a9994fef80cc8a3309ae75fd1a765f729ccef9673af7

Observation 30625b5f-9cdb-4af7-b716-9cdc483e3830 · outbound

This paper cites Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots.

Auditory Intelligence: Understanding the World Through Sound Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.847274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.024131Z digest=sha256:4701024f629fa4e7fe973eebbcc811b1ea297fd8afe33288176913f8c6362995

Observation 1e23078e-4c6b-4e27-8ea9-b43acbe34508 · outbound

This paper cites Automated audio captioning with recurrent neural networks,.

Auditory Intelligence: Understanding the World Through Sound Automated audio captioning with recurrent neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.881036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.170275Z digest=sha256:e28096d0e79c3cb636a681fb36e5bc68afad8f70c213889c6a8aae15005502f1

Observation 1a0665ed-0268-45ac-b880-3125acd53a1d · outbound

This paper cites Clotho: an audio captioning dataset,.

Auditory Intelligence: Understanding the World Through Sound Clotho: an audio captioning dataset,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.676532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.286059Z digest=sha256:ca0abe92e71afd4c3fe629120e31fbd0408318716b8cbfb9f687c13addbd194c

Observation 77818c72-35bb-4d60-bb00-17b96f46a42f · outbound

This paper cites Chatgpt caption paraphrasing and fense-based caption filtering for automated audio captioning,.

Auditory Intelligence: Understanding the World Through Sound Chatgpt caption paraphrasing and fense-based caption filtering for automated audio captioning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.531602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.404359Z digest=sha256:7bff6b03d7ab2d758a294f84fac74423ca66bfad5660b3fa4b54d00c0c316969

Observation 556bbac1-d0bc-4baf-a758-c59642c58474 · outbound

This paper cites Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.629069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.488390Z digest=sha256:56f486da570f7c3c12a7f4262545a06dfb9347c6e5702880a050bb4930567587

Observation 9633a1d8-64e0-48fd-bcfc-fabe39d98b4f · outbound

This paper cites Few-shot bioacoustic event detection utilizing spectro-temporal receptive field,.

Auditory Intelligence: Understanding the World Through Sound Few-shot bioacoustic event detection utilizing spectro-temporal receptive field,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.395425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.572096Z digest=sha256:16e76784b89884ffddffdfb4887fc087f1949739ddc24ba0944a5fa261cd001e

Observation a613f9fe-77bb-4d9b-b2a5-ca500286eb07 · outbound

This paper cites Prtfnet: Hrtf individual- ization for accurate spectral cues using a compact prtf,.

Auditory Intelligence: Understanding the World Through Sound Prtfnet: Hrtf individual- ization for accurate spectral cues using a compact prtf,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.245019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.674078Z digest=sha256:640f0674eb3a0d231fb6552a60dc79a8bf1b7549feeeb9dd00e196cdd3f52124

Observation d6f13d25-fadb-4cf4-90c1-cdf14ff29460 · outbound

This paper cites Filteraugment: An acoustic environmental data augmentation method,.

Auditory Intelligence: Understanding the World Through Sound Filteraugment: An acoustic environmental data augmentation method,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:04.073974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.787521Z digest=sha256:b0ece0dffee83ca18e7cc6a0e5d515b7a17503c9ff41f2172f765eee36af0ccd

Observation 369327be-28db-4ef5-baf5-9b2f03d07e45 · outbound

This paper cites DNN based HRIRs Identification with a Continuously Rotating Speaker Array.

Auditory Intelligence: Understanding the World Through Sound DNN based HRIRs Identification with a Continuously Rotating Speaker Array

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.438903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.901523Z digest=sha256:02e1dc0b418b872d081a2fc3300e89868df1829b8aa61f92c3f6bdaee21b0355

Observation 792542b4-877c-4b7c-922c-55834e3b1dff · outbound

This paper cites AudioLDM: Text-to-audio generation with latent diffusion models,.

Auditory Intelligence: Understanding the World Through Sound AudioLDM: Text-to-audio generation with latent diffusion models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.901509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:58.988369Z digest=sha256:09f63b52d2768c5ca16796c247c09257917971674498b38119e053c4fdca7957

Observation ada6fa31-0d43-4223-a3dc-defab0890554 · outbound

This paper cites Audiogen: Textually guided audio generation,.

Auditory Intelligence: Understanding the World Through Sound Audiogen: Textually guided audio generation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.753031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.102365Z digest=sha256:84e75dedc2235b8a7f1fe47cad877bbc38e134a53e99c59a7179fc82a8beec80

Observation d7c109d0-d67f-4729-bc6d-1c51d4b0c51b · outbound

This paper cites Vifs: An end-to-end variational inference for foley sound synthesis,.

Auditory Intelligence: Understanding the World Through Sound Vifs: An end-to-end variational inference for foley sound synthesis,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.554341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.215508Z digest=sha256:9a9c7307093a8f11d776772a9cd4f1ac053fc3777dcb682930684b3ca698378c

Observation 7bcbe46f-95b9-49f2-91b4-e1d3c2899ae0 · outbound

This paper cites Heavily augmented sound event detection utilizing weak predictions,.

Auditory Intelligence: Understanding the World Through Sound Heavily augmented sound event detection utilizing weak predictions,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.411377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.349962Z digest=sha256:4321392ac2070b35f868197e7b1f1e772bf18f3e89c240a22ff434ffa8d368bb

Observation 21a5f27c-e399-4348-a384-1d74fc22a414 · outbound

This paper cites Filteraugment: An acoustic environmental data augmentation method,.

Auditory Intelligence: Understanding the World Through Sound Filteraugment: An acoustic environmental data augmentation method,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.264445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.472133Z digest=sha256:a72e25fd2b80b359f8761615324aa42427c8cd6be3be77a06d88ef7c65628480

Observation 1973ca4d-8cb7-4cb9-a7b5-7759246abec0 · outbound

This paper cites Frequency Dynamic Convolution: Frequency-Adaptive Pattern Recognition for Sound Event Detection,.

Auditory Intelligence: Understanding the World Through Sound Frequency Dynamic Convolution: Frequency-Adaptive Pattern Recognition for Sound Event Detection,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:03.075348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.555478Z digest=sha256:5ef8e008cd536414675ac37caf39409abe23ec3541ef5a508313cb6cd087865e

Observation ded12c34-a415-43b6-9e1d-3e0d9102ebec · outbound

This paper cites Self training and ensembling frequency dependent networks with coarse prediction pooling and sound event bounding boxes,.

Auditory Intelligence: Understanding the World Through Sound Self training and ensembling frequency dependent networks with coarse prediction pooling and sound event bounding boxes,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.909992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.654198Z digest=sha256:0da4b692a343eefc694ac90c76606af6725f4c662360de946d64b0aa8f646c47

Observation 0ae9c7db-d015-4fed-a50c-4e1f603a90e2 · outbound

This paper cites Diversifying and expanding frequency-adaptive convolution kernels for sound event detection,.

Auditory Intelligence: Understanding the World Through Sound Diversifying and expanding frequency-adaptive convolution kernels for sound event detection,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.747332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.795404Z digest=sha256:eb273a5514ab4a63c5501a4c007a45232be73df5c0cf4629e175d49b509c0cbe

Observation 05d450ad-e4a7-4d4c-820a-035c26b9ba0f · outbound

This paper cites Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution.

Auditory Intelligence: Understanding the World Through Sound Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.245826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.894205Z digest=sha256:404b5c84e2b5ba9c3f87ddc394630fec45196bccd417311c8f8498e2ffaff5d0

Observation 72f54b94-765d-44ca-ac9e-ddc6677757af · outbound

This paper cites Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:01.067562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:54:59.989934Z digest=sha256:73127fc8c04fb2f48ab0b0e115dad9c8e3f6665e85bc689334b360d63d2f5ba9

Observation 773b6fc6-25d1-44e5-86c6-7f4ef1b8a7d8 · outbound

This paper cites Frequency Dynamic Convolutions for Sound Event Detection.

Auditory Intelligence: Understanding the World Through Sound Frequency Dynamic Convolutions for Sound Event Detection

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:00.851246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:55:00.075447Z digest=sha256:7392c5aabcaa2c8daf808db357b237206431b97a189356aa8a8f00679e567bbd

Observation f3f7fd21-81b5-464d-82dc-1d6b272b7750 · outbound

This paper cites FLAM: Frame-Wise Language-Audio Modeling.

Auditory Intelligence: Understanding the World Through Sound FLAM: Frame-Wise Language-Audio Modeling

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:55:00.654739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:55:00.164281Z digest=sha256:df03b7c9f67fd2288a6edbf265c20b9a3516b793102fc29c25b4dec6386c2b42

Observation 2d676abf-5417-479d-85b2-bcc09366cf9a · outbound

This paper cites AudioCaps: Generating captions for audios in the wild,.

Auditory Intelligence: Understanding the World Through Sound AudioCaps: Generating captions for audios in the wild,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.500837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:55:00.299272Z digest=sha256:0c24be4215587541d1218eb8efa6569efe9bb5a1bf1c91ea040a1ca4705b2818

Observation 476396e1-5fa4-4778-a4b8-db33cee5beeb · outbound

This paper cites Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,.

Auditory Intelligence: Understanding the World Through Sound Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:55:02.344444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T21:55:00.437900Z digest=sha256:c9a37dd2d8d95adf1fb78577be5e4754225e241b16066f32b0c3edbe7c4586ae

Pith citing papers

No inbound Pith citation observations are available.