Pith. sign in

Paper Citation Record · LEDGER

The ICME 2025 Audio Encoder Capability Challenge

As of 16 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2501.15302.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.15302 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:27:09.170151Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:13:20.213020Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T16:13:20.931322Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d5742c54-78ce-439f-acd9-f4858c731762 · outbound

This paper cites Neural discrete representation learning,.

The ICME 2025 Audio Encoder Capability Challenge Neural discrete representation learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.727024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.007976Z digest=sha256:6b4f1fb1203e5378f53783618254bb98ada51ed2664bdc1dcf122e8d03758ba8

Observation 483913b3-fe05-46f2-a617-e6a317fdf08d · outbound

This paper cites Finite Scalar Quantization: VQ-VAE Made Simple.

The ICME 2025 Audio Encoder Capability Challenge Finite Scalar Quantization: VQ-VAE Made Simple

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.014071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.014071Z digest=sha256:b0898e7443a28766c22ce21875060c4519b4392f059edece5d3d87a3c46dcbc6

Observation b299c320-1e63-46ff-8682-f3701a879f4d · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

The ICME 2025 Audio Encoder Capability Challenge High-fidelity audio compression with improved rvqgan,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.712327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.019677Z digest=sha256:2228258b69daf529a61dc81a993108f698d5a4836963b9e43b4c4af89916dfff

Observation 1c30ab33-fb6e-44df-9604-748d76e9988f · outbound

This paper cites SNAC: Multi-Scale Neural Audio Codec.

The ICME 2025 Audio Encoder Capability Challenge SNAC: Multi-Scale Neural Audio Codec

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.024591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.024591Z digest=sha256:41ffea279fae14f54f4a08196da91427d3ad85c09debd3ac3dbc73d2f928fbee

Observation ecdc0d7a-ce59-4dab-8e7a-cf9320779054 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

The ICME 2025 Audio Encoder Capability Challenge Moshi: a speech-text foundation model for real-time dialogue

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.030098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.030098Z digest=sha256:73fbeb9bdcea541ece790f119b01c46313ee902433c6c21d1b48a574834a4f4a

Observation cf075be4-d900-495b-b85f-80164947a406 · outbound

This paper cites A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models.

The ICME 2025 Audio Encoder Capability Challenge A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.035252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.035252Z digest=sha256:54f1158eaac2b71e23d9efe88e98f232216f826c827c583a60a51a77cba1d102

Observation dc7dbac0-f5eb-4f43-a2e6-060590170c4a · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

The ICME 2025 Audio Encoder Capability Challenge Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.041296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.041296Z digest=sha256:804728782f29aa0b56088747f643cb1ad4a8e05f87238621f8cc2a19d2681fad

Observation 4297e96e-970e-4e22-817e-13e50c59578a · outbound

This paper cites Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming.

The ICME 2025 Audio Encoder Capability Challenge Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.046367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.046367Z digest=sha256:3be268e6a8904fafde46a43e09f0cbfc1d67e4953fef0574d213580636b3799c

Observation 14e6393b-9283-46b5-b13f-bf5f6b465aad · outbound

This paper cites SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation.

The ICME 2025 Audio Encoder Capability Challenge SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.051470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.051470Z digest=sha256:c89e32f1e748cafd3ba6850a2b1d0a3f631e371187b474c5a5f9840f2d73e93c

Observation ea4c6619-997a-4f60-90dc-9ae8e2f9b9c8 · outbound

This paper cites wav2vec 2.0: A fr amework for self-supervised learning of speech representations,.

The ICME 2025 Audio Encoder Capability Challenge wav2vec 2.0: A fr amework for self-supervised learning of speech representations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.698860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.056593Z digest=sha256:b05e988ca1dcecfb6ccfbe6ba64ccf2f80d5560603695e6907885a16d1061378

Observation c48c1a4a-613b-485f-b7d1-b28ec87a461e · outbound

This paper cites Data2 vec: A general framework for self- supervised learning in speech, vision and language,.

The ICME 2025 Audio Encoder Capability Challenge Data2 vec: A general framework for self- supervised learning in speech, vision and language,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.685160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.061310Z digest=sha256:6f969ea967dde99e93cb63c5731a3495438be00b6eaed0a60762d87114516956

Observation 6d92c814-716c-4c51-9d5c-ee968146577a · outbound

This paper cites Sc aling up masked audio encoder learning for general audio classification,.

The ICME 2025 Audio Encoder Capability Challenge Sc aling up masked audio encoder learning for general audio classification,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.670843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.066268Z digest=sha256:6e984db9cbf05fc57125ffde3054e7dfc8ee86994cf033962039d310b132e9ac

Observation b609f98e-8786-4cfa-9002-31e146495105 · outbound

This paper cites HEAR: Holistic evaluation of audio representations,.

The ICME 2025 Audio Encoder Capability Challenge HEAR: Holistic evaluation of audio representations,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.656081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.071436Z digest=sha256:ec0d79d64a450a3d7ff48ef59f4cbde9908576e564efeded846b0d94d6c17d91

Observation f699b613-476e-400b-81ac-938993fe15d6 · outbound

This paper cites SUPERB: Speech processing universal performance benchmar k,.

The ICME 2025 Audio Encoder Capability Challenge SUPERB: Speech processing universal performance benchmar k,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.641396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.076302Z digest=sha256:acfcbdd94d57d8aa43804f3f05e5a005706059403fa9224edeaba91343bbfb1f

Observation eac09b7b-2251-4600-8aea-229613780c1a · outbound

This paper cites DASB - Discrete Audio and Speech Benchmark.

The ICME 2025 Audio Encoder Capability Challenge DASB - Discrete Audio and Speech Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.081166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.081166Z digest=sha256:ef22af545aac502a48ba33f88a448363cb98d677d778e926e54234613e46a2e7

Observation 2c95c848-6575-4f2f-b5a5-d8a8e783cf0f · outbound

This paper cites Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition.

The ICME 2025 Audio Encoder Capability Challenge Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.086029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.086029Z digest=sha256:856efb231f9448424cfabe50d2c53f17f31a5c382ae2a7f89c74b9d5beb94fa2

Observation 70681049-0a50-4aa7-bd9c-6ebbec8e85f1 · outbound

This paper cites Lib ricount, a dataset for speaker count estima- tion,.

The ICME 2025 Audio Encoder Capability Challenge Lib ricount, a dataset for speaker count estima- tion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.626585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.090811Z digest=sha256:caf827a0415675c7b382f4868ec7ddb94f12f2aefc075b44ed74bcbe35e67535

Observation 60ec8637-d5dc-464a-829a-f0eb4faa7e4b · outbound

This paper cites Voxlingua107: a dataset for spoken lan guage recognition,.

The ICME 2025 Audio Encoder Capability Challenge Voxlingua107: a dataset for spoken lan guage recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.611650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.094633Z digest=sha256:e598b6eddfa2b5f66ef284679bbce392077f8012c8b2949656a67e6d9f4849d3

Observation d4049d2a-85e2-4917-b84b-02feaa2f07a7 · outbound

This paper cites Voxceleb: L arge-scale speaker verification in the wild,.

The ICME 2025 Audio Encoder Capability Challenge Voxceleb: L arge-scale speaker verification in the wild,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.596567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.098417Z digest=sha256:f6a3b64ccd4d5aa5e9a41ad76d658245a9fdf48c49ff3df3c343383836f448d8

Observation 4bfa7132-234c-4247-bc62-793d83ec73ea · outbound

This paper cites Librisp eech: an asr corpus based on public domain audio books,.

The ICME 2025 Audio Encoder Capability Challenge Librisp eech: an asr corpus based on public domain audio books,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.581761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.102379Z digest=sha256:94c94c649b6b19fd31bfb3934aeb38d3cc036f67991d41086fb5d58b326e2a07

Observation 08bbda33-2a61-404b-9116-d7ce1ba5c64b · outbound

This paper cites Speech Model Pre-training for End-to-End Spoken Language Understanding.

The ICME 2025 Audio Encoder Capability Challenge Speech Model Pre-training for End-to-End Spoken Language Understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.106687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.106687Z digest=sha256:bc8eac31f406027a8dabe54579f011d7e56028361a2e948fd562e49ad657cfbf

Observation 9a89e309-af28-4ec2-b079-a918ad385d20 · outbound

This paper cites Vocalsound: A dataset for impro ving human vocal sounds recognition,.

The ICME 2025 Audio Encoder Capability Challenge Vocalsound: A dataset for impro ving human vocal sounds recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.567012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.111099Z digest=sha256:ab0cf48ef2accc3641c07ef2506929a9aaf128d4a19d823238781183d6c7d148

Observation a1f097c8-add5-43fd-81b5-2c4f1967fcb2 · outbound

This paper cites Crema-d: Crowd- sourced emotional multimodal actors dataset,.

The ICME 2025 Audio Encoder Capability Challenge Crema-d: Crowd- sourced emotional multimodal actors dataset,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.551485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.115119Z digest=sha256:18b89fbbb258f252243a725b368c2bf8afbe05abc26439c4aec720fbacdb8128

Observation b199aac7-ee97-4346-ac42-95719c2825ad · outbound

This paper cites spee- chocean762: An open-source non-native english speech corpus f or pronunciation assessment,.

The ICME 2025 Audio Encoder Capability Challenge spee- chocean762: An open-source non-native english speech corpus f or pronunciation assessment,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.535715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.118950Z digest=sha256:c1623f2502c3875812f21b028245b3dabe7470eabf5d08df00c2b6db10e94ad3

Observation f72d2a44-6b37-4c5e-a149-b16d19d43ef7 · outbound

This paper cites Autom atic speaker verification spoofing and countermeasures challenge (asvspoof 2015) database,.

The ICME 2025 Audio Encoder Capability Challenge Autom atic speaker verification spoofing and countermeasures challenge (asvspoof 2015) database,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.521154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.123480Z digest=sha256:61d124cd8347d0c141e345c427d3ceb1300347d41a0973353df27dc636a8f441

Observation 60ee67fa-45d4-45b6-b882-cf47dccb327d · outbound

This paper cites Esc: Dataset for environmental sound classific ation,.

The ICME 2025 Audio Encoder Capability Challenge Esc: Dataset for environmental sound classific ation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.506592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.128144Z digest=sha256:39a43d797b57eb995068262f7749e7effe9ebd4a96bc58ff8ab0cd570bc47953

Observation 647aa645-ec38-4b94-a432-7f7004b675eb · outbound

This paper cites Fsd 50k: an open dataset of human-labeled sound events,.

The ICME 2025 Audio Encoder Capability Challenge Fsd 50k: an open dataset of human-labeled sound events,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.491604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.132952Z digest=sha256:f69b84148bb51f6e90474086b0833ceccdd0916079a720d822f63452c78f03a9

Observation 734b0a95-283c-4511-8f8e-c31e0a3e67a5 · outbound

This paper cites A dataset and taxonom y for urban sound research,.

The ICME 2025 Audio Encoder Capability Challenge A dataset and taxonom y for urban sound research,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.475397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.137432Z digest=sha256:5854638fd3f2b05aad2e43ed6edb23e928f32d7af992e9621f8af6cf2c55c6f7

Observation 0a8327d8-ad6c-4706-b9b9-af6b735cb42c · outbound

This paper cites Sound eve nt detection in domestic environments with weakly labeled data and soundscape synthesis,.

The ICME 2025 Audio Encoder Capability Challenge Sound eve nt detection in domestic environments with weakly labeled data and soundscape synthesis,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.459336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.142173Z digest=sha256:06e1e6982f7701f0cd4df398f9d6d9f0ddce8bfc19aa1e59dccef5cb2774b78f

Observation e0152afd-8583-423a-bc21-81ae392ac2b9 · outbound

This paper cites General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline.

The ICME 2025 Audio Encoder Capability Challenge General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.146676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.146676Z digest=sha256:33ec289ed9dd8cdd73f1a9f4f4d33cb739965bdeb8d51571399785b11d63c10b

Observation 4dd00642-a5e3-4629-bdde-1c92b4c6867d · outbound

This paper cites Clotho: An audio capt ioning dataset,.

The ICME 2025 Audio Encoder Capability Challenge Clotho: An audio capt ioning dataset,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.443976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.151378Z digest=sha256:cec8b5ab77ad905ebad4cafafc99ce4d0eac625509046ad9f5f04f48bf3953f0

Observation 99504da6-da4d-46c3-a1b1-a46ae6aebb4f · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAESTRO dataset,.

The ICME 2025 Audio Encoder Capability Challenge Enabling factorized piano music modeling and generation with the MAESTRO dataset,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.427914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.155999Z digest=sha256:cb129ad8a828e97f708b0d371dc451afa3ee74176e0acf4ee7133483b64c01f3

Observation c7b1239c-9c2c-4c2b-b75e-5b5ea524e3e7 · outbound

This paper cites The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use.

The ICME 2025 Audio Encoder Capability Challenge The GTZAN dataset: Its contents, its faults, their effects on evaluation, and its future use

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.160670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.160670Z digest=sha256:b662a848c5c0a12452fcd411ac58695b536b94eada27f54e3814091aa3357bc5

Observation 28403eb4-d70b-49f0-9d15-1e0ec430e1a5 · outbound

This paper cites Neural audio synthesis of musical notes with wavenet autoencoders,.

The ICME 2025 Audio Encoder Capability Challenge Neural audio synthesis of musical notes with wavenet autoencoders,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:27:09.411329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T14:27:09.165586Z digest=sha256:d35fec0b83f6badfaa06544e4548588c61a9d1be8355363819841138377dd1c8

Observation f443ba02-0d14-4007-91a0-ad1005d83722 · outbound

This paper cites FMA: A Dataset For Music Analysis.

The ICME 2025 Audio Encoder Capability Challenge FMA: A Dataset For Music Analysis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.170151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.170151Z digest=sha256:4ff5931e1ebeede30f962425d132911d5f1d3d1b141d06b9b2737b089cb054db

Pith citing papers

Observation ee2596fe-f80a-466a-af9c-fdd960cdf5e5 · inbound

OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder cites this paper.

OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder The ICME 2025 Audio Encoder Capability Challenge

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:13:20.989094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T16:13:20.213020Z digest=sha256:4e6b48e06b197e9728c526af4d894a39f61ab7e0e093da74bab16960ddf1d7ee