Pith. sign in

Paper Citation Record · LEDGER

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

As of 18 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2506.02181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02181 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:17.610645Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:13.765343Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:32:17.963370Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy36
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 32be8668-a1d0-4b36-9bae-41ad920028b8 · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:25.650868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:13.696269Z digest=sha256:04642064e9a392abcdf5bd6a01dd8a449c450fbd60388e3b7817fa64c02d5fc9

Observation 0b075e5a-482f-47e0-ab8a-9b01b6920469 · outbound

This paper cites Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:32:18.040121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:13.765343Z digest=sha256:68040d3ca1988724e85727e1c6e3ac07ca434b0f213aef2855705fa3bf78ac21

Observation b0cc402e-91c3-4731-bb11-86d9ab1b1021 · outbound

This paper cites Output text is encoded into BPE.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Output text is encoded into BPE

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.519045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:13.851341Z digest=sha256:e3fffee0bef94b11d6d7fba98686962c025ea97817a6247239222b8d36833eb1

Observation 55b8654c-edf4-4196-a818-749efbd5c139 · outbound

This paper cites Time Fig.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Time Fig

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.090547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.077263Z digest=sha256:e885e77c7f0fc23f422ff645e99490a4069e90695af407a7c554711734198f01

Observation df277d58-ba0c-4348-87c0-c04a91f3b2e3 · outbound

This paper cites The encoder layers use a convolution kernel size of31, an em- bedding size of512, and a linear layer hidden size of 2 048.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution The encoder layers use a convolution kernel size of31, an em- bedding size of512, and a linear layer hidden size of 2 048

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.248711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:13.957612Z digest=sha256:08dac3e5419a7a3c34f68e6520141b56388c7025025db5b90d632a5b32d7bcbc

Observation 237cf734-eaba-43a1-a1a0-261ae9493e39 · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:24.899323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.179728Z digest=sha256:47f55da8e1f3b0a93a4f4d4c11bbb3f0635f811a6de729f8e1023cf5181ce1f4

Observation 192b2f02-aa97-4c3d-82e6-7deffcdda717 · outbound

This paper cites This is the first in-depth analysis of saliency maps in relation to fine-grained acoustic patterns across three phoneme classes.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution This is the first in-depth analysis of saliency maps in relation to fine-grained acoustic patterns across three phoneme classes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:25.001122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.125900Z digest=sha256:ab1564643ff4b143f5c63cd02a1c37166fd4ac6d6b60d4c64593dcaf93671857

Observation bfa9bd22-277f-41f7-99d1-40563c5d85ee · outbound

This paper cites Understanding the representa- tion and computation of multilayer perceptrons: A case study in speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Understanding the representa- tion and computation of multilayer perceptrons: A case study in speech recognition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.407824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.927096Z digest=sha256:2f6783b5fe41450ffe0e8929b322953e9a8e39ebd9565402e2bc72b654567e7f

Observation 8427c8e7-6bed-4d11-9105-189ad95a3a56 · outbound

This paper cites Analyzing phonetic and graphemic representations in end-to-end automatic speech recog- nition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Analyzing phonetic and graphemic representations in end-to-end automatic speech recog- nition,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.714209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.217408Z digest=sha256:4147a3ec302cd90c021c662249f9cb6f3f7c9f0bf7db6d49e1a3c53e96467c55

Observation 29201fdd-5aa4-4714-9f74-777854121a46 · outbound

This paper cites Probing phoneme, language and speaker information in unsupervised speech representations,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Probing phoneme, language and speaker information in unsupervised speech representations,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.485514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.303707Z digest=sha256:c27786b9fd715a517969f8ee2b8510a325c67793c435f2a41f8b59a550324b89

Observation dbe9d202-7fd3-4c78-aeab-be3dba4606a2 · outbound

This paper cites Domain-Informed Probing of wav2vec 2.0 Embeddings for Pho- netic Features,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Domain-Informed Probing of wav2vec 2.0 Embeddings for Pho- netic Features,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.312303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.412565Z digest=sha256:d5543471fd8e3a48bc89c12c6237f7f507e36a6d2d179b8a719b899d0a1a1594

Observation d1e2d11f-e127-44a7-88c1-9123d62fd077 · outbound

This paper cites Probing self- supervised speech models for phonetic and phonemic informa- tion: A case study in aspiration,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Probing self- supervised speech models for phonetic and phonemic informa- tion: A case study in aspiration,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:24.127112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.556191Z digest=sha256:d919432eca97b892ac2852e940afe0f3b59e73e98692580428c04616dc995901

Observation c469bcd5-9759-4445-90a0-6495cc0b9107 · outbound

This paper cites Self-supervised speech representations are more phonetic than semantic,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Self-supervised speech representations are more phonetic than semantic,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.912683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.636029Z digest=sha256:b745bf76a171f69898f68f17c374150a41f3da10a74512e5a2798f4b4ea4311f

Observation a9a232c3-1c1d-478b-b710-8fb9d5be4099 · outbound

This paper cites Encoding of phonol- ogy in a recurrent neural model of grounded speech,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Encoding of phonol- ogy in a recurrent neural model of grounded speech,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.745518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.725407Z digest=sha256:c0937b0aa6f8826436964b966121c2e7c31c43cd3c63d47204cc28be66a0d49d

Observation cbaf82b0-f548-412b-b390-a1ef98ef0803 · outbound

This paper cites Learn- ing weakly supervised multimodal phoneme embeddings,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Learn- ing weakly supervised multimodal phoneme embeddings,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.605363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:14.829948Z digest=sha256:b543ef908c24fcbc2f2fcbf90d42b30381d7b1b7d26043c057515a99c1f0a604

Observation 4da04087-fc0b-499c-a2b8-d111cefd052a · outbound

This paper cites Acous- tic characteristics of American English vowels,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acous- tic characteristics of American English vowels,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.306585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.511358Z digest=sha256:69f8f357556ac1c53c9363c663f8a0f8cedbcb7f5e5797e444e34842b894ce41

Observation 235f25a8-d890-43e1-8166-70c1b049d5af · outbound

This paper cites Neuron Activation Profiles for Interpreting Convolutional Speech Recognition Models,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Neuron Activation Profiles for Interpreting Convolutional Speech Recognition Models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.215208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.004400Z digest=sha256:49de18c20acf55f2cfe9523109668b2d91fbdfb11c981dfbb71271372e4e233c

Observation af0cc813-9177-45ea-841e-5d938a432c13 · outbound

This paper cites Interpretable Convolutional Filters with SincNet,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Interpretable Convolutional Filters with SincNet,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:23.047254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.077490Z digest=sha256:78d0ddfdd557a6807a145131d62579c7d2c73ad9ab5701bbbeeb021d493ea147

Observation 1c7f36dc-b57b-4146-8590-64f887c27d60 · outbound

This paper cites Introspection for convolutional automatic speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Introspection for convolutional automatic speech recognition,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.858990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.146016Z digest=sha256:1ff5132354c1f95d1ccabab5d9bdf4f32bca90b01a425100a7b7b6a828b33557

Observation cbad7af6-f918-4f29-b034-15bd8714eaef · outbound

This paper cites End-to-end acoustic modeling using convolutional neural networks for HMM- based automatic speech recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution End-to-end acoustic modeling using convolutional neural networks for HMM- based automatic speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.652588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.227371Z digest=sha256:c6ad3e1c00703711b8fbee152bfd77775b67607aedbb571ca551dd325a0406d3

Observation 2c80d404-16f8-4f7f-876e-d7acd6fc5683 · outbound

This paper cites Gradient-Adjusted Neuron Activation Profiles for Comprehensive Introspection of Convolutional Speech Recognition Models.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Gradient-Adjusted Neuron Activation Profiles for Comprehensive Introspection of Convolutional Speech Recognition Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:32:17.857757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.281363Z digest=sha256:b4c8682a48ecd1c9270c1c59f18d5325ddb04f7b0ae81bebfdafcd23a1b54f81

Observation 34779c97-1dc5-4c41-89fe-784ecac7585e · outbound

This paper cites Directly Comparing the Listening Strategies of Humans and Machines,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Directly Comparing the Listening Strategies of Humans and Machines,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.489366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.360537Z digest=sha256:1ea663cdc027fb35467de56351a207b20f57e7db7f027502d58962d14d71a82f

Observation c0ff9443-cf9f-4f20-b0a5-8365f0174d6d · outbound

This paper cites SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:15.428570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:15.428570Z digest=sha256:ea9982d373457d785b2324b54515512a66b550874295d6819462f698094b875b

Observation b03f39f8-479e-4e1c-85ab-9dd34483a66b · outbound

This paper cites an unresolved cited work.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:32:25.364852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:13.923390Z digest=sha256:be3f24f93bbd321a10fd72753dd56e9f3a8a61429f18d5b8ace66ebea35d9f8b

Observation 9a5d5fc3-9bff-4c7e-9af2-d9bccee1abff · outbound

This paper cites Burst and Transition Cues to V oicing Perception for Spo- ken Initial Stops by Impaired- and Normal-Hearing Listeners,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Burst and Transition Cues to V oicing Perception for Spo- ken Initial Stops by Impaired- and Normal-Hearing Listeners,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:22.124476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.594193Z digest=sha256:2d42a68435f5705f220459f31286c2d764d63f552b8b035996b7b0040f084550

Observation 4c5491e7-8eea-45f3-95e2-2a5e777da75e · outbound

This paper cites Acoustic cues of voiced and voiceless plosives for determining place of articulation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic cues of voiced and voiceless plosives for determining place of articulation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.905743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.692957Z digest=sha256:0d9fbf9d0a23fc0f9947a8021db66029fc67d79daa0b86b70ff302842ad923b0

Observation 1fe81541-41d6-48f7-95b1-c105c5a05685 · outbound

This paper cites Acoustic characteristics of English fricatives,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic characteristics of English fricatives,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.687005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.779334Z digest=sha256:80543be495fa3e9c466a15f415efd5e7cb88360b7c48d6b3a36929c8d6587958

Observation 24f587b8-0313-400c-866b-b6335394924b · outbound

This paper cites Acoustic characteristics of clearly spoken English fricatives,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Acoustic characteristics of clearly spoken English fricatives,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.511061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.889219Z digest=sha256:84f40688b840fd0c7fe20abea44e3a4cc52664c1f15aa5e76d736d2ab96fda52

Observation a2da6397-2f44-4d45-ad16-4085f91b2d35 · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.258248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:15.968308Z digest=sha256:7d87403fd981474f21a6c8180048fa07abce03beebd9dcaa0377dc88c3794bb4

Observation 8a09b4e1-738f-41ad-b7aa-3375bf1f9000 · outbound

This paper cites Introducing Parsel- mouth: A Python interface to Praat,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Introducing Parsel- mouth: A Python interface to Praat,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.075519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.075519Z digest=sha256:99ba3ebcd727512b072c565532ebe85235ae90e0cf90c3215189901a38dd53c9

Observation af314bcb-d75a-494b-bde2-ee3b78eecc9f · outbound

This paper cites Pykaldi: A Python Wrapper for Kaldi,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Pykaldi: A Python Wrapper for Kaldi,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:21.020089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.143157Z digest=sha256:88c0712189aee60419280c96f57c35903212068357e44e51c5854e3759328d0f

Observation 3effd402-33aa-4488-80dd-0fb1fe60c79c · outbound

This paper cites Neural Machine Transla- tion of Rare Words with Subword Units,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Neural Machine Transla- tion of Rare Words with Subword Units,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.786942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.218548Z digest=sha256:11996b5da00cee0fd92da1d7b3babb5396ee9d667834e48062ee29c4054381b0

Observation 592131ef-154a-4488-bb6f-b1564daa0dd9 · outbound

This paper cites SentencePiece: A simple and lan- guage independent subword tokenizer and detokenizer for Neural Text Processing,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SentencePiece: A simple and lan- guage independent subword tokenizer and detokenizer for Neural Text Processing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.523379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.297689Z digest=sha256:34cb56aecde2dcc90c123e43a13ecf7bac8841672c7e2f0100371506aa2c4774

Observation 3f931957-e4ad-4cb7-95f1-c7a131c9a156 · outbound

This paper cites Attention is All you Need,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Attention is All you Need,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.287706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.359174Z digest=sha256:48d46653d1c0495ef23e80b44e648fea8013667cecbb94ef6bdfaa2c6566eb7a

Observation fe1b260f-6a28-4646-9634-947f5a0e244d · outbound

This paper cites Fairseq S2T: Fast Speech-to-Text Modeling with Fairseq,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Fairseq S2T: Fast Speech-to-Text Modeling with Fairseq,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.455994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.455994Z digest=sha256:b5b9cb22694590103d1de769ab697bcb87276f743296736175a587ccbe824f78

Observation 6d166b02-a9eb-4aad-b7b8-5a8562bf286a · outbound

This paper cites Com- mon V oice: A Massively-Multilingual Speech Corpus,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Com- mon V oice: A Massively-Multilingual Speech Corpus,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:20.058654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.536895Z digest=sha256:47929d3447ed93f4c3c782c8687ee76b4e0cc88d8ea9e4e88e180347080c2a94

Observation 5b9d54ab-ea59-49dd-a540-7eca7aa6a668 · outbound

This paper cites Lib- rispeech: An ASR corpus based on public domain audio books,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Lib- rispeech: An ASR corpus based on public domain audio books,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:16.626379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:16.626379Z digest=sha256:c8cff63ac2a6c048ccec8801592a5a938799d21b68db4fc3326cc56bb9f9c21c

Observation db3fa82a-0329-4f53-950d-1afe1a917983 · outbound

This paper cites TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution TED-LIUM 3: Twice as Much Data and Corpus Repartition for Experiments on Speaker Adaptation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.771580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.724921Z digest=sha256:117254a10acae45313959f0e2d56bc152c26b7e6087bd79745b8195fdf5a707e

Observation 824f9cd2-f374-4c14-b28d-3409165a5104 · outbound

This paper cites V oxPopuli: A Large- Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution V oxPopuli: A Large- Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.558003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.814577Z digest=sha256:e181edfe44b5090841eed162268318fff6a1494a5bef5796b295ead182854e2d

Observation 26f4e42a-4f62-481e-9f70-c3a43f6a9d06 · outbound

This paper cites Rethinking the Inception Architecture for Computer Vision,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Rethinking the Inception Architecture for Computer Vision,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.372863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:16.949565Z digest=sha256:8f02b8f71743182ed8974103922da40847ff825870c4fe98c8d1d11ee8468383

Observation 800f6bae-2e69-4255-b17b-ccd2ea389988 · outbound

This paper cites Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Con- nectionist temporal classification: labelling unsegmented se- quence data with recurrent neural networks,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:19.203475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:17.048361Z digest=sha256:237ccca8702d8744de0a7be4bdc015fa4e11d948ec0eb1ce57032ef252735940

Observation b5794d0f-42b7-49d1-aad3-55220a5f8f8a · outbound

This paper cites Adam: A Method for Stochastic Opti- mization,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Adam: A Method for Stochastic Opti- mization,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:17.167561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:17.167561Z digest=sha256:182296812eb433f25bf1c288149a97466eedadfe5a1f9d31a3c8957a44d71fa5

Observation 62b93cdb-19e3-43f4-b3ce-ada0bfe6a740 · outbound

This paper cites SpecAugment: A Simple Data Aug- mentation Method for Automatic Speech Recognition,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution SpecAugment: A Simple Data Aug- mentation Method for Automatic Speech Recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.943826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:17.259642Z digest=sha256:b4b6754c7145b8fb99a878857db86b39fb5cb3bbf1a793c15db14731bef28ae9

Observation 57088e64-c670-4bf6-a24c-a8f1cf64f16e · outbound

This paper cites Darpa timit acoustic-phonetic continuous speech corpus cd-rom TIMIT,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Darpa timit acoustic-phonetic continuous speech corpus cd-rom TIMIT,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.661884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:17.396416Z digest=sha256:ac30e91bafebe0a9a248007643f8de3fab837051feea7e72035a88590bc1d30c

Observation 53152cd0-8d1c-437f-9961-44133f0028fe · outbound

This paper cites Twists, humps, and pebbles: Multilingual speech recognition models exhibit gen- der performance gaps,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Twists, humps, and pebbles: Multilingual speech recognition models exhibit gen- der performance gaps,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.424129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:17.502417Z digest=sha256:3336b013a48d916274e355ecc05d8818c7a046da9a84168161444b918a67f082

Observation a73cfbae-9702-461d-be97-e71fab6edf86 · outbound

This paper cites Burst spectrum as a cue for the stop voicing contrast in American English,.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Burst spectrum as a cue for the stop voicing contrast in American English,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:32:18.247636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:17.610645Z digest=sha256:55dd18895b3dacf903ac7fe7be5504e281f83ad476981eed23931b361788dbe7

Pith citing papers

Observation 0b075e5a-482f-47e0-ab8a-9b01b6920469 · inbound

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution cites this paper.

Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T11:32:18.040121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:32:13.765343Z digest=sha256:68040d3ca1988724e85727e1c6e3ac07ca434b0f213aef2855705fa3bf78ac21