Pith. sign in

Paper Citation Record · LEDGER

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

As of 13 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.24059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24059 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:48.043254Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:41.644202Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:39:49.151368Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact3
  • verified fuzzy23
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a03eef48-8f97-425e-931d-ecfed85ff136 · outbound

This paper cites In speech production, individ- ual variability in both articulation and its consequent acoustics is robust [4, 5].

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition In speech production, individ- ual variability in both articulation and its consequent acoustics is robust [4, 5]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.910034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:41.558755Z digest=sha256:216236c3e5dc6b1224c385a0ae56c0b05f9a1ccf935b7dcb98070d78236e9669

Observation 4b43e443-4b31-4fb0-aec0-873a25c92896 · outbound

This paper cites Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:49.286208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:41.644202Z digest=sha256:6fcb5ed0574e3e9df19d23e62374dfc79658539bb9d9023bc8513b4d236423a2

Observation 19660cb0-be24-429e-ab9b-effb73b85fca · outbound

This paper cites Dataset We use a new single speaker rtMRI corpus with simultaneously recorded audio and video.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Dataset We use a new single speaker rtMRI corpus with simultaneously recorded audio and video

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.432731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:41.990420Z digest=sha256:f5565454c30ed43bca4c3d376af37ee21d8218ebcd673fda88b0f398b310d9f3

Observation 303e5974-c565-4d3d-aa47-6264830f8626 · outbound

This paper cites Phoneme error rate (PER) The overall PER results on the held-out test set are presented in Table 1.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Phoneme error rate (PER) The overall PER results on the held-out test set are presented in Table 1

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.067956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:42.226483Z digest=sha256:adf4cdbcb8d52c25ffc8d542f259d3bbfe1737866b69cf6217335ef0fb2060e2

Observation 8bdbaa1d-acef-4eb8-a05f-8f629f3b9610 · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:55.692447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:42.386227Z digest=sha256:eb5d372aaf856501e42e89cfd578bfef3600cda4bad1a8e895a17fdbf75b4f4e

Observation 4f9464f1-031f-4df4-aa14-49e5721d453c · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:55.539004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:42.468817Z digest=sha256:373dc351e01425cd4c4c9cf159ca779605044f9b87ef8cd392c68e4629e6ce13

Observation b55612bd-f444-4527-91c0-4121046af81d · outbound

This paper cites The architecture of speech production and the role of the phoneme in speech processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The architecture of speech production and the role of the phoneme in speech processing,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.456393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:43.937557Z digest=sha256:da0a1053b6b5fe0f3361066c725ff8f81b5f2b2cd83cd0949d5118f2a22bba7f

Observation 02634c6e-5d51-4923-8e6b-70a56bd6e2f3 · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:54.199624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:44.133282Z digest=sha256:9b4890a6b45b543e22c66dc7a6b8fba82cc5243d00e6a453f47c6a83e03e70ef

Observation ad08ce4c-bade-4fbb-9f2a-e86bc7862945 · outbound

This paper cites Sensorimotor integration in speech processing: computational basis and neural organization,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Sensorimotor integration in speech processing: computational basis and neural organization,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:55.388476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:42.627165Z digest=sha256:5547e4927fb8b99d017421d98ccea242ff0ee4d0acf0e7dba8510275adc3a1d5

Observation cefcdd35-e84e-47e7-83a3-67e13141c8a9 · outbound

This paper cites Perception drives production across sensory modalities: A network for sensorimotor integration of visual speech,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Perception drives production across sensory modalities: A network for sensorimotor integration of visual speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:55.193380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:42.893417Z digest=sha256:05976c00a90d5cb65d22893f85fdf158d0e944ececee30cf2e61a815a7d49279

Observation ba572b53-d737-4711-b1dc-fc85b808dfda · outbound

This paper cites The role of temporal modulation in sensorimotor interaction,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The role of temporal modulation in sensorimotor interaction,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.996114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:43.106505Z digest=sha256:05ae1fb29af0cc25e813cecb35f3362b3615f1ecda811ac73017fd9e5e16a5b4

Observation 85e5706b-555a-4ae0-bd48-ce09cd3fa621 · outbound

This paper cites Variability of articulator positions and formants across nine english vowels,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Variability of articulator positions and formants across nine english vowels,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.823327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:43.308671Z digest=sha256:8bc5a213da81eedda45157ad90ef39cafdd53de65d1d9bd0f38e0ecba7d4834c

Observation 28164159-cfd1-489b-9d53-4391d0aed9db · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:54.596321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:43.525990Z digest=sha256:a520f36eb8417fa32ba804340ab718a419a42dd76f154217927d92e3c22f4566

Observation 3583cc3e-bcdd-47fe-a774-a655e5c0234e · outbound

This paper cites Articulatory phonology: An overview,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Articulatory phonology: An overview,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:43.714834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:43.714834Z digest=sha256:82af88b688fa42b427489e9ce99f6ad2d1f672f128e65f73c744a25d62397bd1

Observation 1db0d81e-9b23-4164-93c8-43726580f67d · outbound

This paper cites Multimodal representations for syn- chronized speech and real-time mri video processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Multimodal representations for syn- chronized speech and real-time mri video processing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.064691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:45.396716Z digest=sha256:cee744fb416f157551711277204e895675fc50ec33d3121776610d88654b4f25

Observation 2e3d5334-098b-416f-8c84-7c5c79b3c398 · outbound

This paper cites Towards speech classification from acous- tic and vocal tract data in real-time mri,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards speech classification from acous- tic and vocal tract data in real-time mri,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.867621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:45.545477Z digest=sha256:f49173045975b9bd998cc3bdc916bf71829dda348d4de1d99f420da103a786df

Observation e7a1966a-fbb3-4274-88fc-a3c1116e54fb · outbound

This paper cites Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:44.345646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:44.345646Z digest=sha256:3b25b83976a7778d257916fb4fa70b18dfcfc06b29df801c40df52b272762f38

Observation f2a4bde9-cf0a-4d2d-8385-60c6cee01e8c · outbound

This paper cites Robust Audiovisual Speech Recognition Models with Mixture-of-Experts.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Robust Audiovisual Speech Recognition Models with Mixture-of-Experts

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:48.950065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:44.552105Z digest=sha256:90e0ea18ea93bce2a92d0420644e6b40be157bc9328196d72e2c027ef97d0cb7

Observation 002a3ab1-5feb-492e-a620-76fbaab3bf76 · outbound

This paper cites Speech production real-time mri at 0.55 t,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Speech production real-time mri at 0.55 t,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.935947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:44.741094Z digest=sha256:385a1c191e3789f2cc082b19824ad83e684e62ea1dbdb448b735742347873327

Observation 84caf0c9-50a7-4904-89ba-54b5dbe728e1 · outbound

This paper cites A multispeaker dataset of raw and reconstructed speech production real-time mri video and 3d volumetric images,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition A multispeaker dataset of raw and reconstructed speech production real-time mri video and 3d volumetric images,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.599113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:44.921151Z digest=sha256:5d1038b2f8276a90d9ff8b18aab3d6bd8640d0c03fa8225bb232760a7b6d9e06

Observation 943031f2-54a5-4935-baf7-120ab28b8497 · outbound

This paper cites The output of the Conformer is decoded by a single LSTM layer, with a final linear layer used for prediction.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The output of the Conformer is decoded by a single LSTM layer, with a final linear layer used for prediction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.668434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:41.818246Z digest=sha256:1dce2998954d7070ca16ccebd04f3cb3d8e383da017efe28459ff1f220c627a2

Observation 53f86a46-9a4c-4b2a-b685-81efa2cfa5ac · outbound

This paper cites Towards Automatic Speech Identification from Vocal Tract Shape Dynamics in Real-time MRI.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards Automatic Speech Identification from Vocal Tract Shape Dynamics in Real-time MRI

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:45.026461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:45.026461Z digest=sha256:a803fa6753f1ffa9b7076e6dc2de94c96c851ae084b475ec1ac07194fb3c3243

Observation 8f0f584e-2244-427e-894a-e5164b58361b · outbound

This paper cites Cnn-based phoneme classifier from vocal tract mri learns embed- ding consistent with articulatory topology.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Cnn-based phoneme classifier from vocal tract mri learns embed- ding consistent with articulatory topology

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.277765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:45.220360Z digest=sha256:f0f728695735fe4ab70011f7478eb53508d7bd676daa2ec1a57f14f0a6d8dc1f

Observation c7435ec9-7a19-48f8-8a38-cca29cc75173 · outbound

This paper cites Eval- uation of a novel 8-channel rx coil for speech production mri at 0.55 t,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Eval- uation of a novel 8-channel rx coil for speech production mri at 0.55 t,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.456383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:46.700802Z digest=sha256:093548e1063fb4542e438c91b71b12e358903c924d2994aab27645f064cc23b2

Observation 3d723bf9-a38a-4dce-b296-52a9c168af75 · outbound

This paper cites Direct articulatory observa- tion reveals phoneme recognition performance characteristics of a self-supervised speech model,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Direct articulatory observa- tion reveals phoneme recognition performance characteristics of a self-supervised speech model,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.651440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:45.684601Z digest=sha256:54cca2b0aeb04f37a2c22a1c744f0c9ee94074d9cf1d58461fefd01a33df4532

Observation 4a4267c1-21c3-48f6-8d48-5aa5897810fa · outbound

This paper cites Usc-timit: A database of multimodal speech production data,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Usc-timit: A database of multimodal speech production data,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.392791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:45.838539Z digest=sha256:9989becdf5cebbe365e93e865087f2e103c0d100e99330dabe3ba820d4ba3c9e

Observation 4a5597dc-eb9f-4ab4-835b-3445f4b1a323 · outbound

This paper cites Database of volumetric and real-time vocal tract mri for speech science.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Database of volumetric and real-time vocal tract mri for speech science

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.163254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:46.054670Z digest=sha256:605e4bff3b0d479201e239feeaa27a2f4943848e812ddd1c47d0222b67a6a9a2

Observation 93f597cf-99dd-42af-a453-18d0d01513ef · outbound

This paper cites Characterization of inter-speaker articulatory variability: A two- level multi-speaker modelling approach based on mri data,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Characterization of inter-speaker articulatory variability: A two- level multi-speaker modelling approach based on mri data,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.879504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:46.187469Z digest=sha256:c674ebb363cd2f32a9954ea476379e5e60f01c0921b8c5e3a81ab6ec9db1116a

Observation 61506fae-6150-4f64-9c98-12aa2c30bd3e · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Conformer: Convolution-augmented Transformer for Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.387249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.387249Z digest=sha256:b8ba0b0b6f04eea8d0c7a128c8a2b0dfd6ec801438d0909c76986125ae6f96f4

Observation 871471fd-e05a-40c1-99e4-0a104d85efe2 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.480237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.480237Z digest=sha256:5787dbda2e124ba1c5a423a7123303f5543cf1b4f6e92e2a0b713dc51524cde7

Observation 6f715e8c-6636-4581-b5c9-a3e2626e3e08 · outbound

This paper cites Simple and Effective Zero-shot Cross-lingual Phoneme Recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Simple and Effective Zero-shot Cross-lingual Phoneme Recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.592923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.592923Z digest=sha256:1376094baf4c3b7f76bad6d1c342a890e3027427e1cabf2591cff4f48d921888

Observation e9410a79-ce06-4001-8534-f1eba9e98040 · outbound

This paper cites Robust and Efficient Medical Imaging with Self-Supervision.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Robust and Efficient Medical Imaging with Self-Supervision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.585835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.585835Z digest=sha256:1ee544cf2b20a8d4a4652ab995a955cd47d0d5acd29d1073ec94a3fdee6bd6e1

Observation 445e246d-f9f9-4650-8dd7-e3230c14340d · outbound

This paper cites State-of-the-art speech production mri protocol for new 0.55 tesla scanners,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition State-of-the-art speech production mri protocol for new 0.55 tesla scanners,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.243216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:46.766700Z digest=sha256:ac922001df9cded7d78f76fc32e44ca1fb8c99b619929ecbf3d61837884905f7

Observation 93c1d063-cc4a-4c4c-88a2-c60ffcfbd0fb · outbound

This paper cites Announcing the electro- magnetic articulography (day 1) subset of the mngu0 articulatory corpus,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Announcing the electro- magnetic articulography (day 1) subset of the mngu0 articulatory corpus,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:49.890156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:46.910754Z digest=sha256:936146ef818005a2e2d530057e53d2a7f0b98edace9d3dbb86e556a3b52fafd0

Observation 75c526be-bcfc-442f-bfee-473c2513d850 · outbound

This paper cites Real Time Speech Enhancement in the Waveform Domain.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Real Time Speech Enhancement in the Waveform Domain

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.014834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.014834Z digest=sha256:43cb4600b8b7f5243f5d8477289fb96e2d61d7adcd2952227fa9fb22896905c3

Observation 1808e0f1-00ca-4fce-bb7b-e83a42107629 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.089969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.089969Z digest=sha256:4def870c0a54cd522c5fe68fa2ecef887e25d2f7fa94e72c685b39c2c1dab100

Observation 951f1270-1d31-4711-90cc-71a59945f422 · outbound

This paper cites Evidence of Vocal Tract Articulation in Self-Supervised Learning of Speech.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Evidence of Vocal Tract Articulation in Self-Supervised Learning of Speech

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:48.383512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:47.199046Z digest=sha256:fea2d82b040c64a6b55c7d3e39a4ba251ddfb33ca0338542f793ba9da5be4b60

Observation 4d4b0c77-17ab-4de1-82bb-8b7beb2f3203 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.361459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.361459Z digest=sha256:e0c334bd2465f0a3dd55d80828307e7f2e20a53602dc31f0868e37fc38ccc969

Observation f2bc14b2-647d-4a89-8c33-86c79e04dbc2 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition A Simple Framework for Contrastive Learning of Visual Representations

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.436332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.436332Z digest=sha256:9bae96e2dd81cf09603ddcb387e05aed07226a3d3654854ac4c3093c849ac5a2

Observation 8a6cf77d-0e33-4af0-8982-9e0ea2a20e6d · outbound

This paper cites Visualizing data using t-sne.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Visualizing data using t-sne

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.724939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.724939Z digest=sha256:ad4d4549e49a9d79308e7873f7bf5aca5aa5949d633f129b4de2bf9381397ae5

Observation c9039109-fc06-40f7-9e0d-fc49b769315d · outbound

This paper cites Toward articulatory- acoustic models for liquid approximants based on mri and epg data. part ii. the rhotics,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Toward articulatory- acoustic models for liquid approximants based on mri and epg data. part ii. the rhotics,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:49.529579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:47.898973Z digest=sha256:0f09ee183960d523696830d719d0f5f146df8da128a6fe3236c11aef12aae88d

Observation 1251a66c-05b6-4b09-ad8a-62368c93a71e · outbound

This paper cites Articulatory characterization of english liquid- final rimes,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Articulatory characterization of english liquid- final rimes,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:48.043254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:48.043254Z digest=sha256:e7252c0b9c053deddf453cd00595e4e11b60fa126fce6e42b089dc8074096e85

Observation c6434ec6-3f83-45d8-b41f-4669a6b8313b · outbound

This paper cites The number of Conformer layers was set to 3.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The number of Conformer layers was set to 3

Reference 256

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.210483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:42.111215Z digest=sha256:f50c7dc41e58aff83df92ded806886a101767f866a8fb19566c600bbbf3f7da2

Pith citing papers

Observation 4b43e443-4b31-4fb0-aec0-873a25c92896 · inbound

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition cites this paper.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:49.286208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T12:39:41.644202Z digest=sha256:6fcb5ed0574e3e9df19d23e62374dfc79658539bb9d9023bc8513b4d236423a2