Pith. sign in

Paper Citation Record · LEDGER

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.24059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24059 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:48.043254Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:41.644202Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:39:49.151368Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact3
  • verified fuzzy23
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a03eef48-8f97-425e-931d-ecfed85ff136 · outbound

This paper cites In speech production, individ- ual variability in both articulation and its consequent acoustics is robust [4, 5].

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition In speech production, individ- ual variability in both articulation and its consequent acoustics is robust [4, 5]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.910034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:41.558755Z digest=sha256:8a4367c55ab424906019b9282abccd6bd4cedbab694b14758c03bd4dc36f1bd4

Observation 4b43e443-4b31-4fb0-aec0-873a25c92896 · outbound

This paper cites Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:49.286208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:41.644202Z digest=sha256:98aa9226cad405054d2b0b40df2d6b28c2f285642d307d4282b2cf53038dc39e

Observation 19660cb0-be24-429e-ab9b-effb73b85fca · outbound

This paper cites Dataset We use a new single speaker rtMRI corpus with simultaneously recorded audio and video.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Dataset We use a new single speaker rtMRI corpus with simultaneously recorded audio and video

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.432731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:41.990420Z digest=sha256:2bc22b5c639642974047e5d035cff0928fff9e54f28ff428b047b4939c919420

Observation 303e5974-c565-4d3d-aa47-6264830f8626 · outbound

This paper cites Phoneme error rate (PER) The overall PER results on the held-out test set are presented in Table 1.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Phoneme error rate (PER) The overall PER results on the held-out test set are presented in Table 1

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.067956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:42.226483Z digest=sha256:afdaf4cf2d92c09c804120a8cfcaf2eb3dc7457196d5064172368fb55788f96e

Observation 8bdbaa1d-acef-4eb8-a05f-8f629f3b9610 · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:55.692447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:42.386227Z digest=sha256:f3e384605716d8e31eca061dbc8d9139609e17ce0f03a763de0cf9c6d968e024

Observation 4f9464f1-031f-4df4-aa14-49e5721d453c · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:55.539004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:42.468817Z digest=sha256:c701ca38bdd124557a4fa49655c62e517eacb3d4c924742b57c299d74491427f

Observation b55612bd-f444-4527-91c0-4121046af81d · outbound

This paper cites The architecture of speech production and the role of the phoneme in speech processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The architecture of speech production and the role of the phoneme in speech processing,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.456393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:43.937557Z digest=sha256:22a9e30817e2f23e8e07bedf209c518c63b37d88510001e96a6c082593bb2c05

Observation 02634c6e-5d51-4923-8e6b-70a56bd6e2f3 · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:54.199624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:44.133282Z digest=sha256:c2307bbdc5eed890129f4ea52ced448cabd2fd2f254521f5f6bea023879176a6

Observation ad08ce4c-bade-4fbb-9f2a-e86bc7862945 · outbound

This paper cites Sensorimotor integration in speech processing: computational basis and neural organization,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Sensorimotor integration in speech processing: computational basis and neural organization,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:55.388476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:42.627165Z digest=sha256:18fd17950b4373dbbc515fe9bf9fcac67cde6b982d457874c6616cc45d12dbbc

Observation cefcdd35-e84e-47e7-83a3-67e13141c8a9 · outbound

This paper cites Perception drives production across sensory modalities: A network for sensorimotor integration of visual speech,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Perception drives production across sensory modalities: A network for sensorimotor integration of visual speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:55.193380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:42.893417Z digest=sha256:f8749e5a109a6bf8f6f666cffa00404b0ed928f6aabf2ab38df554bb64962b46

Observation ba572b53-d737-4711-b1dc-fc85b808dfda · outbound

This paper cites The role of temporal modulation in sensorimotor interaction,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The role of temporal modulation in sensorimotor interaction,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.996114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:43.106505Z digest=sha256:b03837bf131cca6e0aeafc5629413b2fd4db645282be667ce76ee5fbf3fcd4fb

Observation 85e5706b-555a-4ae0-bd48-ce09cd3fa621 · outbound

This paper cites Variability of articulator positions and formants across nine english vowels,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Variability of articulator positions and formants across nine english vowels,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.823327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:43.308671Z digest=sha256:17c0bfae9683369119718a7c1817effd7f2013a931a17bd69502b565e0654de9

Observation 28164159-cfd1-489b-9d53-4391d0aed9db · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:54.596321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:43.525990Z digest=sha256:8dc87deb6c8c9a3755f5e36dbb64f450034c28ac63fcd9bf776582d752cd1809

Observation 3583cc3e-bcdd-47fe-a774-a655e5c0234e · outbound

This paper cites Articulatory phonology: An overview,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Articulatory phonology: An overview,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:43.714834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:43.714834Z digest=sha256:99c30430c614c74e460260b00b59e06f9877b81e428e83665f84e7cbd0686ab6

Observation 1db0d81e-9b23-4164-93c8-43726580f67d · outbound

This paper cites Multimodal representations for syn- chronized speech and real-time mri video processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Multimodal representations for syn- chronized speech and real-time mri video processing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.064691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:45.396716Z digest=sha256:0c6a9be481f625e1ee301415e89d9e5ee094848371fea4eea1fd2f7495de4cc1

Observation 2e3d5334-098b-416f-8c84-7c5c79b3c398 · outbound

This paper cites Towards speech classification from acous- tic and vocal tract data in real-time mri,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards speech classification from acous- tic and vocal tract data in real-time mri,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.867621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:45.545477Z digest=sha256:f15735b148285e54486f3f6d02417851cd8d6dd7849b8389e98197bbf95933e2

Observation e7a1966a-fbb3-4274-88fc-a3c1116e54fb · outbound

This paper cites Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:44.345646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:44.345646Z digest=sha256:b2d11a155bfbff6b1a29bd491a94bcb51fa9e97e4bca2e75516f13fec0e43463

Observation f2a4bde9-cf0a-4d2d-8385-60c6cee01e8c · outbound

This paper cites Robust Audiovisual Speech Recognition Models with Mixture-of-Experts.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Robust Audiovisual Speech Recognition Models with Mixture-of-Experts

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:48.950065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:44.552105Z digest=sha256:1aa77d5141523f107ca016baaf767c0c1692157f9d2fe837ea18fc923a3a305b

Observation 002a3ab1-5feb-492e-a620-76fbaab3bf76 · outbound

This paper cites Speech production real-time mri at 0.55 t,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Speech production real-time mri at 0.55 t,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.935947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:44.741094Z digest=sha256:7773bf5bea152f6b7b237d0a2b1cbb263fc5f0ba59fe52607d4c738dfe7882fd

Observation 84caf0c9-50a7-4904-89ba-54b5dbe728e1 · outbound

This paper cites A multispeaker dataset of raw and reconstructed speech production real-time mri video and 3d volumetric images,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition A multispeaker dataset of raw and reconstructed speech production real-time mri video and 3d volumetric images,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.599113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:44.921151Z digest=sha256:98ec1c79fff5fbc891f133d889b1b1246a4716eac9ac6656d64225475e2be1ce

Observation 943031f2-54a5-4935-baf7-120ab28b8497 · outbound

This paper cites The output of the Conformer is decoded by a single LSTM layer, with a final linear layer used for prediction.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The output of the Conformer is decoded by a single LSTM layer, with a final linear layer used for prediction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.668434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:41.818246Z digest=sha256:39afe8437e9c39f8fbc1edf588f05d4883d35b6e515d347afdd042a5e11d4d13

Observation 53f86a46-9a4c-4b2a-b685-81efa2cfa5ac · outbound

This paper cites Towards Automatic Speech Identification from Vocal Tract Shape Dynamics in Real-time MRI.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards Automatic Speech Identification from Vocal Tract Shape Dynamics in Real-time MRI

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:45.026461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:45.026461Z digest=sha256:d18cb9a2e79e6f9733bf50327d462fdd5d0346da2026ec4bb13857f931562605

Observation 8f0f584e-2244-427e-894a-e5164b58361b · outbound

This paper cites Cnn-based phoneme classifier from vocal tract mri learns embed- ding consistent with articulatory topology.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Cnn-based phoneme classifier from vocal tract mri learns embed- ding consistent with articulatory topology

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.277765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:45.220360Z digest=sha256:b29133702ed503d4fd839cc757655047719c0a062c80f7466347cd9e9f055c89

Observation c7435ec9-7a19-48f8-8a38-cca29cc75173 · outbound

This paper cites Eval- uation of a novel 8-channel rx coil for speech production mri at 0.55 t,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Eval- uation of a novel 8-channel rx coil for speech production mri at 0.55 t,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.456383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:46.700802Z digest=sha256:34c04008953eb84b89cf9bb056f89e82e433754e7b2a1a56a677b0c7d80cf2d5

Observation 3d723bf9-a38a-4dce-b296-52a9c168af75 · outbound

This paper cites Direct articulatory observa- tion reveals phoneme recognition performance characteristics of a self-supervised speech model,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Direct articulatory observa- tion reveals phoneme recognition performance characteristics of a self-supervised speech model,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.651440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:45.684601Z digest=sha256:044bf094c742bde4e5ba43ab6aab55f8dd880d7264f834c7d315315fd9f312e9

Observation 4a4267c1-21c3-48f6-8d48-5aa5897810fa · outbound

This paper cites Usc-timit: A database of multimodal speech production data,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Usc-timit: A database of multimodal speech production data,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.392791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:45.838539Z digest=sha256:792dc9aa85359659c933f62922e20447ec8eb5a119984dcc39f1286eaf2e8b82

Observation 4a5597dc-eb9f-4ab4-835b-3445f4b1a323 · outbound

This paper cites Database of volumetric and real-time vocal tract mri for speech science.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Database of volumetric and real-time vocal tract mri for speech science

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.163254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:46.054670Z digest=sha256:1874fdda2c6a0536a1d8a64806885737253a46b08cb9966c4f80340dbe5c005f

Observation 93f597cf-99dd-42af-a453-18d0d01513ef · outbound

This paper cites Characterization of inter-speaker articulatory variability: A two- level multi-speaker modelling approach based on mri data,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Characterization of inter-speaker articulatory variability: A two- level multi-speaker modelling approach based on mri data,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.879504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:46.187469Z digest=sha256:5b223f49350e9417cb093a47768cb49825f91ec135bb901727b09fce8f6903f5

Observation 61506fae-6150-4f64-9c98-12aa2c30bd3e · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Conformer: Convolution-augmented Transformer for Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.387249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.387249Z digest=sha256:503086a14d277fc8ec8fef5cd540bca9edd782df5250cd99db74fcdd05231c8b

Observation 871471fd-e05a-40c1-99e4-0a104d85efe2 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.480237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.480237Z digest=sha256:04ac32a75aa2bc2942db7d5af5573c3774ebd0602826ae546c1b488f0ee1e526

Observation 6f715e8c-6636-4581-b5c9-a3e2626e3e08 · outbound

This paper cites Simple and Effective Zero-shot Cross-lingual Phoneme Recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Simple and Effective Zero-shot Cross-lingual Phoneme Recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.592923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.592923Z digest=sha256:a76c0f5775507a2bb400014e1bf0dc326945c6f8c6d19141f4de23ebc198b7bb

Observation e9410a79-ce06-4001-8534-f1eba9e98040 · outbound

This paper cites Robust and Efficient Medical Imaging with Self-Supervision.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Robust and Efficient Medical Imaging with Self-Supervision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.585835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.585835Z digest=sha256:42ca1a342ce23a0429607dd7b2b39c5567b545833cfdda98bd109814c6331ff7

Observation 445e246d-f9f9-4650-8dd7-e3230c14340d · outbound

This paper cites State-of-the-art speech production mri protocol for new 0.55 tesla scanners,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition State-of-the-art speech production mri protocol for new 0.55 tesla scanners,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.243216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:46.766700Z digest=sha256:7129bdcd9646e5b110d05279f7f0f30ff22505af2f8b1135cd3ae78ef19a63b9

Observation 93c1d063-cc4a-4c4c-88a2-c60ffcfbd0fb · outbound

This paper cites Announcing the electro- magnetic articulography (day 1) subset of the mngu0 articulatory corpus,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Announcing the electro- magnetic articulography (day 1) subset of the mngu0 articulatory corpus,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:49.890156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:46.910754Z digest=sha256:9712f8e987559ce00c53f68de06354394b5c3ad3084f9c0af3b5eea8d0727079

Observation 75c526be-bcfc-442f-bfee-473c2513d850 · outbound

This paper cites Real Time Speech Enhancement in the Waveform Domain.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Real Time Speech Enhancement in the Waveform Domain

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.014834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.014834Z digest=sha256:d47d826877c19050011fcde9e7200f6da21cce7cca70a483f93122c54752dde3

Observation 1808e0f1-00ca-4fce-bb7b-e83a42107629 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.089969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.089969Z digest=sha256:6d9ce72f7079d631a79af8f93b54544b3d925efd375fdda7b9c2c1fd3fdf9944

Observation 951f1270-1d31-4711-90cc-71a59945f422 · outbound

This paper cites Evidence of Vocal Tract Articulation in Self-Supervised Learning of Speech.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Evidence of Vocal Tract Articulation in Self-Supervised Learning of Speech

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:48.383512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:47.199046Z digest=sha256:0d9a2e078f76d5ff499892871863d6a254851236c2c3fc5abc772116334e5937

Observation 4d4b0c77-17ab-4de1-82bb-8b7beb2f3203 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.361459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.361459Z digest=sha256:2d8ef767aefa7643b8c200e178035fcf90f663e6ce39f96edbe42530c3ddd326

Observation f2bc14b2-647d-4a89-8c33-86c79e04dbc2 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition A Simple Framework for Contrastive Learning of Visual Representations

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.436332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.436332Z digest=sha256:175df87b151564ba28dff21865cf48bd9a8fa1a9097dbd4877fb689dec73497a

Observation 8a6cf77d-0e33-4af0-8982-9e0ea2a20e6d · outbound

This paper cites Visualizing data using t-sne.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Visualizing data using t-sne

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.724939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.724939Z digest=sha256:9ecd9e1d6ea44a4af231c86186add521524769814e0347a73d33f9cd7a7aa807

Observation c9039109-fc06-40f7-9e0d-fc49b769315d · outbound

This paper cites Toward articulatory- acoustic models for liquid approximants based on mri and epg data. part ii. the rhotics,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Toward articulatory- acoustic models for liquid approximants based on mri and epg data. part ii. the rhotics,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:49.529579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:47.898973Z digest=sha256:abe677b0e3391200068ecdf17e3537f3f7efa2454837552f331d9133128fad33

Observation 1251a66c-05b6-4b09-ad8a-62368c93a71e · outbound

This paper cites Articulatory characterization of english liquid- final rimes,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Articulatory characterization of english liquid- final rimes,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:48.043254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:48.043254Z digest=sha256:c887c3d1380140a33ea3f385a93bdd1499f3314fa23d279ced1be83acd25f49c

Observation c6434ec6-3f83-45d8-b41f-4669a6b8313b · outbound

This paper cites The number of Conformer layers was set to 3.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The number of Conformer layers was set to 3

Reference 256

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.210483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:42.111215Z digest=sha256:513760f0081ed38ecef85d3195567da1446b1eae89f9bfc581be6a063d8e13d9

Pith citing papers

Observation 4b43e443-4b31-4fb0-aec0-873a25c92896 · inbound

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition cites this paper.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:49.286208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:39:41.644202Z digest=sha256:98aa9226cad405054d2b0b40df2d6b28c2f285642d307d4282b2cf53038dc39e