Pith. sign in

Paper Citation Record · LEDGER

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

As of 18 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 2 inbound Pith citation observations for arXiv:2505.15061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15061 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:28:03.066429Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:27:58.440889Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:59:56.411210Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c51a7411-f08c-4fc3-a1fc-95f2ca59c90d · outbound

This paper cites SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:58.440889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:58.440889Z digest=sha256:da787e28429b5b6b5887b043838ff1ec2aec88de515964506830c6aa7d381b72

Observation 17b03289-f42a-4ff1-9120-16048fd7ddba · outbound

This paper cites Speech Quality Estimation: Models and Trends,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Speech Quality Estimation: Models and Trends,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.115842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:59.485160Z digest=sha256:f8ebbefae7c8e9a94ae4b432fddbe37b0a6311df798de4b4692d05dad1fc2e0a

Observation 6cb42b73-4b7c-4e0d-b8d1-ca4a168bef27 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:08.125815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:58.531292Z digest=sha256:5e020ed38588132964a76318b27050f4833725a150eee410af853f925565aed5

Observation acdeb8db-a514-4603-be5a-a3073b640b40 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:07.650385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:59.054106Z digest=sha256:4ae1cc0be27dd041f4fb5caa2ec4281a4eec58b9913c9f719f73aeea06b44ba2

Observation 02d4cf45-0602-4757-924f-479bd72ce4c7 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:07.455618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:59.224815Z digest=sha256:f24131fe909d7ade32e8c44ba3c8ad0d8caceee18deb3978237d0940a55fc66d

Observation 6343ec3b-5f95-4d1f-8882-45224c0b29f6 · outbound

This paper cites Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.671159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.021869Z digest=sha256:88b9f10af0f60ed6efe52ed9e80a8109fdc6689793772b701e8a37069bc1ac24

Observation 6dc804ba-5b49-4974-bdae-b4568eb53419 · outbound

This paper cites Speech quality assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Speech quality assessment,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.310950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:59.348457Z digest=sha256:1d9e4a0e17e8631e7d8a95b36273ec14bc0b26e7574f5a77ca4c9165022de070

Observation 81616ce7-c9a7-4040-8ebb-725042a8ae56 · outbound

This paper cites MOSNet: Deep Learning-Based Objective Assessment for Voice Conversion,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit MOSNet: Deep Learning-Based Objective Assessment for Voice Conversion,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.572907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.299044Z digest=sha256:ae3997090191c5bd275b2d6e73ce9f111a40002cc5d22eca7dd7c8352bd1c2e1

Observation 68135304-236a-4e45-b0f4-f58a04201fbe · outbound

This paper cites A review on subjective and objective evaluation of synthetic speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit A review on subjective and objective evaluation of synthetic speech,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.931041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:59.598535Z digest=sha256:06ddc3af099b505efe68bed400188cd1e8501d36570b8ec584244c8eca003b41

Observation 687b92b2-576d-49de-8a40-44c29a674ede · outbound

This paper cites An Al- gorithm for Intelligibility Prediction of TimeFrequency Weighted Noisy Speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit An Al- gorithm for Intelligibility Prediction of TimeFrequency Weighted Noisy Speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.792909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:59.715325Z digest=sha256:2224ed2390f6c1789347afeee351d3b0fc49a9e98770885bce2cda095f14b9d9

Observation de63a51d-0cb1-4d18-93a7-9c8d608a27fc · outbound

This paper cites Mel-cepstral distance measure for objective speech quality assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Mel-cepstral distance measure for objective speech quality assessment,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:59.834694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:59.834694Z digest=sha256:290ce0ab208412d29875cdfc4b9b99847606adcd8538a0a02dcc5691081eacf8

Observation fdf9fadf-3cfe-4cf3-aabe-06025d193970 · outbound

This paper cites The Voicemos Challenge 2024: Beyond Speech Quality Prediction,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Voicemos Challenge 2024: Beyond Speech Quality Prediction,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.136858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.865900Z digest=sha256:1b280dc7b125f88777abeabbd8c2b1a4f00161a766034d02b0ff7c5c9d63f2c1

Observation 3d832d82-6e22-446a-a8ad-6fc86894596d · outbound

This paper cites AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:00.137013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:00.137013Z digest=sha256:1e4b1e4dddda971cf21e0a6cd12dc4bf81799e8b606f261c0ee895f08ff9324d

Observation 112ea42e-d165-4702-93a4-b64fc554ff7d · outbound

This paper cites Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.985689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.998429Z digest=sha256:a642d8fae55510e513134e7f67d428d6df189429b57239052542f011268adeab

Observation 8e877be0-1bf8-40f4-93e3-24bc5b71b792 · outbound

This paper cites DNSMOS: A Non- Intrusive Perceptual Objective Speech Quality Metric to Evaluate Noise Suppressors,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit DNSMOS: A Non- Intrusive Perceptual Objective Speech Quality Metric to Evaluate Noise Suppressors,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.428593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.477433Z digest=sha256:e87cc8bb2798f61e9f0e5086026ef9996b101eb711b7a2ee88c7c48132492b4c

Observation 2d9b6c6b-c28b-4b88-aa88-7e6ee7546c46 · outbound

This paper cites The VoiceMOS Challenge 2022,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The VoiceMOS Challenge 2022,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.306170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.619961Z digest=sha256:3dad5d9b6445dc89566b9b7c4371d95568ab467886fb2d9e550c5ae9d2e5812e

Observation 2482ec1a-57d7-4a76-97db-7b8ccef0660a · outbound

This paper cites The Voicemos Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Voicemos Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.208074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.754218Z digest=sha256:bbf74f935498872a030629a9fcc702dbc1c59d545d84eca2ffbf15e5d9caf9be

Observation 6e572b51-e5f8-42c1-a471-a64f910c57d0 · outbound

This paper cites Generaliza- tion ability of MOS prediction networks,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Generaliza- tion ability of MOS prediction networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.792093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:01.379773Z digest=sha256:93af6031ef248ce3d4a26efe6d201815342f8ee3cb1c81076e0a1013a4555017

Observation 5c1ce8ff-68a5-48db-97cb-bbf5d3200416 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.056204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:00.932454Z digest=sha256:b00a317363f18adafbfa3e61793038a891f9b0080f9643b1f7d83fb9bccaa503

Observation 0e4c87d8-7628-4902-9837-75848e9a1bba · outbound

This paper cites How do voices from past speech synthesis challenges compare today?.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit How do voices from past speech synthesis challenges compare today?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:01.602221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:01.602221Z digest=sha256:9420a2ed59b36f978a7fe981d8bb38d3e084f59e07fc76102ac6e6972d536df3

Observation d2e3d3b2-f745-46f9-929f-7e5ec3b75e7b · outbound

This paper cites SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.916549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:01.102877Z digest=sha256:ad32dad8136f4503a95b6fef95c2eac073f23a64d04fe194faae15952377c421

Observation 2ca48c2f-7018-4db2-8bf8-a2f6be8817d2 · outbound

This paper cites VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:01.173993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:01.173993Z digest=sha256:7119efb83987d3c9ef71be37f2d6fdd3c66f280ff96c9e46d4c2025a33bf904c

Observation 09396629-52f3-4690-b895-9a39664efe2e · outbound

This paper cites NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.862548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:01.260758Z digest=sha256:f92800aed26d9bee05391b19b929354d139db2efa02cfcce5bc5d357f547bbbf

Observation 6453a74d-c0b7-4f19-b1d8-8f0fac00084d · outbound

This paper cites Self-supervised speech representation learning: A review,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Self-supervised speech representation learning: A review,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:02.135446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:02.135446Z digest=sha256:bdc1080b4b471cc287ede3754dfe1dca1c002a292c71eb5a05a67e4cfd8bcc98

Observation c2b7b22f-4ad8-4c73-91b0-4c51f9f0dc13 · outbound

This paper cites RAMP: Retrieval- Augmented MOS Prediction via Confidence-based Dynamic Weighting,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit RAMP: Retrieval- Augmented MOS Prediction via Confidence-based Dynamic Weighting,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.714739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:01.445768Z digest=sha256:c4a7940cd2ab84b70f258a400d5f4dea1845f2fd880593566c1f81d0c1840e70

Observation 936f9241-75e6-43bb-8913-339eb3930deb · outbound

This paper cites The Singing Voice Conversion Challenge 2023,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Singing Voice Conversion Challenge 2023,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.059446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.269699Z digest=sha256:b32c5ca2e8507c0b7a552ed15526dcd25e0014791b90bd7ce69eaade8f51a363

Observation 6294b4b8-8c7e-4460-af43-589f6412482f · outbound

This paper cites The Kaldi Speech Recognition Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Kaldi Speech Recognition Toolkit,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.645976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:01.743471Z digest=sha256:f2958510e72148530ea3508400b57aa4766fb5d0a262ab96b4367331331c66e9

Observation 86aaaa1a-b7c7-4d64-82a6-0341515bb59b · outbound

This paper cites ESPnet: End-to-End Speech Processing Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit ESPnet: End-to-End Speech Processing Toolkit,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.535217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:01.880783Z digest=sha256:47cf0442b34133365395ec67674390885001d997d8c4a94a7295425857fc1ed6

Observation 07f8f61e-1b87-42df-9ecf-7ae14f9ea84e · outbound

This paper cites An End- To-End Non-Intrusive Model for Subjective and Objective Real- World Speech Assessment Using a Multi-Task Framework,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit An End- To-End Non-Intrusive Model for Subjective and Objective Real- World Speech Assessment Using a Multi-Task Framework,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.422256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.016260Z digest=sha256:ae80da015f437258757ad900e548bc99537f7e0a93151f2d1feaa35dc3e54c5d

Observation 0d1a0ec6-69e4-4859-98c9-6c1d8c39dd40 · outbound

This paper cites Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to- Speech Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to- Speech Toolkit,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.228324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.452285Z digest=sha256:5f57ca7d8e5787efb257f29a444f0c410c165abf1b54f1105182b08cc04ed87a

Observation fd44c759-bdfc-4bd8-b71d-c3ecc88d89b8 · outbound

This paper cites Utilizing Self-Supervised Representations for MOS Prediction,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Utilizing Self-Supervised Representations for MOS Prediction,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.259311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.239871Z digest=sha256:a26703491e98b3232c804fb8cc3193146558a63e9e399d3441cc5b7d20f4cd49

Observation 23d6e7d9-910d-4aeb-bd38-8534d5b49be5 · outbound

This paper cites On the NISQA dataset, on average, the WavLM large model [33] and the XLS-R 1b model [34] achieved the best and second best scores on the Sys MSE and Sys SRCC metrics, respectively.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit On the NISQA dataset, on average, the WavLM large model [33] and the XLS-R 1b model [34] achieved the best and second best scores on the Sys MSE and Sys SRCC metrics, respectively

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.871031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:58.911703Z digest=sha256:9a4370b1b7c2bb399c0100c51a18615b503eebd0b4836acbf8768b95b98e254d

Observation d9a2b204-f6f0-48eb-92f0-453313c43342 · outbound

This paper cites MB- NET: MOS Prediction for Synthesized Speech with Mean-Bias Network,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit MB- NET: MOS Prediction for Synthesized Speech with Mean-Bias Network,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.888066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.282769Z digest=sha256:e3be6f2988fabba09711ca47800c8611c6a3c34b1c495825e6a9b82bd8433086

Observation ccae44fd-d877-47cd-ac64-edd7305536fd · outbound

This paper cites LDNet: uni- fied listener dependent modeling in MOS prediction for synthetic speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit LDNet: uni- fied listener dependent modeling in MOS prediction for synthetic speech,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.675473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.320774Z digest=sha256:f2424f3c762f4898d3fa15d97d78d05cca6f9d6e67ed4132298ec8c9ca4213be

Observation 41da7b5e-66d8-4ff5-b61e-2f57760076c8 · outbound

This paper cites Alignnet: Learning dataset score align- ment functions to enable better training of speech quality estima- tors,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Alignnet: Learning dataset score align- ment functions to enable better training of speech quality estima- tors,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.443873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.393441Z digest=sha256:42b2a188f3eeda202ae44d2bfcce7117a7bcc7f1c72f79b6634beb5d1ee881ff

Observation b1532b15-934d-4afe-9c62-277a48251fe8 · outbound

This paper cites VoxSim: A perceptual voice similarity dataset,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit VoxSim: A perceptual voice similarity dataset,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.451526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.779806Z digest=sha256:0edc54136890928052574d851af79d307ce560b73985e9f1b743d6b70ac30ce0

Observation 87b739b1-812e-4f75-9d76-43b334dc497b · outbound

This paper cites A Large- Scale Evaluation of Speech Foundation Models,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit A Large- Scale Evaluation of Speech Foundation Models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.985643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.508013Z digest=sha256:b7d6d5f5b4377e314fcbdda50f0ec2cbf4a7f5fcc7d9a53be8aad92ec7c69361

Observation fc11a089-7fa0-45a6-b6b9-c2ddf825a917 · outbound

This paper cites data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.805292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.558645Z digest=sha256:b6d6914bb33970cc7c4233c6de61a7a96cffcffd3ef3be2b62f4f212f2c93052

Observation e82ef2e9-c752-404e-98af-c00c7a450908 · outbound

This paper cites WavLM: Large-scale self-supervised pre-training for full stack speech processing,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit WavLM: Large-scale self-supervised pre-training for full stack speech processing,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.662765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.605498Z digest=sha256:25eba63d91107c88173de2ed4831efc55cde4dc4004f85f978647c935a375155

Observation 6f803eb1-b3c2-4419-912d-140668090ed4 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Rep- resentation Learning at Scale,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit XLS-R: Self-supervised Cross-lingual Speech Rep- resentation Learning at Scale,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.596764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.647836Z digest=sha256:5ad431707aee4ecc55c983d646d45d60b3de4f58b88f104317bed293ca8e1d18

Observation bbc1f6f4-160d-4abb-816d-fa9e27ff09f0 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.514832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.696851Z digest=sha256:8c2fcb2ad9d6d7fe141c71ff9db56598ab0ec9144dc2d211396eb5c2e55d8e28

Observation e3f0447e-2af3-4381-a9c1-b093fdf438f9 · outbound

This paper cites PAM: Prompting Audio- Language Models for Audio Quality Assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit PAM: Prompting Audio- Language Models for Audio Quality Assessment,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.392977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:02.943151Z digest=sha256:c7ee31771e96ac6367f85904c7b1039ba2ce131905a1c653889f9acce335e5ba

Observation 72a379e7-d0ab-465d-ad96-9a05c8242d2d · outbound

This paper cites Audio Large Language Models Can Be Descriptive Speech Quality Evaluators,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Audio Large Language Models Can Be Descriptive Speech Quality Evaluators,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.305134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:28:03.066429Z digest=sha256:cf36d7cd5fcb6f7cba9e95beaa075ddb85685d48719b860253445db4f25a6168

Observation 964e5b53-eb6d-44da-b0a7-50f33ac1377f · outbound

This paper cites Each sample was rated by 8 distinct listeners.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Each sample was rated by 8 distinct listeners

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.963154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:27:58.708724Z digest=sha256:f98aa4f5f1a9cdf43bb81aeadd3b6ecbfc71017512d80e7e1a1fd2c974533ffc

Pith citing papers

Observation c51a7411-f08c-4fc3-a1fc-95f2ca59c90d · inbound

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit cites this paper.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:58.440889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:58.440889Z digest=sha256:da787e28429b5b6b5887b043838ff1ec2aec88de515964506830c6aa7d381b72

Observation 7b40b310-2311-414f-955d-0e1c5e80c18d · inbound

Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment cites this paper.

Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:59:56.456405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:59:55.643986Z digest=sha256:def90eb962c75eff1144df2bdfcc1ff8f2c9e6a452157f78b14f0080e1b35a2f