Pith. sign in

Paper Citation Record · LEDGER

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling

As of 8 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2607.03670.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03670 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:48:21.716770Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved54
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1e83528-89fe-4745-8984-5e6ce944bd51 · outbound

This paper cites Benchmarking Children's ASR with Supervised and Self-supervised Speech Foundation Models.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Benchmarking Children's ASR with Supervised and Self-supervised Speech Foundation Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:d790349863d315b3dfb10dd10e5bf786c167a0e96662fde37b2a64c72ad1547a

Observation d9781e07-0864-4166-86d4-8601eed7fc10 · outbound

This paper cites Automatic speech recognition (asr) systems for children: A systematic literature review,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Automatic speech recognition (asr) systems for children: A systematic literature review,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:a2b31fe9d4373d46be1baa728fbaf5ec78010dec33f9a53ce4681d9dfbaba33a

Observation 5076ea0f-7273-4e39-aac2-8cd05796920b · outbound

This paper cites Qwen3-ASR Technical Report.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Qwen3-ASR Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:b73834d2ece278fa946ebc0c47db98594b30369bfbbc9d7272d7caa05617cd56

Observation 21c98df7-0684-4e0a-8545-284b905b3fb4 · outbound

This paper cites Canary-1b-v2 & parakeet-tdt- 0.6b-v3: Efficient and high-performance models for multilingual asr and ast,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Canary-1b-v2 & parakeet-tdt- 0.6b-v3: Efficient and high-performance models for multilingual asr and ast,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:8a922610af6451499b9376d39c083f9f1e6fca9443d06c3ece840632c87ba2c0

Observation 43383fbd-3be8-4038-890a-e661e56457e8 · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Robust speech recognition via large-scale weak supervi- sion,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:ecea6871869688847e47af0b1f9a65625226c5d2c276375716c3de262738ba8f

Observation fb4483b9-0404-4f43-9ee8-1a86a9349e0d · outbound

This paper cites Exploring the effect of differences in the acoustic correlates of adults’ and children’s speech in the context of automatic speech recognition,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Exploring the effect of differences in the acoustic correlates of adults’ and children’s speech in the context of automatic speech recognition,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:cd6f6509211c84e6a884e8257b3013e10044f243c102757bf46f360ebfd4ce74

Observation e8f9cecb-fad6-496a-9f97-107ca88d23a2 · outbound

This paper cites A review of asr technologies for children’s speech,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling A review of asr technologies for children’s speech,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:9d95bcd98f74031b3b398188f6c68bc64f62bc9790e3813ed851fd938e068681

Observation 245cebf8-4e02-494f-aa1c-8762313b6bfd · outbound

This paper cites Transfer learning from adult to chil- dren for speech recognition: Evaluation, analysis and recommendations,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Transfer learning from adult to chil- dren for speech recognition: Evaluation, analysis and recommendations,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:90c928a08d524e2fef3aa74832675ac67fe825c3d64f01cfd79d05f22653af3e

Observation 1dd294de-32de-4ae9-8a6b-ad91b48786b3 · outbound

This paper cites Speech production variability in fricatives of children and adults: Results of functional data analysis,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Speech production variability in fricatives of children and adults: Results of functional data analysis,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:14095f59363777a392bdbe6ff220db3c6de7e3d9b5cd80a18fe250806c15018c

Observation b6626d5b-7de1-4307-a50e-80fdc0395a20 · outbound

This paper cites Stop consonant voicing and intraoral pressure contours in women and children,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Stop consonant voicing and intraoral pressure contours in women and children,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:0172add5875bff9a9316bdd0d06c5156528bbc7ebc70f096c9dd61d0914ece00

Observation bc952c76-0c54-4a3d-83b8-615e0989b7ae · outbound

This paper cites Analysis of children’s speech: Duration, pitch and formants,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Analysis of children’s speech: Duration, pitch and formants,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:5b345ec98d6edd420af7a3969b9fefa7eb273fa103798a1fb3686f4763488e11

Observation 6697e24d-39ee-4a51-90ac-3de8bc392b4a · outbound

This paper cites Acoustics of children’s speech: Developmental changes of tem- poral and spectral parameters,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Acoustics of children’s speech: Developmental changes of tem- poral and spectral parameters,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:8df3bbf85f6a7e6460057e8dba0b67cc3c460c60a169255e06dde8db67398087

Observation d3d5cd9a-f03f-4cef-bbd5-4bc01002ae63 · outbound

This paper cites V owel acoustic space development in children: A synthesis of acoustic and anatomic data,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling V owel acoustic space development in children: A synthesis of acoustic and anatomic data,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:0fc0dfcfa9b1be5d06e67ba802b9b1723db2e950f0dfa83ebcf3a0707728f765

Observation c837983b-1073-44e2-be36-b5ad47ff419c · outbound

This paper cites Ticl: Text- embedding knn for speech in-context learning unlocks speech recogni- tion abilities of large multimodal models,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Ticl: Text- embedding knn for speech in-context learning unlocks speech recogni- tion abilities of large multimodal models,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:c96b90fdd3e37545f93ca9b3e8937c639a5344d32b000ac6312d6adc9ba9b969

Observation 757425fc-85d3-4761-b5b5-3bfacae61d5c · outbound

This paper cites Ticl+: A case study on speech in-context learning for children’s speech recognition,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Ticl+: A case study on speech in-context learning for children’s speech recognition,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:cd697b95cdb75402a16eb75cc0a9c2efe1ef4ed203e000443cb7866a34981b9c

Observation 6ca85725-1e1f-4b5c-9fc7-3a6a155b0768 · outbound

This paper cites MetaSICL: Adapting Audiroty LLM via Meta Speech In-Context Learning.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling MetaSICL: Adapting Audiroty LLM via Meta Speech In-Context Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:370b469d773cc80e95334b2439fd6d6aaf23c8a4e7178a40139dee60648723e2

Observation 64788853-d44f-4683-9f8e-9a7dc9575e1c · outbound

This paper cites FSA-GRPO: Teaching Auditory LLMs to Use Few-shot Demonstrations.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling FSA-GRPO: Teaching Auditory LLMs to Use Few-shot Demonstrations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:0a564d7744ab47aa03a20f9fa46f2e548be48fdd5d0f131ad440fc3ce715f37a

Observation cf1645f8-3434-461a-8948-80afd9fc8bd5 · outbound

This paper cites Benchmarking Training Paradigms, Dataset Composition, and Model Scaling for Child ASR in ESPnet.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Benchmarking Training Paradigms, Dataset Composition, and Model Scaling for Child ASR in ESPnet

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:05fdf5acb5d927f368318fb1f187edea02beb1a4337ee973ec55febbd32c797f

Observation 9d7d82cb-5f36-4840-b1d1-07cd311cce38 · outbound

This paper cites Age-Aware Adapter Tuning for Children's Speech Recognition.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Age-Aware Adapter Tuning for Children's Speech Recognition

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:4cf7bf47a9ac7cb9af98587fdfdca93a49b33b37cac8a6fad5918e7d888922d7

Observation 1f356347-e8f2-4821-a670-f2a2dec7f103 · outbound

This paper cites Adapta- tion of whisper models to child speech recognition,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Adapta- tion of whisper models to child speech recognition,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:ecef4691586900fd7c5bf293a79c31bc22ff9cf8970961f3d79066724db14f7d

Observation a5fa2184-e235-4884-bdf9-981fbaa5269b · outbound

This paper cites Kid- whisper: Towards bridging the performance gap in automatic speech recognition for children vs. adults,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Kid- whisper: Towards bridging the performance gap in automatic speech recognition for children vs. adults,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:ccc5e691303bc7a3281619ec081119e23196cee53464be361e41a60d2d96e065

Observation d364dd24-520a-4802-aed7-ba9d68c942ff · outbound

This paper cites Sparsely shared lora on whisper for child speech recognition,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Sparsely shared lora on whisper for child speech recognition,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:c11510f0e90315b13d96e04c96981324b9deb12ab42f421ba36aa5f7a1f922e8

Observation e6eea29a-0f02-4711-b49e-46504bfec4e2 · outbound

This paper cites Mind the shift: Using delta ssl embeddings to enhance child asr,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Mind the shift: Using delta ssl embeddings to enhance child asr,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:766eb5f9a60401db00ba755fc2c7c87642400a5a6349886b82c9eea490e54758

Observation d0cf11e4-c80b-4f40-9872-d593be6c74f5 · outbound

This paper cites DRAFT: A Novel Framework to Reduce Domain Shifting in Self-supervised Learning and Its Application to Children's ASR.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling DRAFT: A Novel Framework to Reduce Domain Shifting in Self-supervised Learning and Its Application to Children's ASR

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:d4f2275c0ba735a3e6adbe907867ab3ef97e1289644784f647e4574fc8cd4384

Observation e27b6f5a-3719-422d-8a9a-c9b369dcee6d · outbound

This paper cites Towards better domain adaptation for self-supervised models: A case study of child asr,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Towards better domain adaptation for self-supervised models: A case study of child asr,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:9c7eb20fea36775ef40af9dec85882ef42d4e991389d52f7b1113f8cc2db7a9b

Observation c4ec6e42-ef45-49c5-adf3-78af585a0e37 · outbound

This paper cites A wav2vec2-based experimental study on self-supervised learning methods to improve child speech recognition,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling A wav2vec2-based experimental study on self-supervised learning methods to improve child speech recognition,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:bc2782d3000d8cc2d280dc7b2f3a5ce5ab1b764685369903d429701a1d9a6284

Observation cc6dff50-6964-4903-89e5-d14399fb5ae7 · outbound

This paper cites MacWhinney,The CHILDES Project: Tools for Analyzing Talk, 3rd ed.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling MacWhinney,The CHILDES Project: Tools for Analyzing Talk, 3rd ed

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:1e84db51ab3a4898d0f9c5c0d641b249285624e6f2c77596411112a2d1feaec6

Observation 0372b923-6ddf-4fb5-b3f0-4a49e9ceb7e2 · outbound

This paper cites Montreal Forced Aligner: Trainable text-speech alignment using Kaldi,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Montreal Forced Aligner: Trainable text-speech alignment using Kaldi,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:903a8086df424d9f53f18f8ddd16da1b6e30fb87527e408a6acfb8332cd2cb9d

Observation 10db3fde-f1e0-4ce7-a1ea-aa0756afa508 · outbound

This paper cites Scaling Speech Technology to 1,000+ Languages.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Scaling Speech Technology to 1,000+ Languages

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:5c073b99bb5f350cbd496d921154552b80bc68deb68b99560c2e00403b3d036d

Observation fe52b86c-87a6-441f-ac0a-5789e6da6947 · outbound

This paper cites A recursive algorithm for the forced alignment of very long audio segments,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling A recursive algorithm for the forced alignment of very long audio segments,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:db1bcccc29b086750f10fda58a0e883103e90f9d82c039e4829171e38224f936

Observation d2054911-28b0-4924-85da-c28ed5d8cb28 · outbound

This paper cites CTC- Segmentation of large corpora for german end-to-end speech recogni- tion,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling CTC- Segmentation of large corpora for german end-to-end speech recogni- tion,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:c1e259d93404a9b583efd5b1fc7664921686a54bc3271bc212eefe14557b60f1

Observation c0ad31d2-13b6-4412-81a7-215ef6771313 · outbound

This paper cites ALISA: An automatic lightly supervised speech segmentation and alignment tool,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling ALISA: An automatic lightly supervised speech segmentation and alignment tool,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:903a58de8f8e856e79188073ad8983f30e158319fc747daf1d75c7f43d288a82

Observation 2868958a-e78a-4300-a523-ed68dca3a9d1 · outbound

This paper cites A linear memory CTC-based algorithm for text-to-voice alignment of very long audio recordings,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling A linear memory CTC-based algorithm for text-to-voice alignment of very long audio recordings,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:86951253ef7eba5e0732832ab11f4021e3bd631f8c7e5f82d4479b7e64c33b65

Observation 361e5e16-74c5-47ae-8b3e-ea557c63ff95 · outbound

This paper cites WhisperX: Time-Accurate Speech Transcription of Long-Form Audio.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling WhisperX: Time-Accurate Speech Transcription of Long-Form Audio

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:d871377908e1c4d6399a015659793609cd1672b435c0388d9899cc1b4b63b6be

Observation 137f1b7d-6156-464a-9c22-2fe1155205f8 · outbound

This paper cites FASA: a Flexible and Automatic Speech Aligner for Extracting High-quality Aligned Children Speech Data.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling FASA: a Flexible and Automatic Speech Aligner for Extracting High-quality Aligned Children Speech Data

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:1e886ffa5f5013e84390c8ae5e181004318b2360fe29da65dae98d7acb02064c

Observation 4ab0584c-1527-4ab4-9135-b0aa1dd0ff87 · outbound

This paper cites Kids: a database of children’s speech,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Kids: a database of children’s speech,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:e6e31781b6b827df6ffbf7fa9de2e5320f9ddadbf7a29bbd09c64954609f53ba

Observation c2845150-f698-4698-8fb6-0caf0a69231c · outbound

This paper cites The ogi kids’ speech corpus and recognizers,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling The ogi kids’ speech corpus and recognizers,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:34c211f5988132c78b4637f48da8cc765c222fe942f7b15cf4a83f7a8f213736

Observation e77428ba-619e-4700-8a0d-b00b4983d70a · outbound

This paper cites Multi-task learning for speech attribute detection of children’s speech.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Multi-task learning for speech attribute detection of children’s speech

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:cf6a9d823ef4807c94f8970b053d23135359c4be556974e2ff46959fd0d2e07a

Observation 6677960c-20d6-42fb-a4bf-baad9bcd6612 · outbound

This paper cites Tidigits,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Tidigits,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:808da8b5c5a3482c9f3a490cf4cbd5cf8f1b56dd7dd5969b68bd65fdc46b659d

Observation c918ce5c-cb6f-4dd7-96e9-1fd550dfa13b · outbound

This paper cites Diagnostic accuracy of sentence recall and past tense measures for identifying children’s language impairments,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Diagnostic accuracy of sentence recall and past tense measures for identifying children’s language impairments,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:281306d1135e9dbf9cc769c12abbc0b5cf73bebb99167f0e9b788342b065ffbc

Observation 56738040-fe85-4093-8b05-ca2823401a80 · outbound

This paper cites My science tutor (myst)–a large corpus of children’s conversational speech,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling My science tutor (myst)–a large corpus of children’s conversational speech,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:34668036b3edbdf3a8a5b302edfa5e6a45d9e4e9cda78b908fd7bce1da4e6651

Observation b1fee90d-852a-4d8b-a88a-5590c11d4cd0 · outbound

This paper cites Children’s speech recognition with application to interactive books and tutors,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Children’s speech recognition with application to interactive books and tutors,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:b222d8be66995f5897b0057d2c5b80868bfade58939d6b90cc6429a9aeb7ed17

Observation 25740fac-9170-42c6-b108-189a31d876f2 · outbound

This paper cites Word-minimality, epenthesis and coda licensing in the early acquisition of english,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Word-minimality, epenthesis and coda licensing in the early acquisition of english,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:35d0b913020e3d0be4776358a2fbb5ca01098163e88ae46a5e5c3347ac67baa9

Observation f8b4f434-7997-4a46-a68b-d5bb3578039b · outbound

This paper cites The pf-star british english childrens speech corpus,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling The pf-star british english childrens speech corpus,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:78c86d18772f67b26aa7f06ebca733fa5d481c58faf03c678be5195d852312ef

Observation a7c23b33-c117-4e8d-9e9b-a0edb811bcc9 · outbound

This paper cites The pf star children’s speech corpus,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling The pf star children’s speech corpus,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:397f3655040e8a0428b7cc9a2beb01f43a003f5a51a004309557deb9efd7464b

Observation f43aae98-6b5d-4792-994d-cab74be4ac31 · outbound

This paper cites Tball data collection: the making of a young children’s speech corpus,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Tball data collection: the making of a young children’s speech corpus,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:ee3e904c828fce81276e5004dc6ebf7abac2c08520bf2c35c19961d257e298ae

Observation 23982a50-fc91-4a06-a6df-2ee947950e9f · outbound

This paper cites The kaldi speech recognition toolkit,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling The kaldi speech recognition toolkit,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:b18e5801e5777694f599cef69d0686878dd2095c9fe3d1fde1e2a6b09647dc65

Observation ae4f7b59-57e0-437d-a786-9881aaf174bb · outbound

This paper cites Automation of Language Sample Analysis,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Automation of Language Sample Analysis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:509bf855b9ca6796fb1a80260bfa70430c78b4e5dfce79628696d70b43a35fdf

Observation 24ef8e9f-a56f-4161-b95b-b5ae0b6e5b7a · outbound

This paper cites Sailalign: Robust long speech-text alignment,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Sailalign: Robust long speech-text alignment,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:b5f3bcf038c1471ac28345f2bea7c648cc57608c0607485427be527178527cb3

Observation 70e2ea42-bee0-4fc8-a0e2-9d903dfdb965 · outbound

This paper cites Automatic long audio alignment and confidence scoring for conversational arabic speech,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Automatic long audio alignment and confidence scoring for conversational arabic speech,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:6795c67402ce1a24c33c854a060ab6b1a29a3dd172c101627eec0a80803a92ed

Observation 419608c9-df5b-4d43-a2be-1219c68c41fb · outbound

This paper cites A post-processing system to yield reduced word error rates: Recognizer output voting error reduction (rover),.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling A post-processing system to yield reduced word error rates: Recognizer output voting error reduction (rover),

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:1be2bf3bc5568db11fed2135b8d5566c7c55610c3b171f69ccbd3ce5eed5b75a

Observation 7dc3eb10-507c-4b9c-8309-8f6895fa5132 · outbound

This paper cites Ensemble methods in machine learning,.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Ensemble methods in machine learning,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:885c758f87a46115a93ab0c8fab645ef9eef8e277a0e125edb9e8d91ee63d57c

Observation 11badae8-27a5-4036-ab28-cf3e05c5420b · outbound

This paper cites Efficient Sequence Transduction by Jointly Predicting Tokens and Durations.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Efficient Sequence Transduction by Jointly Predicting Tokens and Durations

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:b8f59863eadedf4f2968cb5024dc64e96f7d9c4fde2560a5346a3d4846d6f807

Observation df653451-76cf-4596-bad0-1ab5861d3ae5 · outbound

This paper cites Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities.

CHILDES-Aligned: A Curated Children's Speech Dataset via Multi-Model Timestamp Ensembling Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-12T00:48:21.716770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:48:21.716770Z digest=sha256:741a87d1d120470a6a8fad85532c55f45a6f67a81594769630effbe97ec44379

Pith citing papers

No inbound Pith citation observations are available.