Pith. sign in

Paper Citation Record · LEDGER

FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2501.14350.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14350 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:00.417808Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c473e8c4-75c6-40e3-a4de-1e42ab0f59c6 · inbound

Breaking the Barriers of Text-Hungry and Audio-Deficient AI cites this paper.

Breaking the Barriers of Text-Hungry and Audio-Deficient AI FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 148

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:00.417808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:00.417808Z digest=sha256:b78979c96e18151fe2616a6424122fb8b114bc9b2c27b1e7514b4c0f3c756438

Observation 2ef851e9-e40a-4f62-b371-32b5b68955c2 · inbound

Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition cites this paper.

Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:19.311880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:10:19.311880Z digest=sha256:78f8a6a2d78c0f13410099a3c66ed98fe03c0e7acfcc1640ba993a0c658b28fd

Observation 99e48a47-b488-4d56-92f9-75431e2117fe · inbound

Cross-Learning Fine-Tuning Strategy for Dysarthric Speech Recognition Via CDSD database cites this paper.

Cross-Learning Fine-Tuning Strategy for Dysarthric Speech Recognition Via CDSD database FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:18:50.718283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:18:50.718283Z digest=sha256:e6e0f866138dcd310beb512301714e36542f9d74dc49229b59e6d84443e338f6

Observation 7e8d036a-9b51-4646-ae0e-bb34e373869d · inbound

WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation cites this paper.

WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:02.306865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:36:02.306865Z digest=sha256:74a095de20104e750fcbbab246ea02663555002dc92f2391b08764ae49da584d

Observation d967b855-1997-4bd7-af2c-283bdc1ed2aa · inbound

FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations cites this paper.

FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T23:33:04.254815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:33:04.254815Z digest=sha256:e375788ec98d18b71c01f2e2aa92e1f2969afbd79278b1772e1293d79c09ec1e

Observation 3e1af290-f611-4b70-b20c-f2dbcc6ff061 · inbound

SegTune: Structured and Fine-Grained Control for Song Generation cites this paper.

SegTune: Structured and Fine-Grained Control for Song Generation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T08:55:40.021982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:55:40.021982Z digest=sha256:42d453b38a3e9e5f2bb176e0880d9bdb6661f55b671dcee25284527d25b7569a

Observation ec9d05dd-ff1c-4c5a-aa14-4d71a466ce3a · inbound

Existence of the longest arcs for left-invariant three-dimensional contact sub-Lorentzian structures cites this paper.

Existence of the longest arcs for left-invariant three-dimensional contact sub-Lorentzian structures FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-15T13:22:58.005580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:22:58.005580Z digest=sha256:0a4d81ef781e5bd0e1ba2b08745d58f6f8493bfc82e11ad0b2425bd7a31baadd

Observation 82bc387f-687c-4970-9090-2ef3559ae628 · inbound

LLMs and Speech: Integration vs. Combination cites this paper.

LLMs and Speech: Integration vs. Combination FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:45:28.349651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T10:41:22.138517Z digest=sha256:12de02d2382e7dc4f6c83e9058d5e7f04b0454f20f91d7d43d35533bb9fd1639

Observation 061fbd71-59ee-4b54-8252-39203b88512b · inbound

LLMs and Speech: Integration vs. Combination cites this paper.

LLMs and Speech: Integration vs. Combination FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T20:46:17.286576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:46:17.286576Z digest=sha256:62bbcb13c95a2cbfd2b287d1d65b76a8b357bcc37a7b9ac267d43f24cf93019c

Observation 81d205d0-c295-4987-b96a-312761b61059 · inbound

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition cites this paper.

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:41:07.177531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T18:22:08.670559Z digest=sha256:466502b919ede8198c13703394594eae6e0342c230f93258d037de0a52e4df4a

Observation 37d39fea-1205-42a6-8fe9-e7fee4c359f5 · inbound

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models cites this paper.

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:51:09.768681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T16:59:25.973854Z digest=sha256:3904f220c152a8c41e38b342a318d362d4e63d30b89bb35ce7af6f4cc985bedf

Observation 2c4fe42c-885e-41b3-9a43-2b8b69f853a6 · inbound

Dolphin-CN-Dialect: Where Chinese Dialects Matter cites this paper.

Dolphin-CN-Dialect: Where Chinese Dialects Matter FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:16:16.144675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T02:13:42.268229Z digest=sha256:8779f8ff2c1936dd880555147557c50fdb3abd13cca3c15a64b5091fe9ce7a51

Observation 4689b2a8-5cb7-419b-8c22-0623c1cbb594 · inbound

When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai cites this paper.

When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:25:19.816729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-25T04:20:27.456558Z digest=sha256:58247d49e7f1840d714c59fa554aa972732a9bbeeaf9baf035552542134e73c2

Observation 49876449-2efa-40d3-9570-e9d683424d94 · inbound

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis cites this paper.

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T16:23:39.871989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T15:51:21.519785Z digest=sha256:742767ce808a30ebf06ca88fecb03179d0de7429661f89770d1aaa511f0deea3

Observation 0edb7d8a-b937-42d0-a225-51d7743eb8b3 · inbound

Audio-Mind: An Auditable Agentic Framework for Audio Understanding cites this paper.

Audio-Mind: An Auditable Agentic Framework for Audio Understanding FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:13:17.609016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-29T10:03:54.653164Z digest=sha256:4941a2378afde1ade2b031f408caad940eb9f210cd1d46e3a99485dcc32d08e1

Observation 409ae099-8b61-45de-8e73-82819403d8ea · inbound

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation cites this paper.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.699366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:9e20204792e363379275b95d2f4f20a5ca8a8210d59235cd5d622761eea65799

Observation 728c662a-2e2f-40b8-a27a-ce2e589e715b · inbound

Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection cites this paper.

Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:05:01.096878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T18:58:31.318883Z digest=sha256:1888040f65d96a367f9e340194e5edfe2b9e09993f082acf8e30b4562e860b6d

Observation a2f41b5e-9ac4-48b6-af16-2626ff96ffca · inbound

Audio Interaction Model cites this paper.

Audio Interaction Model FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-02T10:46:52.394885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T04:57:05.062465Z digest=sha256:8563c491f017e42dd7af83806e98fc3aa46f7c84840182b9935cb72819b4c1a1

Observation 44bcf260-2066-443b-ac45-90f76c2211ce · inbound

M2S-AVSR: Modality-aware Multi-view Self-supervised Representation for Robust Audio-Visual Speech Recognition cites this paper.

M2S-AVSR: Modality-aware Multi-view Self-supervised Representation for Robust Audio-Visual Speech Recognition FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:17:07.856430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T23:58:21.046264Z digest=sha256:0d99c9674938af0ff4c92db2866905b5781525683a89f247838226099099b04d

Observation 24c79568-2382-4f08-a202-d0419832de3f · inbound

UniVoice: A Unified Model for Speech and Singing Voice Generation cites this paper.

UniVoice: A Unified Model for Speech and Singing Voice Generation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:17:08.272060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T23:56:14.198308Z digest=sha256:6c1647ed60ea22ddc63eae37dd342a0a33ac21d0d287cab8615f9da70276f9a8

Observation b0165a76-7221-4605-89ae-6eee2fa7616e · inbound

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation cites this paper.

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:57:19.608806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T21:14:08.243894Z digest=sha256:21a00c76c90a105daa1d24ff417cf5b26ab115de9a86fcb278d7ca3eaa75eb45

Observation acc687e0-3e9e-45ef-bc4b-ed5ef2f99f97 · inbound

Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving cites this paper.

Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:32.665746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-03T15:17:52.966144Z digest=sha256:ec6a339ad39c999afd47f70e228b40df98673d272ae16d7b967450ebd2b80346