Pith. sign in

Paper Citation Record · LEDGER

FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2501.14350.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14350 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:00.417808Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c473e8c4-75c6-40e3-a4de-1e42ab0f59c6 · inbound

Breaking the Barriers of Text-Hungry and Audio-Deficient AI cites this paper.

Breaking the Barriers of Text-Hungry and Audio-Deficient AI FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 148

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:00.417808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:00.417808Z digest=sha256:7ca65a3575f63eeb549c095e848e2a150383355574627152b4996aab76de15b0

Observation 2ef851e9-e40a-4f62-b371-32b5b68955c2 · inbound

Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition cites this paper.

Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:19.311880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:10:19.311880Z digest=sha256:a7f9345590337cd73759a79e725c0e8c631c9ffb91770a954d806c4a9b00773c

Observation 99e48a47-b488-4d56-92f9-75431e2117fe · inbound

Cross-Learning Fine-Tuning Strategy for Dysarthric Speech Recognition Via CDSD database cites this paper.

Cross-Learning Fine-Tuning Strategy for Dysarthric Speech Recognition Via CDSD database FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:18:50.718283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:18:50.718283Z digest=sha256:20aa300a5c565d433b0a9a89ae9efaa91e530dedc3bf7281556ecf99beed917b

Observation 7e8d036a-9b51-4646-ae0e-bb34e373869d · inbound

WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation cites this paper.

WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:02.306865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:36:02.306865Z digest=sha256:29532cb4c57760ce0328a606c0ce1da6b1ffb8ab844fa881419dc4bd42234886

Observation d967b855-1997-4bd7-af2c-283bdc1ed2aa · inbound

FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations cites this paper.

FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T23:33:04.254815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:33:04.254815Z digest=sha256:129325d052cbcb66e169759659b88bce36c3d300c3a77bf2f8482442fc926947

Observation 3e1af290-f611-4b70-b20c-f2dbcc6ff061 · inbound

SegTune: Structured and Fine-Grained Control for Song Generation cites this paper.

SegTune: Structured and Fine-Grained Control for Song Generation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T08:55:40.021982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:55:40.021982Z digest=sha256:e62b17f37eee0ffc118e0e8a94a8cc586d20c5a51cd27aa2b13940cb6c0131dc

Observation ec9d05dd-ff1c-4c5a-aa14-4d71a466ce3a · inbound

Existence of the longest arcs for left-invariant three-dimensional contact sub-Lorentzian structures cites this paper.

Existence of the longest arcs for left-invariant three-dimensional contact sub-Lorentzian structures FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-15T13:22:58.005580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:22:58.005580Z digest=sha256:a11d127067093d55ba59086ac633de9ae20220f7adb780abfe668c65d22411e0

Observation 82bc387f-687c-4970-9090-2ef3559ae628 · inbound

LLMs and Speech: Integration vs. Combination cites this paper.

LLMs and Speech: Integration vs. Combination FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:45:28.349651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T10:41:22.138517Z digest=sha256:2544833b3f56acee2e173cffbadc5e045c4fe51161c22b3608444f36d8284762

Observation 061fbd71-59ee-4b54-8252-39203b88512b · inbound

LLMs and Speech: Integration vs. Combination cites this paper.

LLMs and Speech: Integration vs. Combination FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T20:46:17.286576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:46:17.286576Z digest=sha256:d62a9f14dd1d5d0bdefb7e434e4bb480e43c35a21fb13db78378f912514cba4d

Observation 81d205d0-c295-4987-b96a-312761b61059 · inbound

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition cites this paper.

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:41:07.177531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:22:08.670559Z digest=sha256:268c4fbee85de56f5fbaeb32d8be3d6d8b24da05d0a8cf3fedb37edc085fa527

Observation 37d39fea-1205-42a6-8fe9-e7fee4c359f5 · inbound

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models cites this paper.

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:51:09.768681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T16:59:25.973854Z digest=sha256:1948710ccd0ffbf4d21b3f9419a4dd62bce6d5bdff640e63e7101a1f959a6f67

Observation 2c4fe42c-885e-41b3-9a43-2b8b69f853a6 · inbound

Dolphin-CN-Dialect: Where Chinese Dialects Matter cites this paper.

Dolphin-CN-Dialect: Where Chinese Dialects Matter FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:16:16.144675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:13:42.268229Z digest=sha256:15ce1ffa4f99e1b9ed9272b764eab2922cddb1fb2a758b03964fd26a10a783c4

Observation 4689b2a8-5cb7-419b-8c22-0623c1cbb594 · inbound

When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai cites this paper.

When Youth Enter the Algorithmic Wild: Discovering and Understanding Potentially Harmful Teen Videos on Douyin and Kwai FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:25:19.816729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:20:27.456558Z digest=sha256:0920b71d4ed5eca4b06e07d4ef52a6931cc2f2a6d9300e3580580a6426cbfa08

Observation 49876449-2efa-40d3-9570-e9d683424d94 · inbound

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis cites this paper.

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T16:23:39.871989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T15:51:21.519785Z digest=sha256:4ed356e1bb315f7cc3497b3b86eacd4f12e1bd07ca99ed2820ca918cf206e29e

Observation 0edb7d8a-b937-42d0-a225-51d7743eb8b3 · inbound

Audio-Mind: An Auditable Agentic Framework for Audio Understanding cites this paper.

Audio-Mind: An Auditable Agentic Framework for Audio Understanding FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:13:17.609016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T10:03:54.653164Z digest=sha256:0c9e3c8cd9780cd32ecc25df30349926106cd77c13dd59de0139211f9e07101a

Observation 409ae099-8b61-45de-8e73-82819403d8ea · inbound

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation cites this paper.

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.699366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:30:11.718647Z digest=sha256:c211af7d20e1bc2e830d4454a08ca2253283d49ab5add65443d84980ce46725e

Observation 728c662a-2e2f-40b8-a27a-ce2e589e715b · inbound

Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection cites this paper.

Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:05:01.096878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T18:58:31.318883Z digest=sha256:8051af0bf9afed6ee04aac51d5f2c88b586419053ba51bf4148221830bcc2dee

Observation a2f41b5e-9ac4-48b6-af16-2626ff96ffca · inbound

Audio Interaction Model cites this paper.

Audio Interaction Model FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-02T10:46:52.394885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T04:57:05.062465Z digest=sha256:4dc37f078521ddd82b63b371037248e99c2acf3dffe24e29695cf21b1280ebeb

Observation 44bcf260-2066-443b-ac45-90f76c2211ce · inbound

M2S-AVSR: Modality-aware Multi-view Self-supervised Representation for Robust Audio-Visual Speech Recognition cites this paper.

M2S-AVSR: Modality-aware Multi-view Self-supervised Representation for Robust Audio-Visual Speech Recognition FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:17:07.856430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T23:58:21.046264Z digest=sha256:c6fba7709e18201ce6d7b4cba58608d974e70efaf1f3f070ae9fb7b4d9ab955b

Observation 24c79568-2382-4f08-a202-d0419832de3f · inbound

UniVoice: A Unified Model for Speech and Singing Voice Generation cites this paper.

UniVoice: A Unified Model for Speech and Singing Voice Generation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:17:08.272060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T23:56:14.198308Z digest=sha256:c6c7354b79fcdda3cffb4d276b539f98f9befc572daac1c89c893103d63ede0f

Observation b0165a76-7221-4605-89ae-6eee2fa7616e · inbound

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation cites this paper.

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:57:19.608806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T21:14:08.243894Z digest=sha256:341944f37e067da90a8f79884aae58cf9eecef584d09f06d658c8de3146e0c45

Observation acc687e0-3e9e-45ef-bc4b-ed5ef2f99f97 · inbound

Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving cites this paper.

Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving FireRedASR: Open-Source Industrial-Grade Mandarin Speech Recognition Models from Encoder-Decoder to LLM Integration

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:32.665746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T15:17:52.966144Z digest=sha256:81190bf948f21eb6afc1a3c6af99e94f99e2f67442916c8a4a198b189371d95c