Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:58:10.657355Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 30 inbound Pith citation observations for arXiv:2506.09344.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:58:10.657355Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:11.838092Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
45 of 45 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 97c7ba52-f118-4652-bb40-c2f3d27a5a33 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3067baf-60f0-4653-b297-95bb30c869df · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Qwen2-Audio Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 682af26a-c8b6-492e-ba35-32edade64f8a · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Image-to-Markup Generation with Coarse-to-Fine Attention
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 374d689e-3863-4d1c-8886-b7d8710bd148 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Chaoyou Fu, Y uhan Dai, Y ondong Luo, Lei Li, Shuhuai Ren, Renrui Zhang, Zihan Wang, Chenyu Zhou, Y unhang Shen, Mengdan Zhang, et al
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 007423dc-82e9-40c5-be88-0607471da412 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70a0d441-835b-48cb-9407-933b084066dc · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation FunASR: A Fundamental End-to-End Speech Recognition Toolkit
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa731db0-4fbb-40fc-9444-209b2dd5062e · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation M2-omni: Advancing Omni-MLLM for Comprehensive Modality Support with Competitive Performance
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608617b7-2ee3-443b-8385-e8a2114b71e8 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 852eed35-e376-4708-b0b5-c93c3bc3d385 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Ming-Lite-Uni: Advancements in Unified Architecture for Natural Multimodal Interaction
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c10826b0-b811-4ac8-848d-710487f0c234 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation EgoTaskQA: Understanding Human Tasks in Egocentric Videos
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d6e47b9-39d2-47d4-8707-2ace687eb27f · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation GeomVerse: A Systematic Evaluation of Large Models for Geometric Reasoning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45760fc5-91b9-4438-85c6-7bc341c48d89 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Every FLOP Counts: Scaling a 300B Mixture-of-Experts LING LLM without Premium GPUs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 241556af-9e91-45bc-b9e4-46a607fba19f · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d6edeb-42af-4c5b-a1e8-f2372a242aa4 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation doi: 10.18653/v1/2022.findings-acl.177.https://aclanthology.org/2022.findings-acl.177/
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ac6a8a02-1cf9-4fc5-ba9e-97f7ff1cced6 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46ecc3aa-3986-4892-9349-3a8f625cc6e7 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2333b94-35d2-4b2f-a9d0-565eff4f8b12 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eafccbb-e235-443d-8afd-81a331039c9e · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation MLS: A Large-Scale Multilingual Dataset for Speech Research
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f69c249c-b6b3-444c-90c8-e4e86784bdca · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf3a7020-ca82-45ec-a2de-6bcdbce815e5 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation CinePile: A Long Video Question Answering Dataset and Benchmark
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c07e59de-411a-49eb-bf6f-bea0b08dda42 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation LAION-5B: An open large-scale dataset for training next generation image-text models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d59c8a55-e419-462d-96a2-21ec867793db · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Solving geometry problems: Combining text and diagram interpretation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9789e5df-5103-4315-9e4b-142f38f3933b · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation MetaMorph: Multimodal Understanding and Generation via Instruction Tuning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b49d0d64-18d9-4b4d-8235-aa46a9201558 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation CoVoST: A Diverse Multilingual Speech-To-Text Translation Corpus
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 32ca496a-6bf1-4228-b6be-b1ab8a8463ec · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Slidespeech: A large scale slide-enriched audio-visual corpus
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36a390e8-a972-421e-8826-460332d0dfae · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a82d713d-511c-41f1-8255-6d2bf0604a2a · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation ISBN 9798400701085
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bb92906-f6d6-4b38-a808-d40f05bc41ba · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Qwen2.5-Omni Technical Report
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 049d0006-7368-460b-a253-3cfe11c2066f · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Vript: A Video Is Worth Thousands of Words
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f598556-2f20-496f-b72b-02e98b4cf9ad · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41c20c2a-d8e2-49e7-88b9-94142f01a28d · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Fan Y u, Shiliang Zhang, Yihui Fu, Lei Xie, Siqi Zheng, Zhihao Du, Weilong Huang, Pengcheng Guo, Zhijie Y an, Bin Ma, Xin Xu, and Hui Bu
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9e141403-3595-47f2-9a8b-2d1e9bd1a02a · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Icdar 2023 competition on structured text extraction from visually-rich document images,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 53a379cc-53de-4149-a087-873789b7006f · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation ICDAR 2023 Competition on Structured Text Extraction from Visually-Rich Document Images
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19cd9c47-a240-4442-9e8b-0599fd5ce481 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4defd50b-aaff-4b29-9951-93ad0d5ea30d · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9d290ed7-dc77-45a3-b4e6-29d20b1418fa · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation How to Train Data-Efficient LLMs
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2e49b1d-b157-49f3-9bcf-4b9f40b7fff0 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Kimi-VL Technical Report
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87b56999-0a42-4485-8fba-9bd04296c0b8 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation doi: 10.18653/v1/D15-1171.https://aclanthology.org/D15-1171/
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 63b21592-3895-4dd0-b2c3-1a91ca1766ec · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Ahmed Masry, Do Xuan Long, Jia Qing Tan, Shafiq Joty, and Enamul Hoque
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df7893b3-3daa-45b7-9005-e5980e6ac6e0 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Chemvlm: Exploring the power of multimodal large language models in chemistry area, 2025.https://arxiv.org/abs/2408.07246
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e9975ef-c8b3-42a2-a663-b1273312b613 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation UniGeo: Unifying Geometry Logical Reasoning via Reformulating Mathematical Expression
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96fbd220-840b-4192-b178-0abc448886f6 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 414aa10c-dfc5-4e9a-aea4-45c25b413918 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bf8485e-b10a-48ca-b658-b84698963118 · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Qwen2.5-VL Technical Report
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 619ad11f-4d92-493b-ac30-edf1ab09139c · outbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5a44a29-2cb8-4c17-8dc1-78f83a0dfabf · inbound
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3abfd34-b80c-4e53-9850-34275c614d48 · inbound
MATE: LLM-Powered Multi-Agent Translation Environment for Accessibility Applications Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81019c43-5f2f-4939-90ec-c12646c4b2ac · inbound
OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 59e6e5ba-9d8b-4c72-a1ef-1317b5dd4ecd · inbound
A Benchmark for Omni-Modal Reasoning in Long Videos Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74fd0d8d-a4eb-4c6d-80d7-b164acdd6b87 · inbound
OmniFysics: Towards Physical Intelligence Evolution via Omni-Modal Signal Processing and Network Optimization Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2d20c179-b2fc-46c8-9617-2633fc13e696 · inbound
Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 69f0c9b0-bad1-475f-857c-98595b03aa46 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 237
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 761672d4-f70a-472c-9291-12db54bc66cc · inbound
Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c9f1329d-81bd-4abe-9103-0004a9efcaa5 · inbound
Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ab75913-f338-49f0-aab7-c3e6c54e60dc · inbound
Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 61e884c7-de6c-4fe5-b20f-6fa77e4eefd9 · inbound
Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation acd9c863-5d3b-4c7f-bc21-31b756b26a82 · inbound
Context Unrolling in Omni Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d2be920c-3348-412f-aa57-6314db1164f4 · inbound
SMoES: Soft Modality-Guided Expert Specialization in MoE-VLMs Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6e5184cd-82a6-4083-9a34-8a156e5a6c1b · inbound
Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation befb3024-ca75-409d-a5ed-cdabcf8ad0aa · inbound
Accelerating Compound LLM Training Workloads with Maestro Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d8a9558d-9a54-4723-b445-58b8bf5b836e · inbound
OmniRefine: Alignment-Aware Cooperative Compression for Efficient Omnimodal Large Language Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a59024a1-6a22-4c83-b3f2-a717820d62e9 · inbound
When Vision Speaks for Sound Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f9ee4736-29c4-4eb7-a044-59007715992e · inbound
UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c38594da-f5fd-434a-8324-fb58ef1d8085 · inbound
Toward Native Multimodal Modeling: A Roadmap Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 178
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation df39790b-feea-4da2-bf12-88b636dfa4a9 · inbound
Addressing Variable Heterogeneity in Distributed Multimodal Training with Entrain Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d996737e-1647-4160-bc43-d7ba78b8ea99 · inbound
MTAVG-Bench 2.0: Diagnosing Failure Modes of Cinematic Expressiveness in Multi-Talker Audio-Video Generation Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3b14cc91-248a-4b17-a5c9-f6a3c32c0e3e · inbound
Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 08c308ba-da7e-4407-9d0a-15ed987b1931 · inbound
PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 02327fc4-30ea-44bc-9b36-2beebe1677fc · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ee1207c0-46e7-4a8e-ab52-f27799133443 · inbound
CogniRoute: Learning to Route Social Evidence in Omni-Modal Models Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 90f74144-9e31-4f1f-aa25-c288041dabe8 · inbound
LiveServe: Interaction-Aware Serving for Real-Time Omni-Modal LLMs Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6d083d17-7830-4795-8cfb-bb967d85bd16 · inbound
AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0bf63540-2e1d-4b51-a92a-1afe5db3e18e · inbound
RedVox: Safety and Fairness Gaps in Speech Models Across Languages Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2dc81df9-a333-482f-b9d6-c0138f50139a · inbound
HorizonServe: Coordinating Request Scheduling with GPU Sharing for Omni-Model Serving Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50206a3c-3d6f-47da-8401-f85120f2b15a · inbound
MMAG: A Multi-Control Mixed Audio Generation Benchmark Ming-Omni: A Unified Multimodal Model for Perception and Generation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.