Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:55.925858Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 14 inbound Pith citation observations for arXiv:2505.20256.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:55.925858Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:07:21.383338Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T17:27:15.771664Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2232d5e1-3447-49bc-a637-195515e6dddc · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Baichuan-omni-1.5 technical report.arXiv preprint arXiv:2501.15368, 2025
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d14e2908-41ad-4ccf-8550-d93ac69167d0 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de29e52b-f9c2-48b2-9b1b-4832bbcb07e3 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Attention-based multimodal fusion for video description
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation db73eb7c-7f3e-4be0-910c-c0d9f3f905d4 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Merlot reserve: Neural script knowledge through vision and language and sound
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1a800756-ad6a-4554-a341-e925a6bc26bc · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration GPT-4o System Card
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a787639-7038-44a8-9486-9a03bd435645 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be76c2b1-c408-4058-bf43-522ab05a4069 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration From Seconds to Hours: Reviewing MultiModal Large Language Models on Comprehensive Long Video Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dba246a-4fb9-4b09-9aae-66b5f621b3a7 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Mavors: Multi-granularity video representation for multimodal large language model.arXiv preprint arXiv:2504.10068, 2025
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec614a37-190e-4dd2-9939-6843c541cbfd · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Lisa: Reasoning segmentation via large language model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8b47e923-4132-4fb2-b4e6-cd716d3c4cb8 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Qwen2.5-VL Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da2c69d7-9b43-48ea-8f86-2b0e4b4ecd34 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89728534-4ffa-413d-875c-af011ef21c29 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab764237-0230-4fb1-8c10-aa1c07c08872 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e81f7b7-2473-478f-ba07-1d29a9a0dee1 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Glamm: Pixel grounding large multimodal model
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b90be1db-5ac2-4685-bb08-193e845dc734 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Florence-2: Advancing a unified representation for a variety of vision tasks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e806b0ba-fb70-439e-bafa-89354ea7f852 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1768ab9d-764b-4ddd-92f2-d163ff7de1c4 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c77d780-6a4c-4f63-8797-8eaa616d254f · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4127edd4-722e-45d1-831c-5edaeb2445eb · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ea6870f-df15-4826-9250-2e6d29f1f584 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Ref- avs: Refer and segment objects in audio-visual scenes
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7dc7196d-1d7d-4a5b-9b90-e1fa8f024dc9 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration VISA: Reasoning Video Object Segmentation via Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0836b924-e8b9-452e-96a7-5763e205d746 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration GPT-4 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56cf3a5c-718e-402e-a6fa-5b117e9945c1 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Qwen Technical Report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37f867b5-ad9a-4d51-9c38-7b886fb81541 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration DeepSeek-V3 Technical Report
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb06e6fa-3c67-441d-a998-be2940b55e26 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 373e72bc-1f8a-40de-bfaa-f65fa07922f0 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e126d7be-81b3-4c64-988b-4074a4487086 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 811d0ee5-5e8d-4291-b3e4-a069be9de93d · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Visual instruction tuning, 2023
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a7f098b-4c37-4905-9a09-dca496a1a4fd · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Deepseek-vl: Towards real-world vision-language understanding, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0fb3a83b-ac11-422f-95b0-30d323ede06c · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Qwen2.5-Omni Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1596d63f-ea15-4435-9835-bba8ac980638 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Omnibench: Towards the future of universal omni-language models, 2024
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7acc720e-527d-4dd3-ae3e-2dbdd88aae6c · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc2370a3-20a0-475c-ab8b-8c05fdd5def9 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed74b537-c464-4e8c-ae1c-f4a2f2fb4fba · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f7eead2-aa70-42dd-8e9d-189cd6d4bd3a · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84479aa3-2f2e-43ac-82ad-7c4c856e9879 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration R1-omni: Explainable omni-multimodal emotion recognition with reinforcement learning, 2025
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 76ba6c11-fdad-4b64-8791-6417ff75765f · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration SAM 2: Segment Anything in Images and Videos
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 091839ed-40c2-4cb8-931b-5ad43caaca4f · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration End-to-end object detection with transformers
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ee01d021-eb3e-42d2-a4e6-7f0ab97fa304 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration MeViS: A large-scale benchmark for video segmentation with motion expressions
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9a828aec-1762-4492-81f7-72aa9620e47a · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Generation and comprehension of unambiguous object descriptions
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2760fb1d-f276-4eb6-8b9e-6aae6a500311 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1dfba77-6695-479a-82fd-bed20612b408 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Avsbench: A pixel-level audio- visual segmentation benchmark
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d024c242-252c-4501-ae1f-6bcb9173d017 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Avsegformer: Audio-visual segmentation with transformer
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2121f94a-ed7e-481f-ac14-118ba329eea8 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Prompting segmentation with sound is generalizable audio-visual source localizer
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fff28671-19d9-4f53-95d3-f7cedab4b6f5 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Language as queries for referring video object segmentation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b616ca86-1092-41be-9e2b-4c47089ceeb6 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Robust referring video object segmentation with cyclic structural consensus
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 28b03452-ecb1-438a-bf2a-f2a1a82b46a0 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration TrackGPT -- A generative pre-trained transformer for cross-domain entity trajectory forecasting
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05bc131d-b4c1-4755-91d1-174dceae61f9 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed5cd32a-4c1b-4cbd-b21c-9fe3843808f9 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration LLaVA-OneVision: Easy Visual Task Transfer
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3fcd63b-ceaa-46f9-bd7f-2fcd422481f8 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41704680-dbff-4445-9e9d-9142addf6d12 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54bdf31a-b38e-46dd-8339-8ec5817828e8 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30c68b51-6c35-4cf8-b986-ded5388dec21 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Mvbench: A comprehensive multi-modal video understanding benchmark
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8162af0-fae5-4271-94fb-4f446fa0a774 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec44756a-9734-4793-b05a-a27563163f71 · outbound
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 748bcc57-ef43-4c6c-874a-623cafbcc7f9 · inbound
Group Relative Policy Optimization for Speech Recognition Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11fa81bf-7cdf-4185-802d-a848fa6ccfb4 · inbound
XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9a349702-83c9-405d-ae4c-4abb8a12968e · inbound
LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 60a8a706-ac62-4312-9138-576e90d3bfca · inbound
Omni-R1: Towards the Unified Generative Paradigm for Multimodal Reasoning Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a35bda6a-7128-40f8-b307-a518e00bcfc2 · inbound
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5cd62825-00f7-443b-b17b-2f24b32fc7a8 · inbound
Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3eb7c43c-ccbd-4e9a-a055-c0581c6ba1b4 · inbound
Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6eb1a9ba-ff0e-40fe-a774-7737a268bd6a · inbound
Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33213efd-ece5-4758-826a-c2693ae4e7c5 · inbound
Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6baa4af0-3229-4b99-b6c1-dfea506983bf · inbound
AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cf8a1670-c68e-4524-ad51-1587028d1389 · inbound
PRIMED: Adaptive Modality Suppression for Referring Audio-Visual Segmentation via Biased Competition Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e46a2ae7-4b41-41ec-9c82-fcdb79ddfe19 · inbound
RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 59631231-82f1-48ef-91b3-68a6c3106ccb · inbound
Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6c937d94-f391-4895-86bf-b614ce2e5c02 · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
Reference 202
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.