Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:15:42.409411Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 96 of 96 outbound references and 12 inbound Pith citation observations for arXiv:2505.15804.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:15:42.409411Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-13T16:55:20.099628Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T15:09:55.322058Z
96 of 96 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dcdf50c0-212e-472b-95d0-367bc152ff3f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 841067a9-ed1a-46ba-a0ee-2455528740e4 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Pixtral 12B
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca00110c-34f1-4ad6-9275-c01dd8fc5f2f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea2d13b7-7fab-409e-bd83-53ac10ced950 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 212cd02f-0c13-4f1d-adda-1d442496413e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen2.5-VL Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd785b30-6e62-4a18-8f3d-bdbe30a01885 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs R1-v: Reinforcing super generaliza- tion ability in vision-language models with less than $3.https://github.com/Deep-Agent/ R1-V, 2025
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c57b4766-fc0e-46f8-b0e5-4965091fba79 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fadc10a-2785-49c0-aedb-977e046774f4 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1d57b38-53f5-49c4-a16d-20c220232ee3 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 675891a1-c4a2-47a4-9012-e0581005d856 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a081f643-7ecc-473a-9653-2dedd8b9ab6a · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 293ee2da-eb8e-4f1b-82cc-912464bb340f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Training Verifiers to Solve Math Word Problems
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4f0255e-e449-42b3-8f9f-641679fe6bcd · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Sophiavl-r1: Reinforcing mllms reasoning with thinking reward
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1901736a-b61f-43ec-9751-af79e224f2da · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50844e53-ff55-471e-b50c-1fd8da55cd7c · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d73b1f5-bc43-45ea-8adc-38778701ecfd · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f68052e0-3672-428a-a044-8e74c7dce9f9 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Geneval: An object-focused framework for evaluating text-to-image alignment
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37009068-0b55-4422-bbfa-5d26f070f066 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs The Llama 3 Herd of Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9b71cb-00c7-43cd-b130-f5b43043dd0e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7e2a54a-18fa-44ae-8c03-ba17156ca081 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Measuring Massive Multitask Language Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60642aa0-19f7-4be2-9edb-31f111399fa1 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Measuring Mathematical Problem Solving With the MATH Dataset
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31f0ea88-2bad-4749-9571-c93c37a31657 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Transformation driven visual reasoning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 917f6bee-c864-4648-aedc-76a5bb7d8451 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85bfca2-c3c0-4ca3-addc-00913ef40ebb · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs GPT-4o System Card
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d30f8d17-a1a3-411d-9322-29d2704458c1 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Omnispatial: Towards comprehensive spatial reasoning benchmark for vision language models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20186f00-b6df-410e-90fb-6bb84a7c64b3 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Gonzalez, Hao Zhang, and Ion Stoica
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a109d7e-59ea-4cba-8521-f65ccbe551f3 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaVA-OneVision: Easy Visual Task Transfer
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e666ddb-be36-425a-86fd-cf359b6216de · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d582b926-e633-4030-af2b-ea3b7dd4edb3 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Mvbench: A comprehensive multi-modal video understanding benchmark
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caddb07a-2955-4f0f-b28c-415281230c70 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Temporal Sampling for Forgotten Reasoning in LLMs
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa3718f8-5a0e-4349-bf35-85bc59e9dbb5 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Sws: Self-aware weakness-driven problem synthesis in reinforcement learning for llm reasoning, 2025
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44d8611a-5f30-48b5-bde4-b16373eea687 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Improved baselines with visual instruction tuning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eec8507e-f0c7-4735-a8fe-825ecadb582b · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f1947f-c76c-4444-8630-cb9fb76e6631 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs IR3D-Bench: Evaluating Vision-Language Model Scene Understanding as Agentic Inverse Rendering
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdf7e325-6210-4536-bd71-a9c819e8432e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25d24bf1-19d6-4345-88ee-c4c100b35c1f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Egoschema: A diagnostic benchmark for very long-form video language understanding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c8831b1-3d61-4d0e-8dc8-4c15ce522f78 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Docvqa: A dataset for vqa on document images
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb33604e-ab4d-4479-992f-5a4abef9a15e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs ivispar– an interactive visual-spatial reasoning benchmark for vlms
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5b910e8-6cbf-4963-97ab-65f3af45f796 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0443644-5b99-45d8-aca1-fcc9c1fdd016 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bf1dcc4-e754-41a0-83dd-12fce5eed07f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Learning transferable visual models from natural language supervision
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 735331ce-7784-402a-accb-e2dadbf6ddb1 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8e46b05-bd85-4e0f-b50b-6367451fa9d2 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd0adcd0-9995-410b-bd76-a4e7d4c46316 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Maniplvm-r1: Reinforcement learning for reasoning in embodied manipulation with large vision-language models, 2025
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e12c1220-c471-427b-a623-b9aab9688493 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Generative multimodal models are in-context learners
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cb7b0e2-9c83-43a7-9cce-1d876cd68385 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcf325dd-e680-401c-9b78-362e0746a881 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a1aefd-278f-4292-b814-a934fda0d30e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d813ab3-513b-41b2-98d7-a677c7f9c0d0 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy, 2024
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b76ae6b-8b34-42b9-8a24-c21f19a89c97 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaMA: Open and Efficient Foundation Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95cbe8c0-37a0-440d-834d-f0fb00ad8cf2 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 489e5f84-5ebd-474e-ba7e-77519bc44865 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Solidgeo: Measuring multimodal spatial math reasoning in solid geometry, 2025
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc84bc8d-8da3-4aa6-a097-e5ddbc760436 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 444d3501-957b-412e-94e2-d863ecb8ec32 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LVBench: An Extreme Long Video Understanding Benchmark
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 232b0cc2-325e-4178-9a77-6c71c261ab59 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Emu3: Next-Token Prediction is All You Need
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6686393-a273-418f-857c-304c2427e61a · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Chain-of-thought prompting elicits reasoning in large language models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 164c55ef-ee9c-4dad-b383-3b7f377d05c5 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs CMATH: Can Your Language Model Pass Chinese Elementary School Math Test?
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d081f85-7c69-48b1-8c65-71e19db85da3 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8c930ac-37a6-4783-9ef2-ee21b8180066 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaVA-CoT: Let Vision Language Models Reason Step-by-Step
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51606719-fb28-47d4-ad15-36ff7ef06158 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddc43479-a4c1-4e89-9009-4e231dfb0a77 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen2.5 Technical Report
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6d0a0dc-20a9-4b37-b7ad-11ab42174270 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepCritic: Deliberate Critique with Large Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35e4b4e5-238b-4731-a32d-1a80201f4aea · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Towards thinking-optimal scaling of test-time compute for llm reasoning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 901e1fc9-dabd-4011-8231-b168b2820735 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e23c3376-8607-4cfd-b530-4cdb6610b899 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 519b80c0-16a8-4165-b300-b17059383537 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1520b89-e6ad-4e5d-95f4-2128216ee265 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba8898d5-c2c4-4162-b806-bda2c55114b5 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d24babf-3c90-4ee8-8d36-656b4965946f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MLVU: Benchmarking Multi-task Long Video Understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67027c9-4458-48db-9a08-8c0e0be4206f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 096c5082-9f0f-4141-9b5d-e7cca97dc84f · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs We employ Qwen2.5-VL-7B as our base model and utilize vLLM [26] as the training framework
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 244a5890-5c75-4fe9-899d-5cf984edde73 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs cube", "sphere
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db0c6a97-a730-43a0-a38f-f16d909cd050 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs large" (radius = 6),
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18b1c8d7-ee0a-4965-9833-d33fe965697d · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs gray", "red
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c323fc22-ff28-4c8c-91f5-c36482bf9bc0 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs rubber",
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 873d6753-d86f-4d12-84b2-2cbb8531affb · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11a8e450-aaaf-4057-9dbc-be23bb60b513 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f7bfbb6-b6d6-44cc-870d-d6db07df822b · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 065dac49-2284-4996-aa2a-0163cf1d7964 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a827d13-3c5d-461a-9659-77a7cccc9283 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e93abd86-47ce-4df2-8ffb-6a464d6f5781 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e66144b6-0ba6-492e-b06e-c3d0a3b6f0b9 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfb56c9a-21dc-4c78-a0d7-f11722c27f22 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab463e98-fd8e-4f31-a007-2314901443fc · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29dbecf5-bbf3-42fb-9721-69d2c6ec808e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61be8813-fb52-4503-92fa-cd11e7f434f1 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 183f12b8-9608-4408-9b86-3b50e85ac1a1 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5820b93-e0f5-4ddf-a0a6-ed01e031abe1 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a87e34da-cf3f-4178-8d74-735548c44d4d · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c93ac4e-86e4-4858-8a80-58a2ab6fa953 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca901aa9-f866-445d-be0f-d3bf6cc630ed · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 808d1f98-d957-4809-8e18-8245b27918ed · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46eb0fbf-5a8a-4bda-aa91-b8eae4758204 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 20b72e1b-ea82-4c20-8442-c828b79e361e · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a894983-d8ce-402e-b519-68e6b63834c0 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6335bd32-dda7-42fb-b6b2-ff2428698243 · outbound
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4cb3a065-2600-492c-a285-090f70fc77d3 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 297
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cad82a6b-df4e-4c32-8a85-f002afae1123 · inbound
Video-R1: Reinforcing Video Reasoning in MLLMs STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23c52071-9374-4517-9e4b-b5b62d315cf2 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 233
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0c118ee-ba2c-4943-91d0-197b451d7758 · inbound
VIDEOP2R: Video Understanding from Perception to Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5ef8303-4565-4db1-a785-e21994e09dc6 · inbound
OneThinker: All-in-one Reasoning Model for Image and Video STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b1a2ca8-b126-4ea0-a620-aba6a117bbea · inbound
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5fe546e-47f0-4f33-b2df-f9ef43d8d2e8 · inbound
Fully Spiking Neural Networks with Target Awareness for Energy-Efficient UAV Tracking STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5b02838-3488-422e-b562-c080880d485d · inbound
Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ae53715-7a4f-4529-9a94-fbda2d4735b4 · inbound
Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd904fd6-2380-4591-afc2-19ec70eb6245 · inbound
Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1bcdc9a-4721-4290-a3e5-7842709a2325 · inbound
Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76164de6-9435-4e6a-a1ef-2695f34cf3de · inbound
From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
Reference 192
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.