Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:20:59.938701Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2505.15447.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:20:59.938701Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:28:56.047745Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T05:56:07.964144Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cf12f22b-1d92-4016-afcf-f6ebf19bc729 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ac58e4-903c-438f-8d0d-c130fca152aa · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 342666bf-c2fd-4629-ad13-108355f1159c · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Sharegpt4video: Improving video understanding and generation with better captions.Advances in Neural Information Processing Systems, 37:19472– 19495, 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bd6c3fb-fa54-444c-8940-686612390029 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.Science China Information Sciences, 67(12):220101, 2024
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7df3f253-5ae7-484a-8a85-caa50585f1e0 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Supervising strong learners by amplifying weak experts
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c265e10-9dcf-4610-9ba4-825985f4bb4d · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f78dde-a786-407f-a9c0-4cac870c7a37 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd2ede75-5c5b-4b78-bb51-b4595d376349 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f85709bc-2a0e-4bd8-bbe6-daaf4d5c108b · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db2ab293-089b-4639-b534-f9246a05f5a8 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning CoS: Chain-of-Shot Prompting for Long Video Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e53acd5-2aaa-4d8f-9a02-327391ec0747 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning M-LLM Based Video Frame Selection for Efficient Video Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 113e3bc5-eadf-4c37-92bb-29c545e23c57 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cf1332b-7be7-4177-b5e0-de0665d5f8ad · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Mvbench: A comprehensive multi-modal video understanding benchmark
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e2e35f4-2831-4741-985a-97dad4e1fcc7 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da9d7c4a-a09d-46f1-aee2-ea29148cb717 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning KeyVideoLLM: Towards Large-scale Video Keyframe Selection
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 720ecb11-afd2-4c2e-a1bb-ad1569845d9f · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 399c84bd-a020-40a0-a430-6bdb27af0921 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8335557e-cb75-467b-b018-6d8712a4a5b0 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a699b2b-43ba-4426-a63a-3c27efa0be05 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Hello gpt-4o.https://openai.com/index/hello-gpt-4o/, 2024
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c66a802e-73d2-4f05-8583-4eb95c48834c · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Ai 2027.https://ai-2027.com/, 2025
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74c24c73-18f8-4c77-a611-7f3807c6e079 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Introducing openai o3 and o4-mini
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbfa3356-28f4-4ebe-b8a7-afdc691d7104 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Learning transferable visual models from natural language supervision
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 436ddb25-59ae-4f29-aed5-494b4d635f43 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Timechat: A time-sensitive multimodal large language model for long video understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1e8c877-96e5-4be1-834f-8bdbff7da2c3 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Trust region policy optimization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9115f05f-a2f0-4b8b-9f4e-6fe819ccecad · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87adb3c3-3276-4d74-b086-2120528e9b70 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d9c289e-2d5a-4082-a8ec-fcc72a4a1192 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbe22634-eae3-4339-8748-7cdf90877cd4 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Moviechat: From dense token to sparse memory for long video understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5fc3f40-2601-4cb8-a4ee-395a76b206cf · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Adaptive Keyframe Sampling for Long Video Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65c61772-c52c-4bbc-9c30-2dd5b0872121 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Gemini: A Family of Highly Capable Multimodal Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be03e616-7c36-4067-ba95-416a5bc8a359 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning LVBench: An Extreme Long Video Understanding Benchmark
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ffb3ee0-50c0-4ea5-affe-48f2ee8e725d · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1445af28-cabc-47dc-8c25-d4d3be3d296f · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning LongViTU: Instruction Tuning for Long-Form Video Understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e3c51c5-4ff5-45b9-bf7d-a5636c50b639 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Number it: Temporal Grounding Videos like Flipping Manga
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a6ad6ee-18f4-468e-92b1-152590b8d05f · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Frame-Voyager: Learning to Query Frames for Video Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8961aedc-0025-46cf-a0aa-5d0cca75b54d · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 686d382c-f69d-4325-9780-7c48d5248f4d · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Long Context Transfer from Language to Vision
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75ba5317-b5ea-4f21-b12a-517e7916add4 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2855bbcd-de4e-45a1-b7e1-4e80a065abd8 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning MLVU: Benchmarking Multi-task Long Video Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28a12429-787f-48e2-805e-b199c7f94349 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning - Check if the occurrence time is mentioned
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4940706e-4e41-4fd1-92a3-f8c3efbb7516 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3571a2f4-bc9c-4d10-a20a-9d55c9d3ba98 · outbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning mRNA" and
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8527e4c0-d110-4e6d-818e-d6e4fa602931 · inbound
Phi-Ground Tech Report: Advancing Perception in GUI Grounding ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 125702e3-0b83-442c-b5ec-d8ac8b39a4d9 · inbound
Swift Sampling: Selecting Temporal Surprises via Taylor Series ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3638ec97-6a7f-47c4-81fd-d04f151aeebf · inbound
Efficient Frame Selection for Long Videos at Test Time with Attention-Based MLLM Selectors ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.