Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:33.223922Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 26 inbound Pith citation observations for arXiv:2505.14246.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:33.223922Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:04:19.981518Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
57 of 57 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation eba71093-4b50-4a74-be73-6e6aa8a27e7a · outbound
Visual Agentic Reinforcement Fine-Tuning Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37f415e6-f94f-4cfc-84e7-0dc74aefcce8 · outbound
Visual Agentic Reinforcement Fine-Tuning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bb37353-e8cb-4ac9-86bc-d6b8d4ebf0d1 · outbound
Visual Agentic Reinforcement Fine-Tuning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ff072ad-0201-483e-acca-62a01329f947 · outbound
Visual Agentic Reinforcement Fine-Tuning Rico: A mobile app dataset for building data-driven design applications
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a7e0fa7-6b81-4ec4-b3ee-2b234a43114f · outbound
Visual Agentic Reinforcement Fine-Tuning ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56f9d6cb-eb9e-4cb4-b5de-f05489c2026f · outbound
Visual Agentic Reinforcement Fine-Tuning OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6bd3552-31f0-4518-b16d-d01d8f677910 · outbound
Visual Agentic Reinforcement Fine-Tuning Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98273d02-832c-49ac-a4d8-b0b721538a11 · outbound
Visual Agentic Reinforcement Fine-Tuning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a671e0c-9e19-4301-8e17-06234cd5b2c1 · outbound
Visual Agentic Reinforcement Fine-Tuning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd2e3d2b-80fb-4e34-b9f7-f70395ea796c · outbound
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5eef7de-a9df-42d2-ac17-2866a3206a6f · outbound
Visual Agentic Reinforcement Fine-Tuning OpenAI o1 System Card
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1798631c-9641-43b4-913d-6c8f4b0fbb2f · outbound
Visual Agentic Reinforcement Fine-Tuning Funsd: A dataset for form understand- ing in noisy scanned documents
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2d898de4-7829-41e0-a13d-1c5972db8f41 · outbound
Visual Agentic Reinforcement Fine-Tuning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03f2efd7-faa7-46c1-8941-40ac41b53eaa · outbound
Visual Agentic Reinforcement Fine-Tuning Are you smarter than a sixth grader? textbook question answering for multimodal machine comprehension
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ec2396d2-1747-4864-a0aa-8cf966d3235c · outbound
Visual Agentic Reinforcement Fine-Tuning Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f8120d5-6b86-41c8-9671-7af117e0f4ff · outbound
Visual Agentic Reinforcement Fine-Tuning LLaVA-OneVision: Easy Visual Task Transfer
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12cec2a7-abdf-4484-a537-a7d49576ec7a · outbound
Visual Agentic Reinforcement Fine-Tuning LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81bd1770-9f80-471c-8d42-e52cec76a3c7 · outbound
Visual Agentic Reinforcement Fine-Tuning Search-o1: Agentic Search-Enhanced Large Reasoning Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b789c3e-1a87-4ba0-9b5e-9ad6d8fe20ba · outbound
Visual Agentic Reinforcement Fine-Tuning WebThinker: Empowering Large Reasoning Models with Deep Research Capability
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 242dc923-cbeb-4a5f-93d5-edf32159a6d5 · outbound
Visual Agentic Reinforcement Fine-Tuning ToRL: Scaling Tool-Integrated RL
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c66baf67-60d2-40f6-a412-24e70996a822 · outbound
Visual Agentic Reinforcement Fine-Tuning MARIO: MAth Reasoning with code Interpreter Output -- A Reproducible Pipeline
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a16924f-b1dc-4240-b031-a515a10c3bdc · outbound
Visual Agentic Reinforcement Fine-Tuning DeepSeek-V3 Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 655ed0a7-707f-4fc9-9dcb-7c24f1f07623 · outbound
Visual Agentic Reinforcement Fine-Tuning Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85a39c18-1c3e-4050-910b-f8a60a7772db · outbound
Visual Agentic Reinforcement Fine-Tuning Improved baselines with visual instruction tuning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c6df9bf-e8c5-4462-9379-efda43819b88 · outbound
Visual Agentic Reinforcement Fine-Tuning Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15e7b348-2b4a-49ad-843b-126a7805723d · outbound
Visual Agentic Reinforcement Fine-Tuning MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f32866b-9bd4-4725-8eee-ecdd2497a2a1 · outbound
Visual Agentic Reinforcement Fine-Tuning Towards end-to-end unified scene text detection and layout analysis
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f83f97ba-dff6-4823-ac40-8ae1fdb0b1d1 · outbound
Visual Agentic Reinforcement Fine-Tuning Docvqa: A dataset for vqa on document images
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d5d6c53-b62d-4137-8bec-5b29142bfd59 · outbound
Visual Agentic Reinforcement Fine-Tuning Openai o3 and o4-mini system card
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 130f2762-e4b3-426d-83ae-f092012e2668 · outbound
Visual Agentic Reinforcement Fine-Tuning Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa4980e5-7038-4d89-a2dc-b53568a92cef · outbound
Visual Agentic Reinforcement Fine-Tuning Gorilla: Large language model connected with massive apis.Advances in Neural Information Processing Systems, 37:126544–126565, 2024
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45b2f3a3-8c62-4438-b0d9-28ac5ba784e9 · outbound
Visual Agentic Reinforcement Fine-Tuning Measuring and Narrowing the Compositionality Gap in Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28131c6f-787c-48c0-993b-cf351e59a231 · outbound
Visual Agentic Reinforcement Fine-Tuning ToolRL: Reward is All Tool Learning Needs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 552cfcd0-d140-4f3c-927d-4e9fc3ff7cff · outbound
Visual Agentic Reinforcement Fine-Tuning Direct preference optimization: Your language model is secretly a reward model.Advances in Neural Information Processing Systems, 36:53728–53741, 2023
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d762b09c-88dd-4453-afd5-5fa0a8d0630a · outbound
Visual Agentic Reinforcement Fine-Tuning Proximal Policy Optimization Algorithms
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66484908-b300-4c1e-b938-d291a55cede8 · outbound
Visual Agentic Reinforcement Fine-Tuning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c15a9e0d-c152-409d-b51a-d801dbb5b727 · outbound
Visual Agentic Reinforcement Fine-Tuning Agentic Reasoning and Tool Integration for LLMs via Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fe74b72-402f-4ed7-a260-a48087a6a253 · outbound
Visual Agentic Reinforcement Fine-Tuning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9b41e95-0d47-481f-b6e2-c4ee79f0e13e · outbound
Visual Agentic Reinforcement Fine-Tuning ZeroSearch: Incentivize the Search Capability of LLMs without Searching
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 660aa28c-4602-46f5-8cd9-c0fb6858fbf3 · outbound
Visual Agentic Reinforcement Fine-Tuning Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7991defb-546e-4abb-b9e3-28a9c4270a6c · outbound
Visual Agentic Reinforcement Fine-Tuning Gemini: A Family of Highly Capable Multimodal Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ce2bc79-6281-4e92-841f-cf235de462ee · outbound
Visual Agentic Reinforcement Fine-Tuning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a40dff6a-fe12-4059-849e-2e610f03f36c · outbound
Visual Agentic Reinforcement Fine-Tuning Musique: Multihop questions via single-hop question composition.Transactions of the Association for Computational Linguistics, 10:539–554, 2022
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c1c7d64-fd79-40a6-a2bd-b1b8d8d2bf3f · outbound
Visual Agentic Reinforcement Fine-Tuning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fa59e86-7134-406c-a9d6-16d7f0b619b9 · outbound
Visual Agentic Reinforcement Fine-Tuning HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8599fc09-f5f1-400f-9b79-f8c25aa03f61 · outbound
Visual Agentic Reinforcement Fine-Tuning Detecting texts of arbitrary orientations in natural images
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6c40227d-ec06-4d68-8082-0a6ea4305e67 · outbound
Visual Agentic Reinforcement Fine-Tuning RlHF-V: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2e6cd8f5-b10e-4e77-ab84-ae3dcf275d60 · outbound
Visual Agentic Reinforcement Fine-Tuning RLAIF-V: Aligning mllms through open-source ai feedback for super gpt-4v trustworthiness.arXiv preprint arXiv:2405.17220, 2024
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af863026-8903-46e6-adf7-f264d3c4b139 · outbound
Visual Agentic Reinforcement Fine-Tuning InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bce2c218-ac62-4c3b-9be2-da4876f22f2b · outbound
Visual Agentic Reinforcement Fine-Tuning InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62d10de-d64c-42de-b2d7-77681a765934 · outbound
Visual Agentic Reinforcement Fine-Tuning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1735659-6f1a-490c-ad60-14d9df918b7a · outbound
Visual Agentic Reinforcement Fine-Tuning Aligning Modalities in Vision Large Language Models via Preference Fine-tuning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7201fc3b-4375-46ef-8839-c0f51fdc0c73 · outbound
Visual Agentic Reinforcement Fine-Tuning Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 782efb4f-3f10-48d7-af85-c042c9b8637b · outbound
Visual Agentic Reinforcement Fine-Tuning the image
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 94189a73-5fb7-42d9-b570-b5eb87aa8083 · outbound
Visual Agentic Reinforcement Fine-Tuning Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7ebc7938-f7eb-42c3-86aa-fb1ab16214ef · outbound
Visual Agentic Reinforcement Fine-Tuning Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b4ca1758-b303-4889-ac1a-bcfa784645eb · outbound
Visual Agentic Reinforcement Fine-Tuning Can You Find Vermeer's Milkmaid at the Rijksmusem?
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c725d446-5861-48db-920f-891c115909de · inbound
Iterative Tool Usage Exploration for Multimodal Agents via Step-wise Preference Tuning Visual Agentic Reinforcement Fine-Tuning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7f07b1-1c5e-410c-aafa-bf50b790308d · inbound
Learning Only with Images: Visual Reinforcement Learning with Reasoning, Rendering, and Visual Feedback Visual Agentic Reinforcement Fine-Tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a1e24f4-4d32-4e6b-a632-3a1e3806b222 · inbound
M2IO-R1: An Efficient RL-Enhanced Reasoning Framework for Multimodal Retrieval Augmented Multimodal Generation Visual Agentic Reinforcement Fine-Tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc79799a-110e-4fd8-acd1-67023cab6372 · inbound
Omnidirectional Spatial Modeling from Correlated Panoramas Visual Agentic Reinforcement Fine-Tuning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 033251b8-d0f4-498b-abd8-e5eb35679d80 · inbound
DeepEyesV2: Toward Agentic Multimodal Model Visual Agentic Reinforcement Fine-Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation af55a94a-d81b-49c0-9995-b46ac10f79fc · inbound
Code-in-the-Loop Forensics: Agentic Tool Use for Image Forgery Detection Visual Agentic Reinforcement Fine-Tuning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2525a52f-b6e8-40ab-afff-3cafdedbf85a · inbound
Code-in-the-Loop Forensics: Agentic Tool Use for Image Forgery Detection Visual Agentic Reinforcement Fine-Tuning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db3d9f0f-b51b-4689-8a98-9c83b448bdbe · inbound
CharTool: Tool-Integrated Visual Reasoning for Chart Understanding Visual Agentic Reinforcement Fine-Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f0b57e88-2868-4f54-b756-d5bfd82a6cf8 · inbound
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning Visual Agentic Reinforcement Fine-Tuning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 323e38f1-cf64-4873-8596-f0686381a9a1 · inbound
POINTS-Seeker: An Open Recipe for Multimodal Search Agents with Visual Memory Management Visual Agentic Reinforcement Fine-Tuning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d9d59b1f-9f42-4f44-a64c-754c24f77a8b · inbound
POINTS-Seeker: An Open Recipe for Multimodal Search Agents with Visual Memory Management Visual Agentic Reinforcement Fine-Tuning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f541607c-fd75-4ef7-a8e9-acc00711daba · inbound
Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs Visual Agentic Reinforcement Fine-Tuning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e2ea8c18-56ec-4367-b21d-c097e4887d49 · inbound
Rethinking Reinforcement Fine-Tuning in LVLM: Convergence, Reward Decomposition, and Generalization Visual Agentic Reinforcement Fine-Tuning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dd30d3a7-b8fa-48d7-a4c5-4859b3c4b05f · inbound
Supermassive Black Hole Winds in X-rays: SUBWAYS IV. Tracing Radio Emission and Unveiling the Role of Winds Visual Agentic Reinforcement Fine-Tuning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0538cd5a-9582-484b-99b4-99161069ed6e · inbound
S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images Visual Agentic Reinforcement Fine-Tuning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6c77e42a-913e-4405-98e4-2bfc083b8fdb · inbound
Perceptual Flow Network for Visually Grounded Reasoning Visual Agentic Reinforcement Fine-Tuning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3a11aa03-4cac-476a-bbe5-0930825189ef · inbound
When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning Visual Agentic Reinforcement Fine-Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ebb3ff76-1be6-49dd-b1f3-dad5b93d9c85 · inbound
When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning Visual Agentic Reinforcement Fine-Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 334ef447-a13b-44cc-bf05-0d65216b2be4 · inbound
When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning Visual Agentic Reinforcement Fine-Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bdd92be5-3bf5-4e97-be2e-a068244c8510 · inbound
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models Visual Agentic Reinforcement Fine-Tuning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 65fa70b1-12b0-4109-a82a-808d652b119b · inbound
Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles Visual Agentic Reinforcement Fine-Tuning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dd367bf8-6560-4f72-be9f-de5dd70f0b21 · inbound
TAPO: Tool-Aware Policy Optimization via Credit Transfer for Multimodal Search Agents Visual Agentic Reinforcement Fine-Tuning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b9c26ddd-41bb-49fc-b1fe-b9f03bffa7a2 · inbound
Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence Visual Agentic Reinforcement Fine-Tuning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 316a23a1-8827-4d56-ae2d-b7e8349b24dd · inbound
SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search Visual Agentic Reinforcement Fine-Tuning
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 73392426-f97a-43ea-962c-b5d9e7e16716 · inbound
VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning Visual Agentic Reinforcement Fine-Tuning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 656c0542-960e-4728-9f4c-95f14bf59551 · inbound
InSight-doc: Agentic Visual Perception for Long-Document Understanding Visual Agentic Reinforcement Fine-Tuning
Reference 140
Source-reported events for the cited work
Unavailable: canonical work link unavailable.