Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2503.16365.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:21:40.483549Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 5854a8e8-56d2-44d4-8836-6ee06d98abf4 · inbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d027aadf-72d6-4160-bbc7-99f50ff05e08 · inbound
A Survey on Vision-Language-Action Models: An Action Tokenization Perspective JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 262
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6e5074be-22d6-4b5e-ab8f-c71c5202bea1 · inbound
InstantEdit: Text-Guided Few-Step Image Editing with Piecewise Rectified Flow JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abb013ec-6241-4b19-a271-d9093a54621a · inbound
Society of Mind Meets Real-Time Strategy: A Hierarchical Multi-Agent Framework for Strategic Reasoning JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eff03f3-e254-4d06-88d9-e3454eef8b1c · inbound
UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4ce6403a-02bc-4462-967f-09a4643747aa · inbound
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 80d67920-a61c-4882-b702-b4cd4cfbe0ff · inbound
World2Minecraft: Occupancy-Driven Simulated Scenes Construction JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 608bc25b-f1c0-4202-b1d9-9aacf3f88b88 · inbound
Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4ea5cb6b-4c02-4bc8-b4cb-0dc5a16a3bc7 · inbound
Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d81108ae-4df3-4b7d-a0d3-65b445c56b69 · inbound
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.