Pith. sign in

Paper Citation Record · LEDGER

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

As of 13 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 6 inbound Pith citation observations for arXiv:2506.21980.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21980 v3

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:18:39.317873Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T07:20:54.642591Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:24:21.251853Z

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6823e212-c87a-4de3-a1f1-22e76dfe2bc0 · outbound

This paper cites High-Speed Tracking with Kernelized Correlation Filters,.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning High-Speed Tracking with Kernelized Correlation Filters,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:42.675512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:36.490583Z digest=sha256:0fd3ed55141a4e611492b64d84c25e261bed74011d8db504fcec774d0fa9c28f

Observation c45c0873-9cbc-4a79-b63b-aaf71d954722 · outbound

This paper cites ECO: Efficient convolution operators for tracking,.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning ECO: Efficient convolution operators for tracking,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:42.453711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:36.750988Z digest=sha256:e72a7e2c6dc6f7b5eda1b3241764d21fc39b392c45a75c94fa881064a694f07b

Observation 57d47af3-446f-49c6-877c-a632b49e44f0 · outbound

This paper cites SiamRPN++: Evolution of Siamese Visual Tracking with Very Deep Networks,.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning SiamRPN++: Evolution of Siamese Visual Tracking with Very Deep Networks,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:42.237028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:36.835779Z digest=sha256:42549c0b192304af950ebff4693594b2d1d6441ae47abee73c3f6f96fc0b64d3

Observation 53265930-aa51-481e-bc98-81a3ed6b4324 · outbound

This paper cites Transformer Tracking.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Transformer Tracking

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:41.943251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:36.960457Z digest=sha256:01cc4937acf724551807e1aca6f37c5b8b9a6759d0ac64232c9a3abd5fbf3139

Observation 1b4712eb-771c-4ffd-bc9e-c3a3335a0d67 · outbound

This paper cites Joint Feature Learning and Relation Modeling for Tracking: A One-Stream Framework.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Joint Feature Learning and Relation Modeling for Tracking: A One-Stream Framework

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:41.631119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:37.142351Z digest=sha256:3b9b122d0035e5708950c32fdff4910df7cf73af55f28e77993370ff8ea26e89

Observation 83ac570b-33bb-4a2e-accf-8dd2e9a0d6cc · outbound

This paper cites Language models are few-shot learners.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Language models are few-shot learners

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:41.386405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:37.275271Z digest=sha256:6e4a035ba07fc3ee6652203cf9f36dd606ffc2b22e6b336e8b3d5725ef54b6d0

Observation cd98c7ac-9640-493c-81aa-5218806e30aa · outbound

This paper cites Visual instruction tuning.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Visual instruction tuning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:41.138987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:37.494877Z digest=sha256:180a18615dc1f4be0d3a81680c1b082e9988d209aa977f3ebfb3db7c40ab650e

Observation 853d58cd-65fb-4489-8823-a2d88a1469ae · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:37.640165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:37.640165Z digest=sha256:f3807e25b896b4f3f8fc05ca6d480947297ef911fbd73f82c98c9e73a38f1595

Observation 875ab6fc-f989-4412-b70f-77e93a694ec5 · outbound

This paper cites Qwen2.5-VL Technical Report.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Qwen2.5-VL Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:37.774926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:37.774926Z digest=sha256:a431f83d8566793f1d78ad3f4b1bf6c4a8af365a08543a7efb9eb0168ae33b8c

Observation 864aedb7-025c-4cb2-bf8c-257e8fb891f5 · outbound

This paper cites OpenAI o1 System Card.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning OpenAI o1 System Card

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:37.903290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:37.903290Z digest=sha256:0379df3c3940ee285c0cab28a5bb0e5b59f75e1ef09ccc92b1ec15b996d2e462

Observation 7c77dbd7-ed54-4eed-8daa-5368d9bd899a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:37.995094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:37.995094Z digest=sha256:88a86811c2eb21af57ae13d61611055d25aa92aa54ab7bf15919621047f8146b

Observation 04e945a1-948e-4042-a7cb-185f16c8ab5e · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:38.262307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:38.262307Z digest=sha256:f7d7078e26b7890b5e5a767249f6aa408e5e4d81ca801dd0f7ac87a34e62b421

Observation 8f16a149-0c7f-4882-863e-7d2d678dfb47 · outbound

This paper cites Got-10k: A large high-diversity benchmark for generic object tracking in the wild.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Got-10k: A large high-diversity benchmark for generic object tracking in the wild

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:40.871319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:38.425406Z digest=sha256:b84c40950c1f2ac66567235f62d8621eb489dffe07e81e7b27b4311dae9371ac

Observation 4996425c-bdbd-46de-a7c4-a1825e9501f7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:38.554959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:38.554959Z digest=sha256:bbf8b7fd328d3852282f5b0ed6c29493947d067eb8569f4e1c941ad17257f80d

Observation ad34a03a-5146-4ddc-99b5-6d510bcd0bee · outbound

This paper cites EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:40.619093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:38.727314Z digest=sha256:e527d7b3301f692c46d9cbf39ab1899c279dfc46b685949486442acc6e532095

Observation 59345c50-d6df-423d-8fc6-9aacd2ec5339 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:38.866409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:38.866409Z digest=sha256:72f03b0c5cf3c66c14897f125452405a3a47ead90dca3a8077f6367ca6c23a31

Observation 7db56ba5-61e6-441f-9f34-411054e36e64 · outbound

This paper cites Generalized intersection over union: A metric and a loss for bounding box regression.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Generalized intersection over union: A metric and a loss for bounding box regression

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:40.235389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:38.983117Z digest=sha256:229470c5d4ebd0d4ef4d8b43c03aee6473f10149cd7b4afd860d460ae41ff969

Observation d1baa265-8dab-40a6-b45b-718a038eceb5 · outbound

This paper cites Efficient memory management for large language model serving with pagedattention.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Efficient memory management for large language model serving with pagedattention

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:39.955767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:39.117650Z digest=sha256:5e87a4c0068394d3a9ab28d0837fc73baead6da9f52a7772fe58ae0f0c118bf5

Observation 296d11de-f1c0-4e9c-a13a-759223b5921d · outbound

This paper cites Improved baselines with visual instruction tuning.

R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning Improved baselines with visual instruction tuning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:39.675939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T22:18:39.317873Z digest=sha256:4aa05af8622a7273d7c777eef078f48ef32c2a5f98c0ded744e018af9d21801c

Pith citing papers

Observation e309a90c-410c-462d-abc7-833fe7e97128 · inbound

OneThinker: All-in-one Reasoning Model for Image and Video cites this paper.

OneThinker: All-in-one Reasoning Model for Image and Video R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:11:26.591440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T02:09:39.820651Z digest=sha256:8805f76ab6e7d602ef28c1ab505a96e782800447d198fdd2dd7a32ff0ceb3f23

Observation d8e3894d-1631-41d6-9596-19a46a6c472c · inbound

Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations cites this paper.

Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:13:22.984143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T20:12:33.250097Z digest=sha256:53d595d5f2742310b53c5a5347145efa99067107413595a7ff66aa2f6885dae8

Observation 089fe88c-1ecb-499f-81a2-338d51b865be · inbound

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding cites this paper.

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:52.937357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T18:38:16.204012Z digest=sha256:15a85b35f1c671d9dcc5575f4749beb23c78834962e1a0f46de92efc0d3ad2c5

Observation 395c07ae-93c3-4105-8e31-77d8e6e840dd · inbound

ReTrack: Evidence-Driven Dual-Stream Directional Anchor Calibration Network for Composed Video Retrieval cites this paper.

ReTrack: Evidence-Driven Dual-Stream Directional Anchor Calibration Network for Composed Video Retrieval R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:23.050130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T04:41:23.980996Z digest=sha256:bc689a31626319b8bd22a62dc28cc1875009e3ff8ef654426fb71ada1dc2df43

Observation e7348589-ec91-488b-b6a9-17f16c6d17a8 · inbound

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval cites this paper.

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:22.673159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T04:41:57.207279Z digest=sha256:078de85d7c293ed16215edea0ac488fd1157f968a241fbdfd322a58361198378

Observation 237d682c-c674-4c7a-919d-5a3f979fd039 · inbound

Dynamic Parsing and Updating Natural Language Specification using VLMs for Robust Vision-Language Tracking cites this paper.

Dynamic Parsing and Updating Natural Language Specification using VLMs for Robust Vision-Language Tracking R1-Track: Direct Application of MLLMs to Visual Object Tracking via Reinforcement Learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:21.253447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T07:20:54.642591Z digest=sha256:395a305e9043c31db2f620bbfab2448c4a0d4c39cd33ee75794869bd56b2233f