Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.536019Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 0 inbound Pith citation observations for arXiv:2507.19875.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T18:00:25.536019Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 100 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7b060917-d2fc-4852-a4e9-0c929326980c · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Exploring Visual Prompts for Adapting Large-Scale Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b6168c0-3fd5-451b-91ce-69a90b892be6 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Qwen Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feac114b-a727-4bb9-ba01-7b223253147e · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking ARTrackV2: Prompting Autoregressive Tracker Where to Look and How to Describe
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31ab4e52-3b86-4a21-9c3b-bc1e9f8aaef6 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Visual objects in context
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dd1badc-254e-41ee-a3ba-7f44c1ff97d2 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Fully-convolutional siamese networks for object tracking
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c458deef-7dc7-4f87-b3d3-1d08557fbc81 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking HIPTrack: Visual Tracking with Historical Prompts
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6664951-cb60-4632-94aa-56def1c74905 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Hiptrack: Visual tracking with historical prompts
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea980acd-dab8-4366-8809-ccbad76b884d · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Robust object modeling for visual tracking
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d8f018-330a-4a16-9f34-107911ef917f · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking The relative con- tribution of scene context and target features to visual search in scenes
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb728ed-985c-4e69-843e-ca843aeb2b6d · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Back- bone is all your need: A simplified architecture for visual object tracking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af9e0079-0496-4bd5-9533-29d214d34c57 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Revealing the Dark Secrets of Extremely Large Kernel ConvNets on Robustness
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 173655c0-8471-4ffb-b9c4-7fe625a453d2 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Transformer tracking
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3e2cd41-3d4d-4834-9634-5aab90328e39 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Seqtrack: Sequence to sequence learning for visual ob- ject tracking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dda1452-603e-469b-9b69-278e31137e65 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking SUTrack: Towards Simple and Unified Single Object Tracking
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43773dbe-f9ec-4cff-909b-deff0e94bff5 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Siamese box adaptive network for visual tracking
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68168ffe-117a-40a8-bb38-bf4d6d3509ce · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Mixformer: End-to-end tracking with iterative mixed atten- tion
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 977012a5-2c21-4eb4-a6f9-4e10441ca09e · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6151760b-09d7-454a-9ac1-62b434f780cf · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b59bdd76-e86e-4ee8-b199-040eadc75684 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Lasot: A high-quality benchmark for large-scale single ob- ject tracking
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9787d40e-5a6e-4c78-81ee-db889399770f · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Lasot: A high-quality large-scale single object tracking benchmark
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22edd8b8-6ce5-4327-9e52-b260742f54cb · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Siamese Natural Language Tracker: Tracking by Natural Language Descriptions with Siamese Trackers
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 097562a7-14eb-456d-9e3d-68b66ceabb53 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Real-time visual object tracking with natural lan- guage description
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba7bb2d-c6ac-44a7-a639-580704c5b823 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Siamese natural language tracker: Tracking by natural lan- guage descriptions with siamese trackers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46e13a0a-1579-47ae-a1b5-a1dafd7713b8 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Enhancing Vision-Language Tracking by Effectively Converting Textual Cues into Visual Cues
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5d49d5c0-266c-46aa-9cb6-dab61f2a3a1c · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Memvlt: Vision- language tracking with adaptive memory-based prompts
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfaa284f-62df-4dcf-88c4-6ad7470bb5a4 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Narrlv: Towards a comprehensive narrative-centric evaluation for long video generation models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6956bd4c-b23c-4673-b875-76a169d735b7 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking CSTrack: Enhancing RGB-X Tracking via Compact Spatiotemporal Features
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d89e1c7f-6819-40aa-aff1-97aad547f407 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Aiatrack: Attention in attention for trans- former visual tracking
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c33d4be3-f5c2-41f3-b167-71479d44de01 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Generalized relation modeling for transformer tracking
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93a04a3d-442f-4d53-ad55-f858ecfda2b4 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Divert more attention to vision-language tracking
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb14e636-5ffc-410d-b9d1-a77f337883d7 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Learning Target-aware Representation for Visual Tracking via Informative Interactions
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98daaa08-b2b2-4287-b051-0c2c9679f50c · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Onetracker: Unifying visual object tracking with foundation models and efficient tuning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a46afc33-de4d-4909-8b30-324264d7ca63 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking A multi-modal global instance tracking benchmark (mgit): Better locating target in complex spatio-temporal and causal relationship
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b4e30215-6014-4ec2-8208-1f21191831da · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Global instance tracking: Locating target more like humans
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8ce4568f-cf5a-4c12-9e6e-1b97b31b4db1 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Sotverse: A user- defined task space of single object tracking
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 203bd55e-7cc5-47bf-b790-74b737a99935 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Got-10k: A large high-diversity benchmark for generic object tracking in the wild
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ede8583e-037f-423a-b501-d7b53fdacb4f · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking GPT-4o System Card
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2895fd57-2cb0-4fc0-aa7b-2b41d85103f9 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 107088b4-3ab2-4a67-a77a-c34de216eba4 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Zoomtrack: Target-aware non-uniform resizing for efficient visual tracking
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 78eb8ad7-c238-4d46-b3f6-3162a9256197 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Multi- modal data fusion: an overview of methods, challenges, and prospects
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 69bfe300-2ea0-496a-8c69-560bc7703344 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Cornernet: Detecting objects as paired keypoints
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fb452ae2-f0aa-4952-9742-fc8afed9e0b4 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking The Power of Scale for Parameter-Efficient Prompt Tuning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc7645a7-9796-4abd-9f0f-c3940de0187d · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking SiamRPN++: Evolution of siamese visual tracking with very deep networks
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4a85f1d4-547e-4ebe-bac4-e6fe7aa1de90 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Dtllm-vlt: Diverse text generation for visual language tracking based on llm
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9ad79807-54ae-4d3b-bb8e-c58ee2ca8bab · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking DTVLT: A Multi-modal Diverse Text Benchmark for Visual Language Tracking Based on LLM
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6b60fe2-130c-4dd0-8cfb-129f4d97ab6f · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking How Texts Help? A Fine-grained Evaluation to Reveal the Role of Language in Vision-Language Tracking
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80bed04c-ce88-4da5-a30e-3d7518189832 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Visual Language Tracking with Multi-modal Interaction: A Robust Benchmark
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb33eb2f-f8c4-4151-a44f-ee6ca2229e8d · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Cross- modal target retrieval for tracking by natural language
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f43e334b-00d2-45f0-a124-e26be12a0705 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Tracking by natural language specification
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a1c4ffdb-d657-4d1d-8c21-1fbbd62da4e5 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Tracking meets lora: Faster training, larger model, stronger performance
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6714be6a-de44-4d61-903c-ca2ce8887242 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking VMBench: A Benchmark for Perception-Aligned Video Motion Generation
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07a508e4-5330-4201-aa48-0450b0c8fce8 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11fd7396-2961-4a85-864e-47a3fd4a0165 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Decoupled Weight Decay Regularization
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c966cb9-8a2a-48eb-94a5-7cbeec8f84cc · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Tracking by natural language specification with long short-term context decoupling
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6f8b7738-ada1-479a-9ec6-523f38a3787b · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5830c71-77d9-45e3-baba-7c2d3e72a7ea · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Unifying visual and vision-language tracking via contrastive learning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0779b1d5-73b4-46c5-8549-6df17a2bbcf1 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Generation and comprehension of unambiguous object descriptions
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 79976080-0973-49b3-a477-4f80fd4cab12 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Textual tokens classification for multi-modal alignment in vision-language tracking
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fe42269f-44fb-45ca-ac5c-ffc338187d76 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Learning target candidate association to keep track of what not to track
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4449fdf6-1f5f-4cbe-b041-ec613717c58d · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Trackingnet: A large-scale dataset and benchmark for object tracking in the wild
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 79a5574f-e0ac-48cf-aab0-dda64beee430 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Vast- track: Vast category visual object tracking.Advances in Neu- ral Information Processing Systems , 37:130797–130818,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c5af3050-d178-473f-89b7-5e73428fbd17 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Faster r-cnn: Towards real-time object detection with region proposal networks
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b391ec34-bb7a-4b16-bc38-97bd873bf123 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Generalized in- tersection over union: A metric and a loss for bounding box regression
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc8af11e-9e20-48ea-968c-7be50078ca59 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Generating semantically precise scene graphs from textual descriptions for improved image retrieval
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5c1f058e-025a-4884-80c9-a491e4c32882 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Context-Aware Integration of Language and Visual References for Natural Language Tracking
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6bf2372-4fc2-4361-bd35-f84ed6e21c25 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Explicit Visual Prompts for Visual Object Tracking
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7c45bfe-6790-4790-9e82-67358170e76c · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Chat- tracker: Enhancing visual tracking performance via chatting with multimodal large language model
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 14d4b30a-8a65-4337-9b36-d4f5e94fa0e5 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking What makes for good views for contrastive learning? Advances in neural informa- tion processing systems, 33:6827–6839, 2020
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bbbed7b5-4ef0-438f-a903-a55d946ff226 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Fast-itpn: Integrally pre- trained transformer pyramid network with token migration
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dbd7e95f-68fd-44d6-b339-179c6817665e · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking LLaMA: Open and Efficient Foundation Language Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bcf2dbb-950a-4d23-8bdf-2636f4cc4953 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Attention is all you need.Proceedings of the Ad- vances in Neural Information Processing Systems, 30, 2017
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 97147d00-61c6-4eb4-b74e-8395c75fa157 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Temporal adaptive rgbt tracking with modality prompt
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 387deadd-f8dd-4f07-87bb-129b691ca116 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Transformer meets tracker: Exploiting temporal context for robust visual tracking
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 72d0a56c-f80d-4747-a621-668380a7ca51 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Unified transformer with isomorphic branches for natural language tracking
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9560beba-56ea-42d6-9f95-9808deaafd35 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Describe and Attend to Track: Learning Natural Language guided Structural Representation and Visual Attention for Object Tracking
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67692cc2-0fe9-4b8c-bd99-3655fb1e7132 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Towards more flexible and accurate object tracking with natural language: Algo- rithms and benchmark
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c2f82135-bd43-4b95-a14a-0743dd896159 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Autoregressive visual tracking
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8573aae1-3df1-4911-aebe-395c538b691b · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Dropmae: Masked autoen- coders with spatial-attention dropout for tracking tasks
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5ec8bf32-ce83-46d2-ad31-8bfc46afd404 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Object track- ing benchmark
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f56365bd-fdf5-45de-87c2-ec0588407938 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Autoregressive Queries for Adaptive Tracking with Spatio-TemporalTransformers
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08292d3a-6347-4310-8ec3-3d7c5a70afef · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Less is more: Token context-aware learning for object tracking
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f42128ac-9417-4ea7-a6bc-e4398784bd5c · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Learning spatio-temporal transformer for vi- sual tracking
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 04bb95d1-433f-459c-9f04-82d0e2e659da · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Learning spatio-temporal transformer for vi- sual tracking
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0b4b19a0-230e-43dd-b401-15f565d5fbdd · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Foreground-background distribution mod- eling transformer for visual object tracking
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7d2b5daf-8b45-4459-b1e3-7513d6eb3080 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Grounding-tracking-integration
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 014d38a9-b19e-4f0a-9053-f91b3f049817 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Joint feature learning and relation modeling for tracking: A one-stream framework
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e748ad6f-bcd2-49f6-bea1-79a96bb6bf03 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking All in one: Exploring uni- fied vision-language tracking with multi-modal alignment
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d0b53a5a-ebf2-45c5-925c-08c7161b930e · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Beyond accuracy: Tracking more like human via visual search
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3124d896-a416-495f-b9a4-0e8185b6b4d8 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking One-stream stepwise decreas- ing for vision-language tracking
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ef1aafe4-0de8-4373-8969-a42c6682dad8 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Hivit: A simpler and more efficient design of hierarchical vision transformer
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 57cc62c4-9c13-4106-8d12-6940d28778dd · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Transformer vision-language tracking via proxy token guided cross-modal fusion
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6194473d-e370-4145-ad3e-3d187e5d9f80 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Biodrone: A bionic drone-based single object tracking benchmark for robust vision
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c4381f3a-f5b6-4ff2-9db2-39b0618eefa4 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Leveraging local and global cues for visual tracking via parallel interaction network
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d76db66e-8426-42c0-a158-608c8bf3fb62 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Towards unified token learn- ing for vision-language tracking
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 389d8f8b-d52d-450d-99a9-6e2cc63643c2 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking ODTrack: Online Dense Temporal Token Learning for Visual Tracking
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b38f2ec4-7338-43a7-9a17-e75e1816ef5f · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Decoupled spatio-temporal consistency learning for self-supervised tracking
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e37c8ada-10ed-47b5-8fdd-2a2c58282a59 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Conditional prompt learning for vision-language mod- els
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f21eb56-ac82-445b-b26e-0b05d4223b36 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking Joint vi- sual grounding and tracking with natural language specifica- tion
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218710c5-93e8-4ae6-aa17-dbbad206feee · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking the ironman in red flying in the sky
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d98f2fbd-536a-4b66-a8b2-298adeaa7136 · outbound
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking plane”. In the corresponding Attl heatmap, the target word “plane
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
No inbound Pith citation observations are available.