Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:20:01.720631Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2411.15459.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:20:01.720631Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a6581c63-6da7-4362-99d1-cb7d782d3f99 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Longformer: The Long-Document Transformer
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd2fb2c-8096-4ff5-8b7e-265da37e7d88 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Fully-convolutional siamese networks for object tracking
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f30c6986-6c06-491d-bba9-9abdf184baf0 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Learning discriminative model prediction for track- ing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f49a0f57-6910-4181-98a0-74b4532cca80 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Transformer tracking
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4fa9e11a-d20c-4404-9ed7-5cb8ec8c3479 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Mixformer: End-to-end tracking with iterative mixed atten- tion
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bf97ae4-53ba-42ea-9ff9-6ee8b80dcc12 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Eco: Efficient convolution operators for tracking
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8181d726-7c28-4ac5-ba72-9e7dc704893d · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Lasot: A high-quality benchmark for large-scale single ob- ject tracking
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae09d6b5-30dd-4372-b1ab-ebb5e54f4606 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Siamese Natural Language Tracker: Tracking by Natural Language Descriptions with Siamese Trackers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57470215-477e-40d7-8267-55067c18dacb · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Real-time visual object tracking with natural lan- guage description
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1d6ea254-fd70-4f8e-a343-e45693357589 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Siamese natural language tracker: Tracking by natural lan- guage descriptions with siamese trackers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c14e1255-552f-4572-9c2e-41e5d40f1dc2 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Hungry hungry hippos: To- wards language modeling with state space models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 52ba6913-936f-46bb-a3c9-ababb534a261 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Stmtrack: Template-free visual tracking with space-time memory networks
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d65da1bc-3c72-46a5-91ba-83adef6e77eb · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Generalized relation modeling for transformer tracking
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6c3e1075-1a27-40e2-98df-385da687db40 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2a870ee-c7f5-4644-89be-513fbf5718b2 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Efficiently mod- eling long sequences with structured state spaces
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0609f4d3-4151-49de-90ff-1cbec8a71231 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Combining recurrent, convolutional, and continuous-time models with linear state space layers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af42b50c-e536-4b5c-87b0-7f2cc8fa3c63 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking On the parameterization and initialization of diagonal state space models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdfa493b-397c-484d-987b-d830f387c830 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Mambair: A simple baseline for image restoration with state-space model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d24d2ca-49b5-4a07-b55b-8a9bd4a5197d · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Divert more attention to vision-language tracking
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 300d66d5-3d69-4d8d-b254-b7218fe9d833 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Demystify Mamba in Vision: A Linear Attention Perspective
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d363e2-8650-4f9c-95e8-a155e8b296c8 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking MambaAD: Exploring State Space Models for Multi-class Unsupervised Anomaly Detection
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 028492ef-d463-4afc-ba58-9a768c3584ca · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking High-speed tracking with kernelized correlation fil- ters
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a9d49dbb-20d5-4d17-a7dd-1e83c68e2e12 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking A multi-modal global instance tracking benchmark (mgit): Better locating target in complex spatio-temporal and causal relationship
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e66ed6a9-f12e-40a1-a5f3-72b75b7224a4 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Got-10k: A large high-diversity benchmark for generic object tracking in the wild
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation abcc008d-4975-4f85-9c1e-529f71ce5104 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking High performance visual tracking with siamese region pro- posal network
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c9d99d7-0374-4249-81a0-f8ac8f4ecec4 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Siamrpn++: Evolution of siamese vi- sual tracking with very deep networks
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5adcd7d7-0fd3-4c44-b089-585491d537f2 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking VideoMamba: State Space Model for Efficient Video Understanding
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6099526f-d93d-4be3-9c6d-dfc03239752e · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Coupled Mamba: Enhanced Multi-modal Fusion with Coupled State Space Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 728fedc6-303c-4b87-8af9-30e0d4c51114 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Cross- modal target retrieval for tracking by natural language
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 98df1291-0654-4c4d-8b92-b1f54915966c · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Tracking by natural language specification
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74d970e1-2326-4893-9a36-9d851db40bcc · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking MTMamba: Enhancing multi-task dense scene understanding by mamba-based de- coders
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 48b519f6-27a5-4f43-9076-c2abea571e20 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Swintrack: A simple and strong baseline for trans- former tracking
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 355e5ee0-c2b1-4e87-bbc8-c4bc297017de · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking VMamba: Visual State Space Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e81b9aad-2ddc-401f-8b03-9df689d56cad · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Swin transformer: Hierarchical vision transformer using shifted windows
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deeb7364-25fe-410e-acb3-bb29b85d01db · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Unifying visual and vision-language tracking via contrastive learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f99182f9-a5e9-4931-ac73-4f59d7fab7da · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Generation and comprehension of unambiguous object descriptions
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd85b160-767a-490e-8d47-3020cd2b1f7e · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Long Range Language Modeling via Gated State Spaces
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4df67e8a-c9ff-485c-84fa-4dbe85e2776f · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Learning multi-domain convolutional neural networks for visual tracking
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fd97c461-71e7-4c1e-9049-546836d36abb · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Context-aware integration of lan- guage and visual references for natural language tracking
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c7707fe5-cfa6-46e8-8e09-f698e646a1e9 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb2cd656-518e-4d52-a50a-1f91343ffbc7 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Simplified state space layers for sequence model- ing
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8163d3bf-69a2-445d-b2aa-6a7fca3820d6 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Fast template matching and update for video object tracking and segmentation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9ab72c60-ba61-4856-b2c4-d80099f52fe2 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Transformer meets tracker: Exploiting temporal context for robust visual tracking
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fae2b7b6-ed6c-4e61-8012-98bcb3e3d0c4 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Towards more flexible and accurate object tracking with natural language: Algo- rithms and benchmark
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4bb6b9a7-8050-43c3-ae9a-b1f873567280 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking MambaLLIE: Implicit Retinex-Aware Low Light Enhancement with Global-then-Local State Space
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f4d9c95-edb4-4d2e-a23e-7baea397b7cd · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Siamfc++: Towards robust and accurate visual tracking with target estimation guidelines
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation dc204745-316b-43b0-a1c7-3439150e22c1 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Learning dynamic mem- ory networks for object tracking
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 40b3d93b-bb3d-4665-8ac0-c45c2f0b7d2c · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Grounding-tracking-integration
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4db2eb79-5d11-4566-9d97-bfd2b92375f5 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Joint feature learning and relation modeling for tracking: A one-stream framework
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7c4f2322-1be1-4234-8f7c-a0e7b4b8c8c2 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Object track- ing: A survey
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f97e2aee-3c2f-4b43-b45f-d1a3edcd877b · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking VFIMamba: Video Frame Interpolation with State Space Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9af8241-1c4c-4c1b-be10-4e786bb577cd · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Learn to match: Automatic matching network design 10 for visual tracking
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f064469c-6e1b-4b67-9fad-8a97627fa7df · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Joint visual grounding and tracking with natural language specifi- cation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4f811161-bb32-43ac-906e-2f40a87e6f66 · outbound
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking Vision mamba: Efficient visual representation learning with bidirectional state space model
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.