Pith. sign in

Paper Citation Record · LEDGER

Adaptive Perception for Unified Visual Multi-modal Object Tracking

As of 10 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2502.06583.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06583 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:01:38.486514Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:20:07.965249Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T15:20:13.646436Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact6
  • verified fuzzy44
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae2b2bb3-95cb-4da3-8082-80449fcb96cb · outbound

This paper cites Backbone is all your need: A simplified architecture for visual object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Backbone is all your need: A simplified architecture for visual object tracking,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.195544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.256848Z digest=sha256:ee56a82fe08db0429ebc52f9d85d4c4a7ae8e6915a08f35399bfc0b82d1b36e4

Observation 6d37118a-2ee4-416a-bb42-55bddf870ca2 · outbound

This paper cites Mixformer: End-to-end tracking with iterative mixed attention,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Mixformer: End-to-end tracking with iterative mixed attention,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.186036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.261223Z digest=sha256:50b66738bffcaffc067aa1a9387070e4eb2e2ece416ec91bca34b52fcd9bf0ad

Observation 774a935a-5458-49a9-9a30-c8e726e2e8e1 · outbound

This paper cites Joint feature learning and relation modeling for tracking: A one-stream framework,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Joint feature learning and relation modeling for tracking: A one-stream framework,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.176512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.264760Z digest=sha256:67434437ff172c169d50a4a8d2dd799e71d40152528162de5dcd39037b55b231

Observation c616e41c-04eb-485c-b31e-aaf76c4797ad · outbound

This paper cites Seqtrack: Sequence to sequence learning for visual object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Seqtrack: Sequence to sequence learning for visual object tracking,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.166811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.268681Z digest=sha256:c7ae5a8d285cf1bf4d283e1d518c693d1e992c5d18c69d346266eedc76acd7ac

Observation c4057e42-c032-4133-9eea-38776be93d97 · outbound

This paper cites Swintrack: A simple and strong baseline for transformer tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Swintrack: A simple and strong baseline for transformer tracking,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.156677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.272577Z digest=sha256:413481b2fb4fd0a2d1db10841956405526f3558fd7f60952111759e6d7f3428a

Observation f78f198c-5177-40cb-8af0-a877820e7249 · outbound

This paper cites Autoregressive visual tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Autoregressive visual tracking,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.146390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.276091Z digest=sha256:a5f5f4080075e95919b4d329236f7199ccc80949db450354f607b25d877c25de

Observation 8108e1d3-9c7a-4079-b285-da49a945e1df · outbound

This paper cites Visual prompt multi- modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Visual prompt multi- modal tracking,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.136277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.280162Z digest=sha256:4aaa8be6d5ce2c5cdecdfc9bd43b2761332babca64ae9c0730b771fc357ff0ed

Observation f1e2e9a1-1af7-4cef-ac7c-22d2c45659a9 · outbound

This paper cites Single-Model and Any-Modality for Video Object Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Single-Model and Any-Modality for Video Object Tracking

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.664936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.284479Z digest=sha256:002863c2beb32d6c18ee6ea887f396a76a430300dc7eb30510c131364d2bafd0

Observation e60ec162-1789-4f98-bd54-6dded71c3a08 · outbound

This paper cites Robust Tracking via Mamba-based Context-aware Token Learning.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Robust Tracking via Mamba-based Context-aware Token Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.289301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.289301Z digest=sha256:9118b93440afc0660795f2ac52ba812a8799f7a227e1498a8d4d0b66a3d2d160

Observation dd84f229-f4bd-42ef-aead-49efbf9852fd · outbound

This paper cites Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.293522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.293522Z digest=sha256:0a9ce68211025c960a8955cc508e7870e33aeb3a130bd121f4d2ea3977759fac

Observation 58b423eb-40aa-447b-b708-d7faf8ac53cf · outbound

This paper cites Curricular contrastive regularization for physics-aware single image dehazing,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Curricular contrastive regularization for physics-aware single image dehazing,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.125360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.297829Z digest=sha256:4dc68638071059e80df69848ed668e7b559b702a3e668667c21a9df897c2ebad

Observation 3127b424-4be4-4df1-a08f-b3163a397d14 · outbound

This paper cites Dynamic group difference coding based on thermal infrared face image for fever screening,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Dynamic group difference coding based on thermal infrared face image for fever screening,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.114649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.301951Z digest=sha256:7457bb0505eca9288fac1707f40618232eba34b5ca4c484bc80f4dd61d457a96

Observation 3abafafd-98ca-473b-8504-fccd1ef64521 · outbound

This paper cites Few-shot learning with long- tailed labels,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Few-shot learning with long- tailed labels,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.104340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.305685Z digest=sha256:384c07e60ba32b50b997f7a3b5b76a182565a8e80d74ed1d4ca047684f930de8

Observation a59b6c1b-ce2c-4e4d-8279-c96f6390d2eb · outbound

This paper cites SHaRPose: Sparse High-Resolution Representation for Human Pose Estimation,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking SHaRPose: Sparse High-Resolution Representation for Human Pose Estimation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.093728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.309654Z digest=sha256:e5d3de6280dd334e6d248ac3e6f024604fba8b03a3f885c367379bbb355204c6

Observation ec604d4e-9319-411b-95c1-ff689e4564bf · outbound

This paper cites An emotion recognition method based on eye movement and audiovisual features in mooc learning environment,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking An emotion recognition method based on eye movement and audiovisual features in mooc learning environment,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.083502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.313467Z digest=sha256:8ad72844bca551bb593a373d5fc92b87ff4b3a3a69726739a88226fb64ad4020

Observation 8c5c0319-ada0-4e34-926c-275af4dc4392 · outbound

This paper cites 3d- guided multi-feature semantic enhancement network for person re-id,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking 3d- guided multi-feature semantic enhancement network for person re-id,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.072896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.317387Z digest=sha256:4d80a6e30e7c88c34616891554b39d6346f1debe928bb8300eab39e64ce981bc

Observation 8bc50acc-e8fb-43fc-99a0-b31dba64f507 · outbound

This paper cites Multi-branch enhanced discriminative network for vehicle re-identification,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Multi-branch enhanced discriminative network for vehicle re-identification,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.061548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.321291Z digest=sha256:b3b6f36de8829d13b6e1634ff459d4c9545d45b28422345aef65f05c6cea4cbf

Observation 75e1fd3c-61ea-4eff-8542-49ef97f0067d · outbound

This paper cites Guided Real Image Dehazing using YCbCr Color Space.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Guided Real Image Dehazing using YCbCr Color Space

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.626079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.325098Z digest=sha256:4a020ad6a240ec163d316e806531edd0a58f57e5aa0cb2d58542fb04e182be61

Observation b0a3ff3b-19c4-438f-97ed-c9fe611ac37c · outbound

This paper cites OneTracker: Unifying Visual Object Tracking with Foundation Models and Efficient Tuning.

Adaptive Perception for Unified Visual Multi-modal Object Tracking OneTracker: Unifying Visual Object Tracking with Foundation Models and Efficient Tuning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.612183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.329265Z digest=sha256:a58daeaa00ab79c9bc4db6795d0985b5699653a720df6b2c03cd274ae3467d4d

Observation b4afefca-df2a-48da-983e-0f1579b8a5fa · outbound

This paper cites Prompting for multi-modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Prompting for multi-modal tracking,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.050806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.333793Z digest=sha256:95316a5b3434f970f8117ca0d5ae67a02e7d6787d952ea1a16c9f79cc0b6897d

Observation 586c5bbc-acf0-469d-a51b-dce848553d4e · outbound

This paper cites SDSTrack: Self-Distillation Symmetric Adapter Learning for Multi-Modal Visual Object Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking SDSTrack: Self-Distillation Symmetric Adapter Learning for Multi-Modal Visual Object Tracking

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.596444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.338768Z digest=sha256:e7ec3ab54acc9a42958ec15caf34e63a342962ca29f11400bfbdcc8fcf402025

Observation 2bc63b7a-a1ee-4a72-9de5-d6e6fbe98bda · outbound

This paper cites Bridging search region interaction with template for RGB-T tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Bridging search region interaction with template for RGB-T tracking,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.040569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.343304Z digest=sha256:5528b76c5edde5b4cd583006464a6d78ea2e6720766d21c77e0ad5202c462330

Observation 796e2777-6c9d-4570-bfd0-15090e63e4a1 · outbound

This paper cites Bi-directional adapter for multi- modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Bi-directional adapter for multi- modal tracking,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.029978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.347554Z digest=sha256:50b30095a96c296153a9a8c79d46530f4c551f1adc8758254f7019ab49f91976

Observation 3ea5e07d-5d7e-4700-92e8-a06ce244ce02 · outbound

This paper cites Spiking transformers for event-based single object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Spiking transformers for event-based single object tracking,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.020710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.351534Z digest=sha256:df9f650075559430c4b4897879bac9d356a2cae52e81ed855efaad1604ccc28a

Observation 8b1f559c-8ee2-4d90-bea0-50eae7c83025 · outbound

This paper cites Lasot: A high-quality benchmark for large-scale single object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Lasot: A high-quality benchmark for large-scale single object tracking,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:39.009603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.355559Z digest=sha256:24e708a61357ecc8e7fe564de2551522124ca6e6ec4f0badf06a5500904eb7eb

Observation c3d2d9ec-e421-4119-b0ba-26dbeb35d237 · outbound

This paper cites Got-10k: A large high-diversity benchmark for generic object tracking in the wild,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Got-10k: A large high-diversity benchmark for generic object tracking in the wild,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.998252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.359442Z digest=sha256:407cb60e2ce9ae04d8330616379a90b02b608207bdbdda8b709366d968e4f9f4

Observation 194a8d21-c048-4cd1-8f62-cf8c9c9d92f2 · outbound

This paper cites Trackingnet: A large-scale dataset and benchmark for object tracking in the wild,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Trackingnet: A large-scale dataset and benchmark for object tracking in the wild,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.986870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.363352Z digest=sha256:3a16b55ebecf60f02c04fde10b1113e91de05be191bcbc1d002acba959ba2593

Observation a73088f9-0cc1-41ce-8626-4b00a8f9b49c · outbound

This paper cites Lasher: A large-scale high-diversity benchmark for RGBT tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Lasher: A large-scale high-diversity benchmark for RGBT tracking,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.973857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.367242Z digest=sha256:4d4c277c4aaedfa5a04bb1e7e2d1a4d397a8ac886d750301bfba16a7dbcfd257

Observation 96f5ea52-3207-43f2-8308-3c3de07a75ce · outbound

This paper cites RGB-T object tracking: Benchmark and baseline,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking RGB-T object tracking: Benchmark and baseline,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.961719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.371104Z digest=sha256:0c3f193548ac7352b4183c4455464fc8b6b7342da3d9054915b1d2b5f7a526bf

Observation ab07304f-7740-4ad3-ab8b-7ecb33720098 · outbound

This paper cites Depthtrack: Unveiling the power of rgbd tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Depthtrack: Unveiling the power of rgbd tracking,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.949221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.374858Z digest=sha256:1d66b90b276f8e9b83e8df76262553dc67f99c784963a22996b5101b40fe66fa

Observation 3aa9dbad-8a55-40b9-93fd-6f09d8839568 · outbound

This paper cites The visual object tracking vot2015 challenge results,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking The visual object tracking vot2015 challenge results,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.936593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.378540Z digest=sha256:46916b3a36e2a32c9cb04ca60175c3a1ab87c534ffea0fe5d6d9d0cb1016c89c

Observation 88f7be3a-626a-4d16-8af3-b4b9962248e6 · outbound

This paper cites Visevent: Reliable object tracking via collaboration of frame and event flows,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Visevent: Reliable object tracking via collaboration of frame and event flows,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.924562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.382631Z digest=sha256:7f4c6253b774980ee402e955b29c69ee6340170ddc6c22678ce238d7f992b76c

Observation 6115b44c-0ead-413f-9f68-3ca71ac655d3 · outbound

This paper cites Unified-io: A unified model for vision, language, and multi-modal tasks,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Unified-io: A unified model for vision, language, and multi-modal tasks,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.386671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.386671Z digest=sha256:9a9f089ea9329c130d97975fec8aebc105b0a00991d319e2d117c94837cb2500

Observation 59d77813-4c96-4129-8b07-9a6faae8009e · outbound

This paper cites Imagebind: One embedding space to bind them all,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Imagebind: One embedding space to bind them all,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.904614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.390455Z digest=sha256:30fd59c993fb2627a78e2d02a6b82a693745f43e1da18d6552901fcd7017d933

Observation a63fec7a-da80-4305-abd7-d572bb5a9d0d · outbound

This paper cites MUTEX: Learning Unified Policies from Multimodal Task Specifications.

Adaptive Perception for Unified Visual Multi-modal Object Tracking MUTEX: Learning Unified Policies from Multimodal Task Specifications

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.393426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.393426Z digest=sha256:acddc198466ba5af0f3585ecaf654c58d86601b96992920657a07eb6ac291827

Observation 766d9f0a-ffd3-4334-8fdf-9469ea52f920 · outbound

This paper cites Siamese Vision Transformers are Scalable Audio-visual Learners.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Siamese Vision Transformers are Scalable Audio-visual Learners

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.397539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.397539Z digest=sha256:053b240a09b4130c079a2c4d5249aac76f1224ca122abe0a86b6247685d354f1

Observation f2bfc63a-4f65-4813-bcf2-2e19670f1d95 · outbound

This paper cites A unified audio-visual learning framework for localization, separation, and recognition,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking A unified audio-visual learning framework for localization, separation, and recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.891981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.401184Z digest=sha256:cd0f7ead9a04eb57b4e8ca3b2eac3c1c82da4065495f75073a3c99e4fa0dfdd7

Observation 0c1a4fa6-a92c-4773-85e1-692bb4cdda91 · outbound

This paper cites Learning visual representation from modality-shared contrastive language-image pre-training,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Learning visual representation from modality-shared contrastive language-image pre-training,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.878140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.404589Z digest=sha256:c21b30fa1200c8bfec4cb3c4f2d7f6fba2026035d649fbda362a848a0ec74cd8

Observation cdfd3f0b-d730-4075-b0d0-e834ec74cfb6 · outbound

This paper cites Decoupled weight decay regularization,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Decoupled weight decay regularization,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.866174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.407603Z digest=sha256:bd1c3b3508460b87bf1f0ac510a4511d74cb02b2aaf9193ca88aa18437fe8c34

Observation 5942d894-1489-408f-988d-916e45a4d3db · outbound

This paper cites Weighted sparse representa- tion regularized graph learning for rgb-t object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Weighted sparse representa- tion regularized graph learning for rgb-t object tracking,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.855104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.410684Z digest=sha256:a78b38899b4f8f886159f4733f8d1993eb6b81d998eb7cb38df3265c8b7f26b0

Observation 0371c9fe-b5e7-4fc6-856f-357ab1008aa2 · outbound

This paper cites Generative-based fusion mechanism for multi-modal tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Generative-based fusion mechanism for multi-modal tracking,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.845254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.413706Z digest=sha256:105b19550f70fcb05a5c25c9653971ff3bd4182c474e28c9fbd13d58e937ab0b

Observation 59dbba3c-c935-4ac3-b145-53800b2d829f · outbound

This paper cites Transformer tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Transformer tracking,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.835252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.416861Z digest=sha256:c2364c1ffd1888c91c6987c2ccdf22f70623940ddfa00bb2cc199a2fc48c6336

Observation e0c0914f-3e9b-45e1-8846-ba04d4acff3b · outbound

This paper cites Learning spatio-temporal transformer for visual tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Learning spatio-temporal transformer for visual tracking,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.420525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.420525Z digest=sha256:e088b73084297fb13d4e9ad5dd36a0e194912abae66c287b2fbda954c94af3cb

Observation e4ff1443-2a31-4fbf-be4a-4dacc6dfc7e1 · outbound

This paper cites Aiatrack: Attention in attention for transformer visual tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Aiatrack: Attention in attention for transformer visual tracking,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.424619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.424619Z digest=sha256:f31819bd8c7932f6d1c208f628ec37b55f2a2c2e031db4b91490b9fee7933568

Observation 8e89baee-71c9-431c-8944-3a5acd28a9be · outbound

This paper cites Rgbd1k: A large-scale dataset and benchmark for rgb-d object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Rgbd1k: A large-scale dataset and benchmark for rgb-d object tracking,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.812467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.428469Z digest=sha256:5eb9f3db5ff3cfa1f6925fc7e0584b3c42dc60675cc4be5892a53aac5d83611d

Observation b0677cfa-d238-45bc-aacf-7fea2b8c6926 · outbound

This paper cites Focal loss for dense object detection,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Focal loss for dense object detection,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.432244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.432244Z digest=sha256:31618c5ddd9612662477ca8402d77474e3d17cb71c37d6b14399714f832191d6

Observation 823456f3-4507-4856-9a7e-96e38f8ab8fb · outbound

This paper cites Generalized intersection over union: A metric and a loss for bounding box regression,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Generalized intersection over union: A metric and a loss for bounding box regression,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.794074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.436214Z digest=sha256:1996e1e8838d15205ef383e49c6812c6d57fda8c8609ec8f5a41c961103e3320

Observation daa29079-8457-4bd8-9f6c-50806d158431 · outbound

This paper cites Depthtrack: Unveiling the power of RGBD tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Depthtrack: Unveiling the power of RGBD tracking,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.782160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.440102Z digest=sha256:0f7c040db257fd6b74bab171137daf789f51bfd427825c0b29d66df7b8a3b8bd

Observation d5690a09-d456-4f20-ab09-aecd738e4914 · outbound

This paper cites Transformer tracking via frequency fusion,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Transformer tracking via frequency fusion,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.770507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.443660Z digest=sha256:de5791202d2e611b7a643c6f0c81439f3c7d7f0de75ef77f9b5b480b69de4ba9

Observation 33dce51f-e2f7-4125-a47a-ad004815f6a1 · outbound

This paper cites Multiple source domain adaptation for multiple object tracking in satellite video,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Multiple source domain adaptation for multiple object tracking in satellite video,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.758488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.447390Z digest=sha256:a05d56b02ea1385a11eea1eaa0c0d7962a0002144f6aa46761b0c08fccb240d0

Observation 7dd02a68-a0b1-451a-9eea-7850cf970597 · outbound

This paper cites Explicit visual prompts for visual object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Explicit visual prompts for visual object tracking,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.747050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.451209Z digest=sha256:0b7bd73d04d4e6c7364696feb6df117617047c6b2b9ebf15446adbf24a90eea9

Observation 6b9fe4db-d580-487f-96d2-73ff40e0dd65 · outbound

This paper cites Autoregressive queries for adaptive tracking with spatio-temporal trans- formers,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Autoregressive queries for adaptive tracking with spatio-temporal trans- formers,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.455226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.455226Z digest=sha256:ad9af7b296fe4abaad43f2e7753c79207c0f06983a49194ba0c2be572dfd78a8

Observation fcdd4def-e275-4da9-9e85-5a9eeb715089 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Learning transferable visual models from natural language supervision,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.726920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.458954Z digest=sha256:3d273827643ef10c593918602af1e71d424d659807b1308ec20f933e7689e470

Observation 83c21fd7-bdef-41bb-8554-d16961611ed2 · outbound

This paper cites Towards modalities correlation for rgb-t tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Towards modalities correlation for rgb-t tracking,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.713885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.462662Z digest=sha256:59403731666cb13d9491110a938bf024efe3cd195e21f51d006dc45dae5fbbaf

Observation d9553625-d158-4c93-a018-9920580c311d · outbound

This paper cites RGBD1K: A large-scale dataset and benchmark for RGB-D object tracking,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking RGBD1K: A large-scale dataset and benchmark for RGB-D object tracking,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.702132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.466373Z digest=sha256:4a26863bed99837165d1203a2688a9fc9f6d6d811641770a9d5f7c3d915f15b8

Observation 64227ba3-8b07-4445-96dc-68280b42ce2f · outbound

This paper cites Siamban: Target-aware tracking with siamese box adaptive network,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Siamban: Target-aware tracking with siamese box adaptive network,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.689883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.470170Z digest=sha256:bfa2ca13da40b40366aa059c9643878ddda059b4413e896e94c8b52a2ad48e98

Observation f0f4888b-998b-4d60-ac2b-34faa6667c2a · outbound

This paper cites Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Reliable Object Tracking by Multimodal Hybrid Feature Extraction and Transformer-Based Fusion

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.557849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.473979Z digest=sha256:45b44eb9a6bed61d5f768e6ec0489e95a63eff71f99522fbf4f71c88c572696c

Observation 63703f02-2ee1-466c-a648-9804ce6f755a · outbound

This paper cites Depthrefiner: Adapting rgb trackers to rgbd scenes via depth-fused refinement,.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Depthrefiner: Adapting rgb trackers to rgbd scenes via depth-fused refinement,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:01:38.677307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.478408Z digest=sha256:e13fcaa7f7d705046bb9f975c1d79f78395d87620522d17ece91798ac257c898

Observation 1ab43ecb-3c4d-4d1e-b0cb-799a3e344e0e · outbound

This paper cites Cross-modulated Attention Transformer for RGBT Tracking.

Adaptive Perception for Unified Visual Multi-modal Object Tracking Cross-modulated Attention Transformer for RGBT Tracking

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-08T15:01:38.539120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T15:01:38.482277Z digest=sha256:bc8ab486d6483ccc485b63a7b309ba6047c0498818ccd4d98d35ca66811075dd

Observation 876085fb-394e-4348-8dda-c6647c7db7cc · outbound

This paper cites RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba.

Adaptive Perception for Unified Visual Multi-modal Object Tracking RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T15:01:38.486514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:01:38.486514Z digest=sha256:f1c0b2cccac3ffcdd949a75c7b21930af561c78123af065552eabca67c75e9ca

Pith citing papers

Observation 16c70ea3-faff-4ac1-8bd3-9e1cdca7e96a · inbound

Explicit Context Reasoning with Supervision for Visual Tracking cites this paper.

Explicit Context Reasoning with Supervision for Visual Tracking Adaptive Perception for Unified Visual Multi-modal Object Tracking

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:20:13.728895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:20:07.965249Z digest=sha256:78506e05e103e4c132a0030b495dbef96a943ce0b94d1dba946b42f8fa296b7c