Pith. sign in

Paper Citation Record · LEDGER

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification

As of 11 August 2026, this Paper Citation Record lists 100 of 102 outbound references and 1 inbound Pith citation observation for arXiv:2501.16811.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.16811 v1

Coverage vector

measured 100 of 102 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T10:34:01.517256Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T22:09:28.433165Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T22:09:30.691155Z

Reference resolution

100 of 102 outbound references displayed

  • verified exact2
  • verified fuzzy67
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation be51c658-c095-4322-ab77-4ae19d16f998 · outbound

This paper cites Spatio-temporal repre- sentation factorization for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Spatio-temporal repre- sentation factorization for video-based person re-identification

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.402053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.402053Z digest=sha256:72a98d3cca55ef6bde91b8635c7b43e3db0e425297e9bf9af2f8734b6f260ee2

Observation 5331d4ee-28fd-4796-873e-5450b91e498e · outbound

This paper cites Salient-to-broad transition for video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Salient-to-broad transition for video person re-identification

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.411694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.411694Z digest=sha256:19b8f57c5d8c4030cc357e2a72202ba0714d8e9c391b1f6b4303819e726fef7f

Observation 56d5aa24-dbd6-4916-9944-84dbd68fdb78 · outbound

This paper cites Event-guided person re- identification via sparse-dense complementary learning.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Event-guided person re- identification via sparse-dense complementary learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.420728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.420728Z digest=sha256:40bb70f66b14e4760347f5de0c4e29f495e68ad6034e3e9ad231ca58266b8959

Observation 90154ed6-1fe0-49d7-a9df-44bf565a5cbd · outbound

This paper cites End-to-end object detection with transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification End-to-end object detection with transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.436696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.436696Z digest=sha256:5303425058fe254266dcaa7930095e89d6a19e36e240f9eb254d00a27c99d9e5

Observation ecdd2d22-8b68-49ec-a390-b9c5ee31bea8 · outbound

This paper cites Video person re-identification with competitive snippet- similarity aggregation and co-attentive snippet embedding.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Video person re-identification with competitive snippet- similarity aggregation and co-attentive snippet embedding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.445164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.445164Z digest=sha256:f25a3f9060ac37c7a24d6e4e20409ca6e33f7f59e2688142b7bb072f7d04c6b9

Observation 65b52554-710d-4ba4-8361-254266588c79 · outbound

This paper cites an unresolved cited work.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.455961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.455961Z digest=sha256:683779d505de742f3262e5df0f56add630cab4c3fbf94ed85168de4ff30c5215

Observation 433ae3d1-7382-441d-aca6-116b6032def2 · outbound

This paper cites Diffrate: Differentiable compression rate for efficient vision transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Diffrate: Differentiable compression rate for efficient vision transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.467257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.467257Z digest=sha256:5a610d18e0d9f3a6f023beb8ef65569693df13f8436b0802f27f4aa73ae6516c

Observation 00f87cd1-c2b5-4f55-b84a-5f9968001dca · outbound

This paper cites Reality3dsketch: Rapid 3d modeling of objects from single freehand sketches.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Reality3dsketch: Rapid 3d modeling of objects from single freehand sketches

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.487323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.487323Z digest=sha256:b4b980a8e91b8080ef9b4e2e2ef520e234944adb3ab63406c29690396c54a096

Observation 90c2b3f2-02d5-4622-905b-d27dad233155 · outbound

This paper cites Abd-net: Attentive but diverse person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Abd-net: Attentive but diverse person re-identification

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.495338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.495338Z digest=sha256:c03adc1aa9517845c7b6f022ab0f34f550886d7a0663237b36144dc135a7b649

Observation e96a7da9-8e65-4853-bd8f-95f28a4ac6fb · outbound

This paper cites Deep3DSketch: 3D modeling from Free-hand Sketches with View- and Structural-Aware Adversarial Training.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Deep3DSketch: 3D modeling from Free-hand Sketches with View- and Structural-Aware Adversarial Training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.504460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.504460Z digest=sha256:846530cd6ca487fff8cfbf061a47e6b4fb65dab2ee666b50efc022030356765f

Observation 66d5905d-e82f-4c19-ba91-494ee22e5e33 · outbound

This paper cites SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.512010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.512010Z digest=sha256:2905e3327f45391bd8261395590f93fe340a69cf84c6aff0fb0afdc2709f5522

Observation 1c5017d8-2451-4f00-bb52-7ab188e989c6 · outbound

This paper cites Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.519569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.519569Z digest=sha256:d981051544191972363dc7dc96ebfd8e2409e6fcebcca36464e80207c3fec231

Observation eef637cf-3f38-4267-ae1b-5f4386982dc5 · outbound

This paper cites Sam-adapter: Adapting segment anything in underperformed scenes.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Sam-adapter: Adapting segment anything in underperformed scenes

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.530575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.530575Z digest=sha256:4a85318279e22560c3b4b3d0a19796053102c134f880b8b845dc49b0103eea8c

Observation ebd7810f-a237-4ada-9c16-2cfd45f69fbf · outbound

This paper cites Masked-attention mask transformer for universal image segmentation.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Masked-attention mask transformer for universal image segmentation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.546756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.546756Z digest=sha256:c32b583ac253b6766c18ff2b0ed6ac950a347ccf8df8edf9c6f6e280012c0da0

Observation 3c1a51ba-db07-409b-9872-9cac7dab837b · outbound

This paper cites Towards accurate post-training quantization for vision transformer.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Towards accurate post-training quantization for vision transformer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.554536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.554536Z digest=sha256:1e348417106b59522f9bbe8ab74a19be555cfa5c91b5790ac178d71e284b275e

Observation 0f9310be-8449-4ff2-ab3d-872ad65edcf0 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.564210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.564210Z digest=sha256:a8abdbcfe1c6ca429f0743247096e5c5f6dd0dcb662e1c0b9d81b7ded9967b34

Observation f6aab266-ce41-40d9-85be-40e8511a2972 · outbound

This paper cites Eventful transformers: Leveraging temporal redundancy in vision transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Eventful transformers: Leveraging temporal redundancy in vision transformers

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.572592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.572592Z digest=sha256:4e79a6c7282b3b2a03674eaaf6b87da09b7338953290be5fb477ecc45282115c

Observation 54c3c877-9305-460e-8706-b375752924b5 · outbound

This paper cites Video- based person re-identification with spatial and temporal memory net- works.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Video- based person re-identification with spatial and temporal memory net- works

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.582628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.582628Z digest=sha256:58318c3dca2eb4843cc15e0cb2bb6d5006fa706370a51ccbd1b45106b167f8f2

Observation 656a2bd3-e1a5-49af-a9fc-74650e0963de · outbound

This paper cites Motion adaptive pose estimation from compressed videos.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Motion adaptive pose estimation from compressed videos

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.608030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.608030Z digest=sha256:d8b28850523c73b812fe2d2e6ae248ef19d902f2bc53cd81f3901ca283a679eb

Observation 0480bfbf-e940-47fd-9373-93d6ac9a172c · outbound

This paper cites Sta: Spatial-temporal attention for large-scale video-based person re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Sta: Spatial-temporal attention for large-scale video-based person re- identification

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.625173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.625173Z digest=sha256:9d11b6361f20e37d7e3b3de880eba41062f704bfdc9144bc1f151f14490d7784

Observation f5749dd6-6132-4cd6-9556-a1ecc125c45d · outbound

This paper cites SparseFormer: Sparse Visual Recognition via Limited Latent Tokens.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification SparseFormer: Sparse Visual Recognition via Limited Latent Tokens

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-10T10:34:01.847747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.630751Z digest=sha256:de92c1b28b98199ea278508042d7087476a2d3eb284d11cfb45fb2fd5019a448

Observation 3f4a3073-e7f5-4911-b942-92c7be60e771 · outbound

This paper cites Appearance-preserving 3d convolution for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Appearance-preserving 3d convolution for video-based person re-identification

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.639133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.639133Z digest=sha256:98e573bb408992762835108ca6f142a357efa307d725bbdf8d4a224e806e59be

Observation d4c22ab1-99fb-4af4-b6f4-033f63cdd966 · outbound

This paper cites Flatten transformer: Vision transformer using focused linear attention.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Flatten transformer: Vision transformer using focused linear attention

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.646378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.646378Z digest=sha256:ae09ebc3c10eeada726ef5f7742aaa9ab5b85363ad04ffd3365127f40a501def

Observation acb84bea-0f12-432f-985d-9a08d8dfd446 · outbound

This paper cites Deep residual learning for image recognition.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Deep residual learning for image recognition

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.655144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.655144Z digest=sha256:0f786126130b6a90b212e90b367db0c808a8ea5c4ef44edb9f76c5efaf11c452

Observation 4d63f547-094d-41fa-b1c5-28796c2e2061 · outbound

This paper cites Transreid: Transformer-based object re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Transreid: Transformer-based object re-identification

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.661142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.661142Z digest=sha256:734c87ae203dbf8a45b433777be97a2241c1bd1fe905600555f6f10ed4401ae4

Observation c1442edd-9dfb-4961-9f3d-beb9a487176d · outbound

This paper cites Bicnet-tks: Learning efficient spatial-temporal representation for video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Bicnet-tks: Learning efficient spatial-temporal representation for video person re-identification

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.679988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.679988Z digest=sha256:4087528102c8354c9ba383f843e2915b804326680f2febf1d782dd175fa470a4

Observation 46249f9c-6e96-4396-a7bc-f6ea6c22c8d4 · outbound

This paper cites Temporal complementary learning for video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Temporal complementary learning for video person re-identification

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.687909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.687909Z digest=sha256:4ceb23b90caecc3665e27462facba7e50002ed5d0e7cf58651d484452a0ecb19

Observation 72827808-4fda-42c7-86af-a46a71c96b7a · outbound

This paper cites Vrstc: Occlusion-free video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Vrstc: Occlusion-free video person re-identification

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.661983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.693101Z digest=sha256:017789e7a63f63eae5f12fb1575df32884119ce14f4178aa5cecd857cf4297e8

Observation 9bc71934-e46b-4d09-a639-34b90dd9c969 · outbound

This paper cites Orthogonal transformer: An efficient vision transformer backbone with token orthogonalization.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Orthogonal transformer: An efficient vision transformer backbone with token orthogonalization

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.620085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.698633Z digest=sha256:f3801bd584f77ec8c07d34d6234ec7a2f6a349269ec8290846d6307431c331fc

Observation 1517c3fa-2eb8-4df8-88d9-3fbec8c5e1a6 · outbound

This paper cites Reasoning and tuning: Graph attention network for occluded person re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Reasoning and tuning: Graph attention network for occluded person re- identification

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.581239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.703468Z digest=sha256:eac401a4ad9a01fc7ef2d3ac87ff762df7b251819cc22bb01af47debc7a2236b

Observation c2634ffa-1e9e-4648-83f9-efd63555e76d · outbound

This paper cites En- hancing person re-identification performance through in vivo learning.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification En- hancing person re-identification performance through in vivo learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.553697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.708947Z digest=sha256:3048b9c19950edf58d45fa811cfc9da2de430374307556675c2d1c408230f994

Observation d36ecef9-e7c6-440e-935a-58f65b41a0f7 · outbound

This paper cites Discrete Latent Perspective Learning for Segmentation and Detection.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Discrete Latent Perspective Learning for Segmentation and Detection

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-10T10:34:01.796964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.724221Z digest=sha256:ebb6d3d9f235fc0766a81e2884e05b1168f08ec8dd7899f6e0a93c4628ea987e

Observation 7b913b73-bda9-4bde-957e-29ba3d671c27 · outbound

This paper cites Fast decoding in sequence models using discrete latent variables.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Fast decoding in sequence models using discrete latent variables

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.522693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.741056Z digest=sha256:a8583040aba1577d41e36b76444e7cf7591fe0dec2e5f195fd4f831a78e5a0d1

Observation c4540df0-bd28-42ee-8f57-48ffd3778bba · outbound

This paper cites Spvit: Enabling faster vision transformers via latency-aware soft token pruning.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Spvit: Enabling faster vision transformers via latency-aware soft token pruning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.485282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.754533Z digest=sha256:055fcb8bef51d84d1dc5790bc22737172e23850a43b74fd0e85653381ccc06f1

Observation 2bd61ef7-1de5-4404-b5a0-54fe6bd6d05b · outbound

This paper cites Global-local temporal representations for video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Global-local temporal representations for video person re-identification

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.447056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.766983Z digest=sha256:df1573796309c68cd5fff7cc43963683ce7e5b8177a241cc145ba2f195a1522f

Observation 6afe1780-49fc-4568-ba2c-0b287a391b1b · outbound

This paper cites Multi-scale 3d convolu- tion network for video based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Multi-scale 3d convolu- tion network for video based person re-identification

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.419340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.783289Z digest=sha256:20377ddfb3575ef10c659ff572eed0f4f32f199d5f13376887f46be4d3fcc1fd

Observation 81f863f8-b2ba-4b09-9d35-36c03e7cb53a · outbound

This paper cites Diverse part discovery: Occluded person re-identification with part-aware transformer.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Diverse part discovery: Occluded person re-identification with part-aware transformer

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.374263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.791397Z digest=sha256:ba0360e7c9df0340822330cd561e3f51da9282b23552d9de9af24503dd21ab28

Observation 8d2f5ad6-4692-4481-94fe-254745f1180a · outbound

This paper cites Svitt: Tem- poral learning of sparse video-text transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Svitt: Tem- poral learning of sparse video-text transformers

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.340068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.800153Z digest=sha256:e48541d1bb4f47a3eb4d3bae1d1a865048d39d10f22967f9de40ec93d4bc5c96

Observation 58878bff-9f9c-4afe-a621-d8b63b922a9b · outbound

This paper cites Efficientformer: Vision transformers at mobilenet speed.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Efficientformer: Vision transformers at mobilenet speed

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.308115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.807634Z digest=sha256:daef7114317bdc1c721a1e54f5558b473f920f4b46ec1d278a5155c645a5020e

Observation 849ba5bd-9fbd-4a2a-a3fd-c1585051c39a · outbound

This paper cites Evit: Expediting vision transformers via token reor- ganizations.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Evit: Expediting vision transformers via token reor- ganizations

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.274361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.820143Z digest=sha256:4fe1485a0a7e1a30cf69d2f54db69572da92db78d975c41126c6901ba5664300

Observation 3c36b440-baa5-418a-a0da-b9b18c4916ea · outbound

This paper cites Not all patches are what you need: Expediting vision transformers via token reorganizations.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Not all patches are what you need: Expediting vision transformers via token reorganizations

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:00.836138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:00.836138Z digest=sha256:ddc25e704ccddf0799f01df0e841c16e829eef357974ebd07d87614674fc7974

Observation cde5d3dd-fed4-471b-87f7-cccb5775bfdd · outbound

This paper cites Supervised masked knowledge distillation for few-shot transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Supervised masked knowledge distillation for few-shot transformers

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.208933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.845032Z digest=sha256:578f7e5d092355798f634ca9edcf44380526f0f4d0885e1a1228c0f2da384ce4

Observation d0bd5b39-b890-4dba-9249-d399d71bbefd · outbound

This paper cites A versatile model for packet loss visibility and its application to packet prioritization.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification A versatile model for packet loss visibility and its application to packet prioritization

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.177601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.855441Z digest=sha256:79c900e359f56a5d068ec370daee742fa540f23c559da71b1164b73bde208e17

Observation 9f6afe2c-12cc-44d1-a3c9-2ed433e9d0a5 · outbound

This paper cites Learning modal-invariant and temporal-memory for video-based visible-infrared person re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning modal-invariant and temporal-memory for video-based visible-infrared person re- identification

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.122754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.865209Z digest=sha256:588671b00ce5e038126def58724b2f999f2965d8fa8d05d0158b5af3a35e904d

Observation 4f6bf0da-4e50-448a-a1de-2e66d2fb3113 · outbound

This paper cites Video-based person re-identification with accumulative motion context.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Video-based person re-identification with accumulative motion context

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.088878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.875964Z digest=sha256:77a49e3c74d6270b2b26dbafac4fb1f268e4973a3422fdac0170740e412ef38c

Observation f4d99539-83c9-4f2b-8c1f-82c38fbe9e05 · outbound

This paper cites Spatial-temporal correlation and topology learning for person re- identification in videos.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Spatial-temporal correlation and topology learning for person re- identification in videos

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:04.034501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.882999Z digest=sha256:0368b7d0c032638dd4e5b24998c6142b323932aff513fd49eb3dac07897d4fda

Observation c6fa7295-253d-468f-a89e-dd9337dd5827 · outbound

This paper cites Fre- quency information disentanglement network for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Fre- quency information disentanglement network for video-based person re-identification

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.999890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.888267Z digest=sha256:6f01874142b15fb067d1c0e446a8ed7a3574f9459ef6669e0fe327e66e43b54c

Observation 54402849-bfa5-4d3f-952c-485f8bc285b6 · outbound

This paper cites Deeply coupled convolution–transformer with spatial–temporal complementary learning for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Deeply coupled convolution–transformer with spatial–temporal complementary learning for video-based person re-identification

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.974301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.897975Z digest=sha256:92fb3dc046d8a9dc205a42c653466e43f74c40d867aa2eb7b0894e7bb7079330

Observation abdb1110-f433-42b1-b98b-5c1a6d9e8df6 · outbound

This paper cites Watching you: Global-guided reciprocal learning for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Watching you: Global-guided reciprocal learning for video-based person re-identification

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.940849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.915330Z digest=sha256:ac3dcd885c2b9c27e91f4a687e5803f46741ff60ba2fbd740481ce71fcc63430

Observation 7b1846f5-bab4-4e15-91f6-5f4efc4a705c · outbound

This paper cites Noisyquant: Noisy bias-enhanced post-training activation quantization for vision transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Noisyquant: Noisy bias-enhanced post-training activation quantization for vision transformers

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.899867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.924429Z digest=sha256:87a8d7a79905df77da964965ad14a98081a3939e8ca5c2ec05c9c2bb9638b8f3

Observation 1d5f7440-c864-4760-9294-77c39c6b51a7 · outbound

This paper cites Post-training quantization for vision transformer.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Post-training quantization for vision transformer

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.870990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.930507Z digest=sha256:f1ded742f817fe9ba777482cb54fc4f73bad44cbbb2abd564ff4ed29c3f94cda

Observation ffc1e182-1a4c-48d5-a5ad-84c3e9cf45cc · outbound

This paper cites Label-guided attention distillation for lane segmentation.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Label-guided attention distillation for lane segmentation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.833979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.943975Z digest=sha256:852c077c6d89ca607a86676fc7368d0693ea0e02d490bee849904d620eec025f

Observation 008f223d-c2e5-40c2-b586-8e9b171ae2b3 · outbound

This paper cites Learning based multi-modality image and video compression.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning based multi-modality image and video compression

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.809597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.953962Z digest=sha256:f80ff674e71306edae58e80d3c7c1788fbbcc996c5b16a2152d6224e4d941148

Observation df999c87-0a0f-47b2-af23-0d45ca647b49 · outbound

This paper cites Ppt: token- pruned pose transformer for monocular and multi-view human pose estimation.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Ppt: token- pruned pose transformer for monocular and multi-view human pose estimation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.782259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.961135Z digest=sha256:0a9bd98324efbc6d523d9322d3503f4a95f1f6f0799299c79ede4ddb60315d95

Observation 9772c7e4-0af2-4a67-916a-7fee1d309a52 · outbound

This paper cites Re- current convolutional network for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Re- current convolutional network for video-based person re-identification

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.759689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.966551Z digest=sha256:8ae1309249f6763cd5572dd1c7099b961f57fd43a7ddaff4aa1f28670863989a

Observation 1d7d963a-c84f-4d36-a925-bf5f54620d50 · outbound

This paper cites Deep spectral methods: A surprisingly strong baseline for unsupervised semantic segmentation and localization.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Deep spectral methods: A surprisingly strong baseline for unsupervised semantic segmentation and localization

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.714484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.971980Z digest=sha256:b27c238a4b3485889b55bcdc2db000df7a51af596dfdd20b22ff0bb7806cec65

Observation 02aaef4e-ed0d-4ca7-9789-230b272472b6 · outbound

This paper cites Adavit: Adaptive vision transformers for efficient image recognition.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Adavit: Adaptive vision transformers for efficient image recognition

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.670264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.979085Z digest=sha256:3c84fc36b2393e2eeac60abd7eceba82be683b8dc255fb12e0328a6e06b6851c

Observation 047631c8-c271-44d9-913c-f170ded2615e · outbound

This paper cites Counterfac- tual attention learning for fine-grained visual categorization and re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Counterfac- tual attention learning for fine-grained visual categorization and re- identification

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.641634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.987111Z digest=sha256:3efba5b3315bf71a292036866d64138478841b9712bc4d69297b4caf9830d825

Observation 4aa80c06-eaaf-4974-81f7-cd986963a15e · outbound

This paper cites Dynamicvit: Efficient vision transformers with dynamic token sparsification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Dynamicvit: Efficient vision transformers with dynamic token sparsification

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.601500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:00.995065Z digest=sha256:366dcc21fadb4802d02d213cc969be455dd1ea1501942760c9d53bd5472e99b1

Observation 4b95de68-333b-43fd-a312-13f738dbcf05 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification SAM 2: Segment Anything in Images and Videos

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:01.004449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:01.004449Z digest=sha256:88f137af594f0115e48d145cade616780c8bece24c25e5388da94dae0c3ccea9

Observation 2e020797-d634-47ca-aded-044083e250ea · outbound

This paper cites Co- segmentation inspired attention networks for video-based person re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Co- segmentation inspired attention networks for video-based person re- identification

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.553221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.020476Z digest=sha256:444a39d955d3095974511ceada69ce07a2a77ac4e1b65e897a2549daef42c801

Observation 94818aaa-7c94-457d-bed4-b8360c2c15e8 · outbound

This paper cites Patch slimming for efficient vision transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Patch slimming for efficient vision transformers

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.522141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.029785Z digest=sha256:94237f8228643c9fb42994a6c7f00b414119c9531fa494e9be44b4b42489f730

Observation ebb60d0e-9153-4c83-b7f6-e9a0d53dc7c0 · outbound

This paper cites Multi-stage spatio-temporal aggregation transformer for video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Multi-stage spatio-temporal aggregation transformer for video person re-identification

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.481702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.045746Z digest=sha256:4cb28e3a13f8da736770bfff42b4412146c03c94196bcfc07b00630a8260d881

Observation 921d8ada-28a8-41f3-a01e-09366a84df14 · outbound

This paper cites Training data-efficient image transformers & distillation through attention.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Training data-efficient image transformers & distillation through attention

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.446338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.053739Z digest=sha256:69d99d02e4c0def4bb36918355c4dbf30e21a3aef00a3fd474236c8736903df6

Observation 4f2d0bb4-95a6-41d1-9d37-f8429ae60f87 · outbound

This paper cites Efficient video transformers with spatial-temporal token selection.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Efficient video transformers with spatial-temporal token selection

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.405956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.070129Z digest=sha256:5767d11435fd87e6195355dcf06d5ffa1bfbdaf8419745839490f0fd02482033

Observation a7e11d25-784c-45bc-90d0-8ae283a164d5 · outbound

This paper cites Pyramid spatial-temporal aggregation for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Pyramid spatial-temporal aggregation for video-based person re-identification

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.369877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.082826Z digest=sha256:0e8ee611ef92c00cef74e55e857cb54bcb6ce2979bb4224ad7510a0539ae0b2f

Observation 307c01be-5d1a-4f31-aeca-b5319d7704a0 · outbound

This paper cites Joint token pruning and squeezing towards more aggressive compression of vision transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Joint token pruning and squeezing towards more aggressive compression of vision transformers

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.332193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.095693Z digest=sha256:57174d495a322e4e6643d42a8ce8dbe1245955d990256b8630111e699d91073f

Observation 0aa13f07-9948-4b24-873c-767070e4e605 · outbound

This paper cites Overview of the h.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Overview of the h

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.299580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.115936Z digest=sha256:00d1dd23b8567067e1d30bef2e97c8779b5c4830f3f36a0a4fe9bb83029fa2e7

Observation 12292abd-d6f3-456e-a744-a2af9f2b9a37 · outbound

This paper cites Cavit: Contextual alignment vision transformer for video object re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Cavit: Contextual alignment vision transformer for video object re-identification

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.261838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.128704Z digest=sha256:a908bd9006d135a082140d61b6460b8dd2aebb2beac2eac6d475e150297d1979

Observation cf4bce51-30f0-43f1-af45-aab88f3b625f · outbound

This paper cites Tinyvit: Fast pretraining distillation for small vision transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Tinyvit: Fast pretraining distillation for small vision transformers

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.236096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.140111Z digest=sha256:882a7cb63e06b2709f3c7ff4b0dbd4147c8d53f950285a9b6561856f1dc866e6

Observation 8a4f1536-7211-4faa-b76d-70eb0606caf7 · outbound

This paper cites Learning resolution- adaptive representations for cross-resolution person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning resolution- adaptive representations for cross-resolution person re-identification

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.186474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.149464Z digest=sha256:92277d8401f653108d0209a7a56fc0d500069bb4fd2eb79ba2204d6e63170677

Observation 3db7aa8e-8600-4560-85b4-a18c9a2ae359 · outbound

This paper cites Temporal complementarity-guided reinforcement learning for image- to-video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Temporal complementarity-guided reinforcement learning for image- to-video person re-identification

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.140192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.155537Z digest=sha256:7c20321830004b16b07e5531e4a5feb04626ff1546843b88604601362ccd509f

Observation 31e0a4cd-b324-453f-81e6-ab3037f547d7 · outbound

This paper cites Adaptive graph representation learning for video person re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Adaptive graph representation learning for video person re- identification

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.110319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.171234Z digest=sha256:cf553abd243ae2b9777562733fb1e157a3d4b2cffc24f2984eab77a5e4eca285

Observation 894f8484-1cd6-4cb9-bf6c-b0f81ca7ca5f · outbound

This paper cites Segformer: Simple and efficient design for semantic segmentation with transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Segformer: Simple and efficient design for semantic segmentation with transformers

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.084236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.180921Z digest=sha256:86985bbfac8c131967da44c6c2baf5bea550671fdcc634a3aab6a2b6332a09b9

Observation e897bd96-5ac3-438c-9781-cdc94afa0989 · outbound

This paper cites Learning multi-granular hypergraphs for video-based person re- identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning multi-granular hypergraphs for video-based person re- identification

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.049824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.191992Z digest=sha256:53ffce66088c306ab2b682d17ffbb9c6905de2fba6581d54a0dd7c77873e6f61

Observation aa82aabd-5183-453a-ac1d-84549556343e · outbound

This paper cites Spatial-temporal graph convolutional network for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Spatial-temporal graph convolutional network for video-based person re-identification

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:03.023721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.223155Z digest=sha256:1db8ff567def76cc13e0563fb5e42c44b25e4e973d58dae6ddef3ed4dd2b5865

Observation d1300d57-c14c-4878-b45b-e29a0df9bdcb · outbound

This paper cites Stfe: A comprehensive video-based person re-identification network based on spatio-temporal feature enhancement.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Stfe: A comprehensive video-based person re-identification network based on spatio-temporal feature enhancement

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.990324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.239166Z digest=sha256:c8e87fceb26dd6a59410b7e054157d0442ccfeedd9581acb8cb38006694a6135

Observation a8e9a7b2-4c4b-41b0-9b1c-cd466559acd9 · outbound

This paper cites Shiftaddvit: Mixture of multiplication primitives towards efficient vision trans- former.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Shiftaddvit: Mixture of multiplication primitives towards efficient vision trans- former

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.950869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.253538Z digest=sha256:00cb5e697048d3ae6a14e6fabd1c3ba51b8c5491843c7b6f9155a3e033d5b253

Observation d64b79cf-b33f-4b05-9b3b-d6d84fbf11f1 · outbound

This paper cites an unresolved cited work.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-10T10:34:02.912112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.273595Z digest=sha256:05114773574334cebcf0f63d704d78a91539b96e7d8e44b7f281e43f5a8fa699

Observation 18378f3d-7930-4126-9b1c-acfc7e872773 · outbound

This paper cites Tf-clip: Learning text-free clip for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Tf-clip: Learning text-free clip for video-based person re-identification

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.868769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.289444Z digest=sha256:0049ef5ddd2efc9484003901fa22c0b3eb3f783f228898a707dfeca63e8ad74e

Observation 98e448b3-cf8c-443d-9d51-808a34e17e79 · outbound

This paper cites Metaformer is actually what you need for vision.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Metaformer is actually what you need for vision

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.829017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.303231Z digest=sha256:4c8e001faf82c32231ada77ecc51696a2e12cc059cb8aa86eea4e4e38fd42ad4

Observation d9437572-3996-44cf-9677-5dc8e707cb0d · outbound

This paper cites Ptq4vit: Post-training quantization for vision transformers with twin uniform quantization.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Ptq4vit: Post-training quantization for vision transformers with twin uniform quantization

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.805325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.320177Z digest=sha256:e145a4be0e7bf445110f5b926a621accc65157fc315881062d52ecea99b0db10

Observation 903b4ee1-7d31-409a-85e6-e0761d354f86 · outbound

This paper cites Resmatch: Referring expression segmentation in a semi-supervised manner.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Resmatch: Referring expression segmentation in a semi-supervised manner

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.766980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.328611Z digest=sha256:2e059b4c9e4ebb30dc8979d3e2a3f83940580afc6a6c9b2d442d8d15a1d6326b

Observation c6465497-0326-48a4-9472-6a3c67ae4bd7 · outbound

This paper cites Minivit: Compressing vision transformers with weight multiplexing.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Minivit: Compressing vision transformers with weight multiplexing

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.721102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.337366Z digest=sha256:56a9f07939f92702cd7a6dc7ddc76d9b399f25ce4cdb260e10d0461c7dc8c79f

Observation cf8422f9-5cbc-401c-8289-d1aae970818c · outbound

This paper cites Magic tokens: Select diverse tokens for multi-modal object re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Magic tokens: Select diverse tokens for multi-modal object re-identification

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.494523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.344601Z digest=sha256:7302c8c36dca85ee70e3956993e71cc4f83ab7f712687da1377d0575671f6192

Observation 06b90b88-2f2d-4c88-a479-36a41fc82c4d · outbound

This paper cites Learning bidirectional temporal cues for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning bidirectional temporal cues for video-based person re-identification

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.469527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.350272Z digest=sha256:7ce646b0ff48e35e501340ac652f1eaec67ceb0a64ebddc663a2eae3ac08ee4a

Observation 8bbb94ba-6850-468c-9bb4-4e9b1cef9310 · outbound

This paper cites Multi- granularity reference-aided attentive feature aggregation for video- based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Multi- granularity reference-aided attentive feature aggregation for video- based person re-identification

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.444522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.358747Z digest=sha256:7212e0139f6d7d87f3617b861adb205d13b33e4be95704fdba5e76dce37bd5e2

Observation 0c313ad0-a03e-4e60-b764-b4dc5a424a5b · outbound

This paper cites Structure- aware cross-modal transformer for depth completion.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Structure- aware cross-modal transformer for depth completion

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.413916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.367458Z digest=sha256:904f85cd033dc8d99580d3e96d136be2a528894c31666224e284f56b85827107

Observation 9b79195b-35f0-48b5-bf5f-b905e06fcb14 · outbound

This paper cites Attribute-driven feature disentangling and temporal aggregation for video person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Attribute-driven feature disentangling and temporal aggregation for video person re-identification

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.385764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.374584Z digest=sha256:39f01398261a3c3ea859d078c0d4e471da92d7c2b0a8544f544de12fb03cfc4e

Observation d2aa7ecb-ecf0-4824-a9b9-ebbd4bf1d86f · outbound

This paper cites 3d human pose estimation with spatial and temporal transformers.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification 3d human pose estimation with spatial and temporal transformers

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.351170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.383392Z digest=sha256:7c4c569ab32a8642f536a2b8e5617183b3c7deb48cbc07368c511fcbbfef36d3

Observation 48e554cd-45a1-4044-9b37-dad578f6e048 · outbound

This paper cites Person Re-identification: Past, Present and Future.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Person Re-identification: Past, Present and Future

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:01.388890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:01.388890Z digest=sha256:c0ecb30384cb1ed2783f68fd627271b90622949f816d9fbeb04572aee4fcd8f3

Observation 95d447e4-cb39-44ef-afde-87efb668f158 · outbound

This paper cites Joint discriminative and generative learning for person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Joint discriminative and generative learning for person re-identification

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.316212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.420122Z digest=sha256:4a9a6a72b94cd5afaa9abc02c676d2a81fdef565e299d0eece4f4cda0d867283

Observation eda5782f-e048-46b1-94a8-a659335c6634 · outbound

This paper cites Omni-scale feature learning for person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Omni-scale feature learning for person re-identification

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.274860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.429972Z digest=sha256:a05a5565bd08507f2f72067de6e499e32114c6b170f983ca966655b46a635c60

Observation c593dd8b-0ae2-4291-a6e3-eb601c39fd4b · outbound

This paper cites See the forest for the trees: Joint spatial and temporal recurrent neural networks for video-based person re-identification.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification See the forest for the trees: Joint spatial and temporal recurrent neural networks for video-based person re-identification

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.248839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.440553Z digest=sha256:d9a0bef2e71b50e51c20b370c1a973bebabbe262a6f04b531a419af31df57323

Observation 97d4ea06-622f-4d46-99b0-c209172392c3 · outbound

This paper cites Llafs: When large language models meet few-shot segmentation.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Llafs: When large language models meet few-shot segmentation

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.207072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.449686Z digest=sha256:3f77d1bace759befc27f627a98c5273f48c66e21a7dc588291f01b681f8e737e

Observation 84fc00dd-55fa-4b52-9401-0faccc751e8b · outbound

This paper cites Continual semantic segmentation with automatic memory sample selection.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Continual semantic segmentation with automatic memory sample selection

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.175854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.457857Z digest=sha256:05bd9fe3a1e297a0dbae5d16e2d87dc33eb9e7a846f6f3be5ed5252ce3714b94

Observation 9f95fbcf-1407-4e71-bec4-57dbd7de63f7 · outbound

This paper cites Learning gabor texture features for fine-grained recognition.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning gabor texture features for fine-grained recognition

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.146109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.469095Z digest=sha256:2282db16c3e0e3f242b8a4801ee53621425e2b51d4d40cc545cf5ff800a3075f

Observation cc7719c1-30c2-4586-8b78-e6b0371cee33 · outbound

This paper cites Addressing background context bias in few-shot segmentation through iterative modulation.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Addressing background context bias in few-shot segmentation through iterative modulation

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.117096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.484070Z digest=sha256:60b3a1abc2091d2331a264eb4a843baf2239061e218568c3f6efe76ddf23fd26

Observation 25b2325f-84a0-44fc-a614-f97f6590f3d8 · outbound

This paper cites IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T10:34:01.502054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:34:01.502054Z digest=sha256:b0efb54f665b3ca763476b22821eb3c5b206b9aa9dd82a3fa40c64a6ad1cd24c

Observation 0f656cd4-37b5-44fa-884d-ddfe7b1a491b · outbound

This paper cites Learning statistical texture for semantic segmentation.

Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification Learning statistical texture for semantic segmentation

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T10:34:02.082473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T10:34:01.517256Z digest=sha256:7b79687b8107378d011e0df0856f54729116b5b961fc7e4f3a29f5a2ef72d373

Pith citing papers

Observation 3b8a5e8e-1997-4d87-8d47-ed4499b58348 · inbound

HD-VGGT: High-Resolution Visual Geometry Transformer cites this paper.

HD-VGGT: High-Resolution Visual Geometry Transformer Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:09:30.695324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-14T22:09:28.433165Z digest=sha256:93c0e29955174b2a7e45691b7b171030a6e9828dbeda3fb711592f3e8e2fc524