Pith. sign in

Paper Citation Record · LEDGER

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding

As of 18 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 0 inbound Pith citation observations for arXiv:2608.08832.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.08832 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:28:17.123060Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

72 of 72 outbound references displayed

  • verified exact7
  • verified fuzzy44
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 21627b4f-3927-41b5-9214-74ad654f6b7b · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.834104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.834104Z digest=sha256:980e56971d5eba46ffcfffc1ecf07171039c0534084d18599b67bc8f320a4f35

Observation 9dbc080d-cd63-4e17-8f0d-05eb4ce0b1a0 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding DINOv2: Learning Robust Visual Features without Supervision

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.839108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.839108Z digest=sha256:3c33ac0b1ecf7727222b22daeff6a43b39a72d105bfa8275c53818186edade43

Observation 89888f0a-8424-4215-86b8-9844169e6c63 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.843829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.843829Z digest=sha256:f24a188880373925ace14aa94647f39040b09524dc8d0fa2bff423c10afd0be1

Observation 7cbda7ac-2907-4e37-8bc4-7c711c03fa47 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding SAM 2: Segment Anything in Images and Videos

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.848229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.848229Z digest=sha256:ea13d9978fb4b53c578718e1a6011565549bf66852a3295ecec302ebc550e9bf

Observation b91af1f3-ef7a-45e0-b3f4-bfb6217fae44 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding SAM 3: Segment Anything with Concepts

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.852686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.852686Z digest=sha256:217a7d51cf275eca8e807ae3e260b9a1ee5c92ecc45d6984cfb3da401042458f

Observation aa859a75-8b50-443e-89e8-8f6f93efed3d · outbound

This paper cites OpenFedLLM: Training Large Language Models on Decentralized Private Data via Federated Learning,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding OpenFedLLM: Training Large Language Models on Decentralized Private Data via Federated Learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.352097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.857124Z digest=sha256:374384b7584d3620b2e285c2c4814a20ef22c3cae1b63b58695bc9d81dbbd2c7

Observation 939425c4-6526-47c3-9d9f-523d5baae060 · outbound

This paper cites Feature Coding in the Era of Large Models: Dataset, Test Conditions, and Benchmark.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Feature Coding in the Era of Large Models: Dataset, Test Conditions, and Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.862213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.862213Z digest=sha256:5e4fa63e794cad5dac39c0fe68e64a1f7013600bef21c29eaf58a67136baa5c9

Observation d902ed8c-4478-4416-9c8d-c8964fa2d8ec · outbound

This paper cites Vision transformers need registers,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Vision transformers need registers,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.328800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.866841Z digest=sha256:a910ed61d23dcfc1a5f6bd8101e7c2af1690444772ff8dfb1d01b6bc103b1725

Observation e4a93c31-862c-4a68-a664-c4400ce27532 · outbound

This paper cites DT-UFC: Universal Large Model Feature Coding via Peaky-to-Balanced Distribution Trans- formation,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding DT-UFC: Universal Large Model Feature Coding via Peaky-to-Balanced Distribution Trans- formation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.314995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.870949Z digest=sha256:25d63b4a6a5dab4823f3eb688187bf1e6395ec047cf6d4d270d52e21493e58b7

Observation 6ebd86b1-fdfd-4769-a3b5-aa607fcee88d · outbound

This paper cites Transform-Free Feature Coding via Entropy-Constrained Vector Quantization,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Transform-Free Feature Coding via Entropy-Constrained Vector Quantization,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.302157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.874952Z digest=sha256:67bfee24b406d8500d7d80963ebae3d12d003ed144c5c862149e57922df2814e

Observation c4811298-0eec-42c5-9751-dbf07471513b · outbound

This paper cites VVC and VTM 11.0: The Versatile Video Coding standard and its reference software,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding VVC and VTM 11.0: The Versatile Video Coding standard and its reference software,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.287727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.878931Z digest=sha256:41ddb40a5d210a2be43c2d02d2d5adb056cba7a6ac2357cf9f3982a36fb01aab

Observation 90b4033d-06e7-4926-bf3d-648bf7b21ebf · outbound

This paper cites Variational image compression with a scale hyperprior,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Variational image compression with a scale hyperprior,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.883230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.883230Z digest=sha256:0cac493d8a4f8e456cb85e6490fab3769195a01e8aaf034eab2412e1d4ad8a6c

Observation 56d989bc-5cac-4e74-854d-bf991784ef43 · outbound

This paper cites ELIC: Efficient Learned Image Compression with Unevenly Grouped Space- Channel Contextual Adaptive Coding,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding ELIC: Efficient Learned Image Compression with Unevenly Grouped Space- Channel Contextual Adaptive Coding,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.266177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.887163Z digest=sha256:3ab753905d5593121365b2f93a75770af6b8132c387baf074a0d88595ffb183c

Observation 6d009baf-32a1-491d-b96d-f2f3fd0f1df1 · outbound

This paper cites Nonlinear Transform Coding,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Nonlinear Transform Coding,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.254409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.891335Z digest=sha256:8864f034163674297877f53c037459637e4a0e4d156f2f93c55a98a33a8c5111

Observation 551287e9-e6c8-47f8-a31a-3f154ff577c4 · outbound

This paper cites End-to-end optimized image compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding End-to-end optimized image compression,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.237069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.895604Z digest=sha256:4522098048769534d734b815ffc61be58deede21ae1b510df980349f377aa1fd

Observation f4150ee7-a3d6-4679-80cb-e6253686f51b · outbound

This paper cites Joint Autoregressive and Hierar- chical Priors for Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Joint Autoregressive and Hierar- chical Priors for Learned Image Compression,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.224770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.899489Z digest=sha256:db75d68acb1a2fff398158cd6d39c3715b7f6229426fbd96b8e5905737762839

Observation 6e2043f8-194c-4a35-b34e-92f90e99a23b · outbound

This paper cites Learned Image Com- pression With Discretized Gaussian Mixture Likelihoods and Attention Modules,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Learned Image Com- pression With Discretized Gaussian Mixture Likelihoods and Attention Modules,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.212189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.903731Z digest=sha256:7c36f4ab6cc1396dc885bd253fb28045cd26a16fceeae1cffd9f489d7b5954a7

Observation 456db18d-a295-45c7-80f1-dc53c73113bc · outbound

This paper cites Enhanced Invertible Encoding for Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Enhanced Invertible Encoding for Learned Image Compression,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.907809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.907809Z digest=sha256:11360586db60db8de8df61d188dba2a2be01fe563428f6b1b46a5ddcf5a5e550

Observation 9dda892b-b615-4de1-8476-6a5307414997 · outbound

This paper cites an unresolved cited work.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:28:18.199838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.911683Z digest=sha256:2ada4b7e08be1d78d553c58ab93463ae4d1da2aa86c9c48946618e9363fd5c22

Observation e869c0b8-a275-4d0d-b5dd-15749c24edbc · outbound

This paper cites End-to-End Optimized Versatile Image Compression With Wavelet-Like Transform,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding End-to-End Optimized Versatile Image Compression With Wavelet-Like Transform,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.187529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.916050Z digest=sha256:2a2c46fe3212eaab298b08ac3a5c15c5e1dd3cf23533a69545c51424cfccb8db

Observation c60bccf0-1881-4ed5-b89d-d61830f30ba9 · outbound

This paper cites Transformer-based Transform Coding,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Transformer-based Transform Coding,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.170835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.920195Z digest=sha256:e039df6f4b8bddd3433f2a436ac729efec13b40ece263ffd2e531cf7f2de98fe

Observation d240ac1c-c1a9-49ec-be1f-e45302cf02aa · outbound

This paper cites Learned Image Compression with Mixed Transformer-CNN Architectures,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Learned Image Compression with Mixed Transformer-CNN Architectures,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.150887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.924231Z digest=sha256:8b2caa9c2f4fe073e153d7abe283e0484f87c6fff286a2039732600090f11a33

Observation 4e3c0b76-88cd-4ce3-b1c1-1b8e05493a37 · outbound

This paper cites Frequency-Aware Transformer for Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Frequency-Aware Transformer for Learned Image Compression,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.138152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.928006Z digest=sha256:6afbc2bd1f89be1cdc73a596b480ce93fbbdad78681a998d136def80b0e668bd

Observation cbfd6ee5-ab6b-4a81-9bcf-29d7f32431eb · outbound

This paper cites an unresolved cited work.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:28:18.122481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.931832Z digest=sha256:81b684875addb5655b3c4ab9b8abd0e70f9679b36e252a8db062f36ccf0ee6f0

Observation 97fa5ad3-31b8-4297-8200-bda2683c4a9b · outbound

This paper cites an unresolved cited work.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:28:18.104671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.935651Z digest=sha256:cbee406cdf1be95b8da327cb49c993ab67fdf55379e377baa51bc351774a151a

Observation e308bd92-d009-4014-9769-7332b9d82d69 · outbound

This paper cites MambaIC: State Space Models for High-Performance Learned Image Compres- sion,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding MambaIC: State Space Models for High-Performance Learned Image Compres- sion,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.090519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.939716Z digest=sha256:4e0fadadffaea936709fb89a424398ecec0cfa7c7c6dcf7d3206fc962850a987

Observation f0b7900a-c0e1-4a6c-b8f1-3b4b4f7ba270 · outbound

This paper cites an unresolved cited work.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:28:18.077627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.943481Z digest=sha256:4150c35e5db911f7eea541ed3a5aad3258e68cafe07cdeffb634d310bc53e372

Observation cb7e82b7-b74e-4608-a24b-e1f8a356b96e · outbound

This paper cites an unresolved cited work.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:28:18.063841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.947520Z digest=sha256:21dd9fcd84c43cd04eca9cc162c5ed54706e52358b25c457e609a428e23271a8

Observation abb60aa7-69d1-4bd6-a082-ded678c992ff · outbound

This paper cites Linear Attention Modeling for Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Linear Attention Modeling for Learned Image Compression,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.049871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.951629Z digest=sha256:d5dea7eeef89b36581b7b4f142ba22be9d7f9695a3abb8a30623af068a4bc891

Observation beb0e65f-ad60-43b1-b943-503aa9d72ed8 · outbound

This paper cites Distilling complexity-scalable learned image compression models via neural architecture search,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Distilling complexity-scalable learned image compression models via neural architecture search,

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:28:17.451049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.955365Z digest=sha256:fefe7af69a8c7d4fa37bfdc8799b69971c321c8d2eb1aa48d61fc30198b39c50

Observation 703f94dc-5a50-40cc-9e63-345a59541011 · outbound

This paper cites EVC: Towards Real-Time Neural Image Compression with Mask Decay,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding EVC: Towards Real-Time Neural Image Compression with Mask Decay,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.033085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.959515Z digest=sha256:7ddc77177aa15f552d66d7d0e8814cddaa832566af48c4f7167e9c995c25fd95

Observation 51ffdcd3-a94d-4059-8928-dd9a3e364a24 · outbound

This paper cites Towards practical real-time neural video compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Towards practical real-time neural video compression,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:18.016688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.964275Z digest=sha256:a8fff108db3c2413d0abdeff0b61c7accee38fb4b8876e21d680fbbbf156ad0f

Observation 4dfe1035-d4a9-4311-8acd-d8751958e51d · outbound

This paper cites Checkerboard context model for efficient learned image compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Checkerboard context model for efficient learned image compression,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:16.968303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:16.968303Z digest=sha256:8b5b76f9dba8403cdd680c923c7df743c07f7d9b9304d7be08d2b6e20b11be1f

Observation 2e245178-9ece-4ad5-953f-3a60d2ac94f3 · outbound

This paper cites Channel-Wise Autoregressive Entropy Models for Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Channel-Wise Autoregressive Entropy Models for Learned Image Compression,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.989513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.972243Z digest=sha256:df659b01687e80b1139b1e4d535925757efbe2051e3977182a674902d79289a6

Observation 8d877eca-af07-4632-be6c-2eb59d3a39a2 · outbound

This paper cites MLIC: Multi-Reference Entropy Model for Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding MLIC: Multi-Reference Entropy Model for Learned Image Compression,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.975265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.976154Z digest=sha256:7f622826d31be339c242db85bdac32eaaa32b0a259df1e457378abaec46b3ae4

Observation 6ab54182-7915-4327-aa61-fc77241f2ca6 · outbound

This paper cites MLIC++: Linear Complexity Multi-Reference Entropy Modeling for Learned Image Compression.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding MLIC++: Linear Complexity Multi-Reference Entropy Modeling for Learned Image Compression

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.960813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.980713Z digest=sha256:bc16e0fe48708d9193c5fdcabcca19bc403576ad5da621514699068623597d48

Observation e295b671-29e3-4846-88ac-bb8df6767450 · outbound

This paper cites an unresolved cited work.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:28:17.941544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.985036Z digest=sha256:75509a39b6ae573df8b2d013c2bc69cfadc2f46feb64c30c4f7a967c96cf9a38

Observation 3842660c-f2db-41fb-b3be-9d98d8b06440 · outbound

This paper cites Contextformer: A Transformer with Spatio-Channel Attention for Context Modeling in Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Contextformer: A Transformer with Spatio-Channel Attention for Context Modeling in Learned Image Compression,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.926371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.989128Z digest=sha256:986087c8e1c6fd573e34aa4b7be3b6037d65600f15b562383247a99804017f44

Observation 02ddbd9f-6784-42b1-b113-17551b6b6cad · outbound

This paper cites Efficient Contextformer: Spatio-Channel Window Attention for Fast Context Modeling in Learned Image Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Efficient Contextformer: Spatio-Channel Window Attention for Fast Context Modeling in Learned Image Compression,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.911307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.993445Z digest=sha256:8a0da32e388ea91845556eb4a0813fd58a4d8326bb50f05f7122fbd0f04906a1

Observation a55f8753-c8e9-4756-8121-2509bd40cdc3 · outbound

This paper cites Learning Context-Based Non-local Entropy Modeling for Image Compression.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Learning Context-Based Non-local Entropy Modeling for Image Compression

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:28:17.390663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:16.997430Z digest=sha256:05bcf13f1d0cfc15e5053e6498a241de7c85ab2e1b19e6dd5b2aa308ca07f11e

Observation d9a32c24-c0c0-422b-87a6-bdd03a6b9e29 · outbound

This paper cites Video coding for machines: A paradigm of collaborative compression and intelligent analytics,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Video coding for machines: A paradigm of collaborative compression and intelligent analytics,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:17.002209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:17.002209Z digest=sha256:b0b411597601b8c1d85e647f47822ebed86ff7549ea8afeb04287eb27be0cbb9

Observation bb0e904f-477a-4685-84e2-14abc38ad38e · outbound

This paper cites Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:17.006371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:17.006371Z digest=sha256:7e76f230d53db18ff023a52f5d2b9f8c4127f9d15c670348b4cbf7b66b655019

Observation 860a3d97-6f2c-4c43-8a08-976e8120f90f · outbound

This paper cites Human-Machine Collaborative Image and Video Compression: A Survey,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Human-Machine Collaborative Image and Video Compression: A Survey,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.874885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.010402Z digest=sha256:9e18fa07f53ac5cd3e259658ed5af6edb2eed29af58629360339723b33a870a2

Observation 935b4398-0a34-4246-a43b-e38bcf9c6e26 · outbound

This paper cites Call for proposals on feature compression for video coding for machines,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Call for proposals on feature compression for video coding for machines,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.861515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.015042Z digest=sha256:d3ffd57bd230e526fc3f2e4dedce9b3d1c959b8de635ab65815b3250ebea03c4

Observation 943bfbc4-d6ff-41f5-9379-0a149aea57ef · outbound

This paper cites Toward Intelligent Sensing: Intermediate Deep Feature Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Toward Intelligent Sensing: Intermediate Deep Feature Compression,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.847811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.019123Z digest=sha256:52d65dc4dbc898c05794f3a29dab61253931a181eda5c97968a40fb161cf222e

Observation 768d4133-f79e-4ea3-ad56-31a9cc4bdd16 · outbound

This paper cites AlphaVC: High-Performance and Efficient Learned Video Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding AlphaVC: High-Performance and Efficient Learned Video Compression,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.833549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.023105Z digest=sha256:87b92c6d05f4a6be6e1c232e722fc8188b617e84a3526133a28ae9c3221c0a2f

Observation 9458dc65-8a80-4c7c-aea1-ddfc5243a664 · outbound

This paper cites MIMT: Masked Image Modeling Transformer for Video Compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding MIMT: Masked Image Modeling Transformer for Video Compression,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.815715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.027284Z digest=sha256:ad87e9fdf7c29c52c0a8b40aae5b471cade5c626b905ebe3ca45a43ca23fef40

Observation 08b8f5bb-8617-418b-ae51-2988cb3a2ac9 · outbound

This paper cites FLA VC: Learned Video Compression with Feature Level Attention,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding FLA VC: Learned Video Compression with Feature Level Attention,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.800885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.031467Z digest=sha256:a808b869148b2b0c2fd2111e60a2fe59bde87a848d2f06aebe142dba9a68780a

Observation 7627733e-df5e-4365-a91a-8fcdbcbd0004 · outbound

This paper cites End-to-End Learnable Multi-Scale Feature Compression for VCM.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding End-to-End Learnable Multi-Scale Feature Compression for VCM

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:28:17.374236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.035641Z digest=sha256:23e4e8cb61f910a5443fb645092af9fa6dc110c05ef2258212cd02dcefb51dfb

Observation 3e1728ab-a82c-4327-ad98-ffbd80e69da0 · outbound

This paper cites Learnt Mutual Feature Compression for Machine Vision,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Learnt Mutual Feature Compression for Machine Vision,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:17.039504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:17.039504Z digest=sha256:ca6b73699a8f7934e44b8d9ad881f343fd58eed10baf5a2cdaa851c4a607438d

Observation d84eda67-0909-4c3a-8ff2-c0436e8849c7 · outbound

This paper cites Sensitivity-aware bit allocation for intermediate deep feature compression,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Sensitivity-aware bit allocation for intermediate deep feature compression,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.785887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.043322Z digest=sha256:e69c347e11f3126fc72d3fe49e15ef4ff3af34546db57f489383fe780073dc2d

Observation 7117eb8e-228b-404e-9fdb-6728a4bb43f6 · outbound

This paper cites Feature compression with 3d sparse convolution,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Feature compression with 3d sparse convolution,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.774974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.047282Z digest=sha256:6c1fffa7e9641a02d20cbdd6dc339035fda9e77045293d9f8601107b5d66c99b

Observation bae0d763-9a58-44fd-a042-ae828d7b0cc5 · outbound

This paper cites Latent-space scalability for multi-task collaborative intelligence.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Latent-space scalability for multi-task collaborative intelligence

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:28:17.291194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.051052Z digest=sha256:f5412b0e79a50653d05c068f2fca4521543959447a06233d4e70604b65aba890

Observation 1dda6253-80be-4d95-bbc3-b9233421539d · outbound

This paper cites DMOFC: Discrimination Metric-Optimized Feature Compression.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding DMOFC: Discrimination Metric-Optimized Feature Compression

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:28:17.274287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.055164Z digest=sha256:544640cd5bf6a3472c7d808a8d4d1525057740ca0fc1e78f141cfe6b7a1ed056

Observation cfc761cf-65fe-4dab-adc1-2a885fcf2ccf · outbound

This paper cites IMOFC: Identity-Level Metric Optimized Feature Compression for Identification Tasks,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding IMOFC: Identity-Level Metric Optimized Feature Compression for Identification Tasks,

Reference 55

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:28:17.255720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.059144Z digest=sha256:c3c61815ae24acd8bed17931fb5a227ff5aa062f14cbfe30466e6f13ee880abb

Observation 0b83ca89-6881-431a-ba99-78f5ece06793 · outbound

This paper cites Rethinking joint optimization in feature compression: Insights from person re- identification,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Rethinking joint optimization in feature compression: Insights from person re- identification,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.757758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.062696Z digest=sha256:78475dfeb55e9ce02e097bb70e0a4aa7f3264d3380904e0b7cb89f583e78f0af

Observation f0e1aded-7d99-4157-ac5b-68e7ae79db57 · outbound

This paper cites Compressed feature quality assessment: Dataset and baselines,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Compressed feature quality assessment: Dataset and baselines,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.745222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.066411Z digest=sha256:7c638bf116795609831a7152868c460afb672113afd1aff63c718b119fd8c940

Observation 499f9489-4402-4934-8642-95c47953b211 · outbound

This paper cites Scalable Facial Image Compression with Deep Feature Reconstruction,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Scalable Facial Image Compression with Deep Feature Reconstruction,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.732926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.070170Z digest=sha256:80db676f2dd81425edd282f537ef0979e8a6fc82fab2b1ba000b4315af9e574e

Observation 4912f3fd-e53b-46bf-ac1d-59b27c09c5a0 · outbound

This paper cites Towards coding for human and machine vision: Scalable face image coding,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Towards coding for human and machine vision: Scalable face image coding,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.719949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.073796Z digest=sha256:c66b063d3d2aa1426df02bb7509e6f6ba5a0e6e35da2fc228e71e38d5eb38b5a

Observation 0849ef8a-fd75-4628-b71b-a318dd48d995 · outbound

This paper cites SSSIC: Semantics-to- Signal Scalable Image Coding With Learned Structural Representations,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding SSSIC: Semantics-to- Signal Scalable Image Coding With Learned Structural Representations,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.706713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.078041Z digest=sha256:640a45fc1b6731d6dd779a5efe7e4ea985979c39b2bca63c5097ddd3261efe88

Observation bdce8b0d-d284-4cc1-9e04-d16fb9f6a895 · outbound

This paper cites Semantics-to-Signal Scalable Image Compression with Learned Revertible Representations,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Semantics-to-Signal Scalable Image Compression with Learned Revertible Representations,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.693082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.081309Z digest=sha256:25041ee32d68c53e3318588a7b36ce660ca9b3e287d4bf65563fd603e8b5100c

Observation 44e616b9-1565-43d2-90eb-80ee70860ec8 · outbound

This paper cites End-to-End Learned Scalable Multilayer Feature Compression for Machine Vision Tasks,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding End-to-End Learned Scalable Multilayer Feature Compression for Machine Vision Tasks,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.679952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.085658Z digest=sha256:f46f18d70578a8a91d9a15a98527ba8cc841fd9fffcaf18db2a3264059af9dbd

Observation 070edd12-6c2c-4d65-9784-437d3aae9105 · outbound

This paper cites Towards Large Model Feature Coding.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Towards Large Model Feature Coding

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:28:17.173662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.089597Z digest=sha256:6387ecb8dbbe16e308e45bffca200fb9b6982c48f22556c77ad1852eedbb6fa1

Observation c0b33474-a197-4a1d-99c0-188e1984ff0c · outbound

This paper cites Compression of Self-Supervised Representa- tions for Machine Vision,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Compression of Self-Supervised Representa- tions for Machine Vision,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.666988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.093678Z digest=sha256:fefda069bb8553a0d290f64aaec7eeb04a9f3aa54c67e7e464469b7a8b5afc98

Observation 55b220b0-124d-4db5-bd54-ef0a450c3f62 · outbound

This paper cites Cross-architecture universal feature coding via distribution alignment,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Cross-architecture universal feature coding via distribution alignment,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.653727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.097363Z digest=sha256:fd160c6f76f62fe47ecb8e071ab4579d458599b216cd5a9b12ba82dc3bd80617

Observation ac260cbf-b0bd-4b1e-a6aa-26e7caa4bd3c · outbound

This paper cites Multirate Neural Image Compression with Adaptive Lattice Vector Quantization,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Multirate Neural Image Compression with Adaptive Lattice Vector Quantization,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.640353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.100994Z digest=sha256:2fac7dbfda3d0c77a0702de0a07e2b14412ec3e3532b970b1dbf4208c588cda6

Observation c036d20e-7b4f-48bb-8232-fa377de73869 · outbound

This paper cites Zhang, Y.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Zhang, Y

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.621704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.104725Z digest=sha256:4f1bbf892ede16387e4767a057db6c1bb21d6d1f3c6d95321a4f06b1ce2d3faa

Observation e0d5622e-ccce-4f18-ad62-7b1a3dec94fc · outbound

This paper cites ImageNet large scale visual recognition challenge,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding ImageNet large scale visual recognition challenge,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:17.108778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:17.108778Z digest=sha256:adf52398e7a040f1fd373b6115d0b96f368b51499593071d2894ad7aa2297f5a

Observation 236bced9-8652-4f10-a1b3-e79b125d3496 · outbound

This paper cites Microsoft COCO: Common objects in context,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Microsoft COCO: Common objects in context,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.602260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.112341Z digest=sha256:320ee6d4efc7148cabadf3c44966e5018bd8205fba8f3310bcae63e32fad4484

Observation 6d007847-6d5f-4793-a138-42f34cd52e5e · outbound

This paper cites Scene parsing through ADE20K dataset,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Scene parsing through ADE20K dataset,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:28:17.590339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-14T04:28:17.115755Z digest=sha256:9cb327b1e0c8fd57147ea201ba3c8c7b6be57c9a62934e62e0c2aee2e0055360

Observation 01648681-8daf-415b-9864-c927c8aaa655 · outbound

This paper cites Calculation of average psnr differences between rd- curves,.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Calculation of average psnr differences between rd- curves,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:17.119375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:17.119375Z digest=sha256:69b65ae08d56f635b7748724ee56290f98f61b3b5eb1ffc4dc98cf0ec6d84dcd

Observation 12e70c59-373e-4b12-974e-2b2c7ad7515d · outbound

This paper cites Diffusion Transformers with Representation Autoencoders.

Visual Token Codec: Unleashing Spatial Redundancy for ViT Feature Coding Diffusion Transformers with Representation Autoencoders

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-14T04:28:17.123060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:28:17.123060Z digest=sha256:a1bd548fa665cf1122a8a6ede80be116844c941f144df03a48af479fcdbca6c4

Pith citing papers

No inbound Pith citation observations are available.