Pith. sign in

Paper Citation Record · LEDGER

GViT: Representing Images as Gaussians for Visual Recognition

As of 18 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2506.23532.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.23532 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:47:13.856071Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact3
  • verified fuzzy28
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9945d05e-7020-4d1e-9ab0-bd728522523a · outbound

This paper cites Slic superpixels compared to state-of-the-art superpixel methods.

GViT: Representing Images as Gaussians for Visual Recognition Slic superpixels compared to state-of-the-art superpixel methods

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.892758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.014039Z digest=sha256:930660f0eaca95ab7a04284702d4e702251ced116d294ed071a7e43d9a636cd8

Observation 2c890d52-e762-428e-9657-dad3cf888b27 · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

GViT: Representing Images as Gaussians for Visual Recognition BEiT: BERT Pre-Training of Image Transformers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.069323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.069323Z digest=sha256:7673103357a485eb2bd1a3284948c0f11d2711d0b0fe37d85bb075c88a47d571

Observation 0092c7e7-3d0d-4b55-b20f-f251c06928f2 · outbound

This paper cites an unresolved cited work.

GViT: Representing Images as Gaussians for Visual Recognition Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:47:18.746717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.138184Z digest=sha256:8f51902750bb0811388ad7f131f39b48864339061f00ea2d221ea75f86b40204

Observation 5910e295-c639-4bc8-80ed-03a55f33dc01 · outbound

This paper cites Class-Discriminative Attention Maps for Vision Transformers.

GViT: Representing Images as Gaussians for Visual Recognition Class-Discriminative Attention Maps for Vision Transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.202139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.202139Z digest=sha256:76e25e52989b4167770758879b4edf37c28ddbe49d14699a758ad0383b1f3ad0

Observation fd5ef612-735c-428c-a5a7-dfeddb721ae2 · outbound

This paper cites Generative pretraining from pixels.

GViT: Representing Images as Gaussians for Visual Recognition Generative pretraining from pixels

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.235366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.235366Z digest=sha256:8acbded26494d59f82bbb76521f8f7687348ea739519c6460ccddf4f828410c0

Observation 936e8e54-4ae0-45ba-b2b1-bb233b9fc377 · outbound

This paper cites ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators.

GViT: Representing Images as Gaussians for Visual Recognition ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.283757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.283757Z digest=sha256:151910813c64e067519c4ae4c2cf7d365b2fe02925c20bb0483816a3a741c4d4

Observation f4c5c823-5540-4b5a-b51a-db3931724776 · outbound

This paper cites Vision transformers need registers.

GViT: Representing Images as Gaussians for Visual Recognition Vision transformers need registers

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.656085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.350695Z digest=sha256:985b2d4f19b067b734e045c15a73a4d5d5fcf39fce09e90a15db364108f59fe7

Observation 37e73944-bae5-47a0-860c-c20c99361830 · outbound

This paper cites Scaling vision transformers to 22 billion parameters.

GViT: Representing Images as Gaussians for Visual Recognition Scaling vision transformers to 22 billion parameters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.395995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.395995Z digest=sha256:895b416cd83ef1e40abc9f4e38fc65700c28e9abbdb7147fad188af1c10f55bf

Observation fdd7961f-9de7-4c5f-be6a-aead69cc1528 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

GViT: Representing Images as Gaussians for Visual Recognition An image is worth 16x16 words: Transformers for image recognition at scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.450427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.450427Z digest=sha256:af0837e3a97b7da09f70347b06e6a848a8e292138a65d428d49bff994bf763c0

Observation 9c3db3c5-a6d1-4b5d-ba0d-e706919159f9 · outbound

This paper cites Adaptive slot attention: Object discovery with dynamic slot number.

GViT: Representing Images as Gaussians for Visual Recognition Adaptive slot attention: Object discovery with dynamic slot number

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.466824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.569998Z digest=sha256:3f0146b539b9ebe3bc0d8e97c6b505fa752db888466780013ab96d623a4e072d

Observation f66e188e-2881-48ee-bffe-e431d0f62f50 · outbound

This paper cites 3d gaussian splatting as new era: A survey.

GViT: Representing Images as Gaussians for Visual Recognition 3d gaussian splatting as new era: A survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.364413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.604742Z digest=sha256:9a8a3c200b1103b1245f1b133c2533fed21d55765fabf3a68db1f56a704c5398

Observation c1fc16db-617d-46d9-86a7-162eed514be1 · outbound

This paper cites Efficient graph-based image segmentation.

GViT: Representing Images as Gaussians for Visual Recognition Efficient graph-based image segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.238363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.656350Z digest=sha256:36abe6b206d9f3670d6739faf5906dae8550809d4f056fe218cb2374ebb4202c

Observation 579af9de-e56a-451e-b5b9-2714f0311f26 · outbound

This paper cites Understanding the difficulty of training deep feedfor- ward neural networks.

GViT: Representing Images as Gaussians for Visual Recognition Understanding the difficulty of training deep feedfor- ward neural networks

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:18.119540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.719489Z digest=sha256:6b8afdc22b38d85c42b7669db0d1c1f90a7b688180826c4f749c03a3819ac03e

Observation ff0bb9f3-9761-48f2-affb-e7a0de90b98a · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

GViT: Representing Images as Gaussians for Visual Recognition Explaining and Harnessing Adversarial Examples

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.771826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.771826Z digest=sha256:741d729f8440620770a0bb586e8529353259df942dea641f89f2de1909542393

Observation b192aebd-5bf7-4416-977f-feba17913d7e · outbound

This paper cites Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour.

GViT: Representing Images as Gaussians for Visual Recognition Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.840094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.840094Z digest=sha256:8a325da288847c7e292e595fc24ff1eca98f88a35e9c65fe8fea1149f6f26f10

Observation 3c6a61b4-d77f-49b5-bc12-b73b59d2e83d · outbound

This paper cites Faster neural networks straight from jpeg.

GViT: Representing Images as Gaussians for Visual Recognition Faster neural networks straight from jpeg

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.992426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:11.888156Z digest=sha256:9b72617ae482d1d03b9977b60fbdae224456c0022bdb777e69e9c56791b436da

Observation dc0b8bf2-6a07-4b45-a0b1-4398de97033a · outbound

This paper cites Masked autoencoders are scalable vision learners.

GViT: Representing Images as Gaussians for Visual Recognition Masked autoencoders are scalable vision learners

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:11.946548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:11.946548Z digest=sha256:0194bf86131dfcd20478b2486e6acf4f6e01ab1e51c61ac7710e4f9b518787fe

Observation 261a297a-38f0-42d1-a87a-84055dd8365c · outbound

This paper cites Deep residual learning for im- age recognition.

GViT: Representing Images as Gaussians for Visual Recognition Deep residual learning for im- age recognition

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.880417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.019522Z digest=sha256:78c76f1a6fe85b9fc281ed48dda39f5091bd1bea033002d3ca3c2d704314f974

Observation 846f3813-b848-4d05-990d-3c5ce50e77ea · outbound

This paper cites Bytes are all you need: Transformers operating directly on file bytes.

GViT: Representing Images as Gaussians for Visual Recognition Bytes are all you need: Transformers operating directly on file bytes

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.757233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.058765Z digest=sha256:d3105451538f2ecedbf54c19e15654454f7ec9ca7801c63ae54cc5668318d83b

Observation 200ecfb6-b37f-43f2-9086-224e9134a633 · outbound

This paper cites 2d gaussian splatting for geometrically accurate radiance fields.

GViT: Representing Images as Gaussians for Visual Recognition 2d gaussian splatting for geometrically accurate radiance fields

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.643782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.080524Z digest=sha256:f9b9a97822d31a381b510164baf1374985275b97ef18550ae8164407aad0e681

Observation fa411bc1-4037-4db5-8286-9e0aeb480988 · outbound

This paper cites Perceiver IO: A General Architecture for Structured Inputs & Outputs.

GViT: Representing Images as Gaussians for Visual Recognition Perceiver IO: A General Architecture for Structured Inputs & Outputs

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.543035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.133865Z digest=sha256:c6180da4bf3530f7664f09477afc4ee1ce8dc455f33f374268cfa1e97c93bd5a

Observation 8b67a148-fc63-4ce8-97dd-954a45b029db · outbound

This paper cites Perceiver IO: A general architecture for structured inputs & outputs.

GViT: Representing Images as Gaussians for Visual Recognition Perceiver IO: A general architecture for structured inputs & outputs

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.398396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.186546Z digest=sha256:165280d671079857786aaf03049e6e35bfc4a22c4670daa9c7c6669d6456ae4e

Observation b1e8f434-2fcf-4c5b-a4b2-4ecf418d0a53 · outbound

This paper cites Perceiver: General Perception with Iterative Attention.

GViT: Representing Images as Gaussians for Visual Recognition Perceiver: General Perception with Iterative Attention

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.294073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.228976Z digest=sha256:e1d7a20ffe629c39557527216f691698a5172bb69076327c8c4cddba0c5de09b

Observation ea2f901f-f3cd-4cd1-80f6-79f9c737a85f · outbound

This paper cites Superpixel sampling networks.

GViT: Representing Images as Gaussians for Visual Recognition Superpixel sampling networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.179891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.286148Z digest=sha256:3b40a2980c4e9a260a4ab4d02f9c15afd4780080b5d9f951e3e8cb77bdabb6c8

Observation 3738d5c8-68ff-4a31-be72-2ea36e1b9ea0 · outbound

This paper cites Unsupervised image segmentation by backpropagation.

GViT: Representing Images as Gaussians for Visual Recognition Unsupervised image segmentation by backpropagation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:17.066064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.336586Z digest=sha256:84a511e0d3b5819d70c1cc44bb25e30e5b0293c8219976193bc4a860889529ca

Observation 56d5519e-e377-4930-93b9-f89ca3ee78ca · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

GViT: Representing Images as Gaussians for Visual Recognition 3d gaussian splatting for real-time radiance field rendering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.962101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.396900Z digest=sha256:7c7d06dcc2bcd10dee8753c450d42317de47ecd9e283e124fd65cacd2ee3fa7f

Observation a6efc2d2-c4bc-4755-a961-289d45fa081b · outbound

This paper cites Seac and the start of image processing at the national bureau of standards.

GViT: Representing Images as Gaussians for Visual Recognition Seac and the start of image processing at the national bureau of standards

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.825354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.452282Z digest=sha256:d7d989fe6f4257290fa0977c9c89cc9589e6342d790afc3eae5a1afb8ca9cf1a

Observation 7ab25380-ec07-491d-b429-42da834e935e · outbound

This paper cites an unresolved cited work.

GViT: Representing Images as Gaussians for Visual Recognition Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:47:16.628991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.508278Z digest=sha256:fee8ad397f0e648ded486d8ec8632a057ba108c7c947e9e9cb922cdb442e87d5

Observation 6ba2d68d-75c3-4501-8c4f-b387e64c880e · outbound

This paper cites Kutulakos, David J.

GViT: Representing Images as Gaussians for Visual Recognition Kutulakos, David J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.415159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.567226Z digest=sha256:36d58beb30adc91e4122c5d396b485f2278b095c55d5d00f50d04065aa3c9a9f

Observation a6cba2a0-19ec-475c-a0c7-e0cc0eefc094 · outbound

This paper cites PyTorch Distributed: Experiences on Accelerating Data Parallel Training.

GViT: Representing Images as Gaussians for Visual Recognition PyTorch Distributed: Experiences on Accelerating Data Parallel Training

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:12.598568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:12.598568Z digest=sha256:b28d30151c592d1a9b9452a47656b73707c23d29e7dece51d6682cbac8c940ae

Observation 2c815d31-fe80-4758-a764-4601fe534137 · outbound

This paper cites A convnet for the 2020s.

GViT: Representing Images as Gaussians for Visual Recognition A convnet for the 2020s

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.205710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.646382Z digest=sha256:f5d5a973c4bcd20bc71f69363832a60d8ac9c4ef30cb7bd90cb7d8b64f7c2723

Observation bdd6e5d4-4c34-4d8b-9461-9dcca0ac6c66 · outbound

This paper cites Object-centric learning with slot attention.

GViT: Representing Images as Gaussians for Visual Recognition Object-centric learning with slot attention

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:12.704280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:12.704280Z digest=sha256:d7e47dd4e6ab7c925a643205266872843ca38f63f835bcd8c41672130bb2a1e4

Observation d926a667-08f5-4c40-bb5b-b0e508e55a00 · outbound

This paper cites Sgdr: Stochastic gradient descent with warm restarts.

GViT: Representing Images as Gaussians for Visual Recognition Sgdr: Stochastic gradient descent with warm restarts

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:16.093469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.746140Z digest=sha256:e86feec2740b0a03d20f891fe7a17ebe21b067bb3f3d08f07d304d4efef3c812

Observation 30cabcaf-6846-453f-a2e5-9f7095397d4b · outbound

This paper cites Decoupled Weight Decay Regularization.

GViT: Representing Images as Gaussians for Visual Recognition Decoupled Weight Decay Regularization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:12.793127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:12.793127Z digest=sha256:103cb9a2c4575bddcc1ac1911d9ab77ae4aa2b50c8cba63665a9d7b2c908430b

Observation f9c6bfc4-7a58-4297-a4ed-650674a643db · outbound

This paper cites Enhance the Visual Representation via Discrete Adversarial Training.

GViT: Representing Images as Gaussians for Visual Recognition Enhance the Visual Representation via Discrete Adversarial Training

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:47:14.362286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.882028Z digest=sha256:e9baf7ee5cc626845281e59d0894437e912c75e5f95b3a951327b115ded3037e

Observation cf18a7bb-9b6b-4674-8f32-bc1ff44f1e3c · outbound

This paper cites An image is worth more than 16x16 patches: Exploring transformers on individual pixels.

GViT: Representing Images as Gaussians for Visual Recognition An image is worth more than 16x16 patches: Exploring transformers on individual pixels

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.980100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.940750Z digest=sha256:84b1496c52bfcf7f3a701dfd6c2a34a081a6feda6428323d8e77813f04998d43

Observation 42d1f6b9-685f-462f-9500-ad2e411fcf97 · outbound

This paper cites an unresolved cited work.

GViT: Representing Images as Gaussians for Visual Recognition Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:47:15.875334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.965051Z digest=sha256:e32e111c0b41fa7ca2bc341cb4e339a4cc5ce181abac58c64931f2124cb08e13

Observation e517d215-d6b7-45cc-b96d-efacca6aeec9 · outbound

This paper cites Rgb no more: Minimally-decoded jpeg vision transformers.

GViT: Representing Images as Gaussians for Visual Recognition Rgb no more: Minimally-decoded jpeg vision transformers

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.695604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:12.993161Z digest=sha256:db9bbd20fc5b54b04bc9a4ff666a12812aef912ed5e6df0b9a4fab70173d5dca

Observation 51ff19dd-a76a-45ab-9c72-8799982fd377 · outbound

This paper cites Gaussian Masked Autoencoders.

GViT: Representing Images as Gaussians for Visual Recognition Gaussian Masked Autoencoders

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:47:14.169384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.040502Z digest=sha256:587389bf919e21675cb8c8f499c459294d22085b369b6c349f802c9a90196b1e

Observation 326b0d3b-41f2-4a27-9b57-1129e725fce4 · outbound

This paper cites Learning a classification model for segmentation.

GViT: Representing Images as Gaussians for Visual Recognition Learning a classification model for segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.493510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.079031Z digest=sha256:48ed5d94c762f492dc30d763aa619e4a303d81362a7ccdee403d9e5d6472316b

Observation 0836e8a5-4ef1-445b-bacb-576c6956ee30 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

GViT: Representing Images as Gaussians for Visual Recognition Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.101183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.101183Z digest=sha256:1bf4a811e3f622bc602c4bce443bf4c8b45deac43ba189451f1f89541bd606dd

Observation c0ea71a1-20eb-4917-9fc1-5c0ec8c287e8 · outbound

This paper cites How to train your vit? data, augmentation, and regularization in vision transformers.

GViT: Representing Images as Gaussians for Visual Recognition How to train your vit? data, augmentation, and regularization in vision transformers

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.305740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.151121Z digest=sha256:448f2ab8650e08e2263b0e9336173874ffe22a982186047f56614120602c7157

Observation a4d8a551-bbbe-44e0-a939-83fbef92fc93 · outbound

This paper cites Superpixels: An evaluation of the state-of-the-art.

GViT: Representing Images as Gaussians for Visual Recognition Superpixels: An evaluation of the state-of-the-art

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.212064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.212064Z digest=sha256:e15ade5fd0432ec0a02d950cab403323c826aa22522d1ee095656dfcb5487e31

Observation 5340c48a-b528-4227-a4de-b16e52df9857 · outbound

This paper cites Single-view view synthesis with multiplane images.

GViT: Representing Images as Gaussians for Visual Recognition Single-view view synthesis with multiplane images

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:15.139091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.236571Z digest=sha256:0adff1abb4c09b0de70c7ebaeac054840172a568aa6a669169184f21d7d63f96

Observation 635841d9-9702-4ce9-8f0c-7eb7efbe101f · outbound

This paper cites Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion.

GViT: Representing Images as Gaussians for Visual Recognition Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.279370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.279370Z digest=sha256:d7de0a83ada96024d89bc9b37dae3842dac618139e6ff9de9b9d08499e6e4970

Observation ffd14c20-ccef-4ae2-ab4f-20b68397693d · outbound

This paper cites Matching networks for one shot learning.

GViT: Representing Images as Gaussians for Visual Recognition Matching networks for one shot learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.354326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.354326Z digest=sha256:217457b10621b2a2463a6ba8c0f29e4abb9ed592bd2ac64738438f448e098329

Observation e1d282c4-f13d-41c1-801a-ac2fb39a08f0 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

GViT: Representing Images as Gaussians for Visual Recognition Image quality assessment: from error visibility to structural similarity

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.357382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.357382Z digest=sha256:7156da52096e84d838913d5ed11767c293a29e092173d9d78e49944e26b77515

Observation 06c8f80a-1f85-441c-86cf-7007377b6e37 · outbound

This paper cites Beyond Language Models: Byte Models are Digital World Simulators.

GViT: Representing Images as Gaussians for Visual Recognition Beyond Language Models: Byte Models are Digital World Simulators

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:47:14.007426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.372641Z digest=sha256:8506296a2a7dc74be80bf247d2f46cf6de90e1e09b1dd4db94f73e1a9a525b5e

Observation 5135e461-daef-48a1-bcdd-0af3d01951a0 · outbound

This paper cites Learning in the frequency domain.

GViT: Representing Images as Gaussians for Visual Recognition Learning in the frequency domain

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:14.983179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.442882Z digest=sha256:ea079f1dba4c810fa906c9e2dc583086435388da0f612c967e1c93b2adb2b6d1

Observation c8e13307-e537-4105-9de6-390e916900c0 · outbound

This paper cites gsplat: An open-source library for gaussian splatting.

GViT: Representing Images as Gaussians for Visual Recognition gsplat: An open-source library for gaussian splatting

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.546287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.546287Z digest=sha256:f20911bb89a253e0a24d37de24870bd9111c0664fa6a1eda3dfc861416f60548

Observation 3daf0d46-cb1e-448f-a58b-7f7da4af2f18 · outbound

This paper cites Vector-quantized Image Modeling with Improved VQGAN.

GViT: Representing Images as Gaussians for Visual Recognition Vector-quantized Image Modeling with Improved VQGAN

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.608125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.608125Z digest=sha256:79198cf0c8a762bd21a51ef56d9dc8c97629e0746a2e14c5702a593c81eb22d3

Observation ba8285a7-b6e8-454e-b4eb-2ab2b9d54e0a · outbound

This paper cites A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond.

GViT: Representing Images as Gaussians for Visual Recognition A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:47:13.697596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:47:13.697596Z digest=sha256:180e6b3f4003721e089d2737c2cc58965d47e58e4969b86b21a2fc96477ba97b

Observation 8da1c4bd-7218-4bd1-8d90-9ac99f21af3d · outbound

This paper cites Gaussianimage: 1000 fps image representation and compression by 2d gaussian splatting.

GViT: Representing Images as Gaussians for Visual Recognition Gaussianimage: 1000 fps image representation and compression by 2d gaussian splatting

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:14.800213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.794092Z digest=sha256:b17d1feef2485bcd75327a4f089f057847219fe6db601bfbad04051a59887215

Observation 9d935b03-d3ab-47a7-9c6d-74efc9df2911 · outbound

This paper cites Self-supervised learning of object parts for semantic segmentation.

GViT: Representing Images as Gaussians for Visual Recognition Self-supervised learning of object parts for semantic segmentation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:47:14.591969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T21:47:13.856071Z digest=sha256:24184b6247345fe9c3e1f24bbbb1e8e1b1efb57b2890668aae369ef05cb128e4

Pith citing papers

No inbound Pith citation observations are available.