Pith. sign in

Paper Citation Record · LEDGER

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding

As of 16 August 2026, this Paper Citation Record lists 100 of 287 outbound references and 0 inbound Pith citation observations for arXiv:2608.12748.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.12748 v1

Coverage vector

measured 100 of 287 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:09:18.495392Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 287 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved98
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d5ada3e7-5e40-4954-926e-02eec0f49532 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.131394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.131394Z digest=sha256:26bb27f7372698914d4d65b3677a63d1751dfd5574e0db6e56ff036dcf89823c

Observation a81588f9-4d07-4de5-8d0e-2e7ba330ccfe · outbound

This paper cites Applied Sciences , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Applied Sciences , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.135472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.135472Z digest=sha256:322cf4a34e2ba6992c8c2fab9acf241c624ca003d99bb97ed178478c32e93edc

Observation 7a1532a6-cb00-4695-8031-9a8beedf3cce · outbound

This paper cites Signal and Data Processing of Small Targets 1993 , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Signal and Data Processing of Small Targets 1993 , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.138865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.138865Z digest=sha256:dc394ae6154b586c312509b6dfb481f30182acb57a5636119c6436cda9894d6e

Observation 3febe7c0-0ae0-40c8-80fe-e7aed9825ef5 · outbound

This paper cites Sensors , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Sensors , volume=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.143478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.143478Z digest=sha256:29b91af39c1443550c5966dab57e4beca7fad1b1b2f738f533ed276df742cbf8

Observation f5a86670-90e2-4c0a-aba4-449af04cb6c5 · outbound

This paper cites Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , pages=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.147495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.147495Z digest=sha256:29e88449820004c607e0254a0a38367c965115634e7337627a43de8cf5cd1629

Observation cb5ff312-aa1b-4059-9149-a9c07845df58 · outbound

This paper cites Sensors , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Sensors , volume=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.150786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.150786Z digest=sha256:c6b383b7671154099cf5166c1184459424316a4fae56b54b36b1ac2369dbe8bd

Observation 6cac7bdf-ef22-4c44-96bf-66fdb7166eae · outbound

This paper cites 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.153828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.153828Z digest=sha256:a6300a5c6fdbf232f502f5405cf5ac2f5b9bae775ca8075955d71f7708a4edd2

Observation c637ee4e-6a73-4ed3-acfd-8baccf73b7a7 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.158526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.158526Z digest=sha256:056a821e69637523ef0c248af5b15f7a032ebc3c3232205f075aedb533b32bd0

Observation ee327802-dd29-482f-9599-43871f64980d · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.161713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.161713Z digest=sha256:1b6bd73f3819520c1d9fb70457b9a43f4917b0a614108ca64cadf323fa06fbcb

Observation 6713d00b-68a6-407a-ac83-7f0e7df22efd · outbound

This paper cites International Journal of Computer Vision , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Journal of Computer Vision , volume=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.166091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.166091Z digest=sha256:88987ec2ef4ef1cb5254bd54ee1615612e82764e67dc7b15feb5e412856fe4ea

Observation 198ced30-915a-4898-b3d7-56b6f1096078 · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.169235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.169235Z digest=sha256:2d08a9cc29f8a8936d93a903c8d0bbb8cacece4ee03bb795385bb32e4e8aa068

Observation 1ef4a677-338f-4948-8ef1-88dd4e14e538 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.173688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.173688Z digest=sha256:b2dcd6c46b32b7499c3518a59c2d8b15edd13ff91a18d80f5b4eac4b78d0c151

Observation f40c3e1e-bc8a-4f09-ae70-5f1976c6468d · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.176903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.176903Z digest=sha256:d4aeeef4cce4796372c36e62247d757cda404ecfe4fe23c2c8e91994672ce55d

Observation a44725a2-b514-444d-b8f9-0221777dcbf7 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.180236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.180236Z digest=sha256:7fedb324d8b7596d3760339ad05da9070fee132361cb26e5ff6eeb260c81cdbb

Observation 2384866c-174a-4bfc-b306-6b373233115a · outbound

This paper cites International Journal of Computer Vision , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Journal of Computer Vision , volume=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.183249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.183249Z digest=sha256:3edfd84470a4a06c2cfa2563eaf02e3c37d8196831269ac7a7bc0df3fdcf7268

Observation 3150fbd9-0146-4dc5-9f1e-09ffd12aff8c · outbound

This paper cites Infrared Physics & Technology , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Infrared Physics & Technology , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.186095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.186095Z digest=sha256:ca0caa10d567917601473f26f2f18e658820d0399973ae48545ee4e9546cf9a5

Observation a41111da-8d77-4612-a8ce-e501f6f99bfb · outbound

This paper cites Kaur et al.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Kaur et al

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.189270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.189270Z digest=sha256:6cf5ffbb447fb77941af262c30f076554190edcef642531f7b1d1ebe4ddddc7f

Observation ea7d1d10-5292-4302-826b-a1e991a320aa · outbound

This paper cites , author=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding , author=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.192080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.192080Z digest=sha256:67e6a925e9f0730a16d1c62286b369c9724bf3a45fac57976c50e34e5e8bda3c

Observation 738f01d4-a2f8-47b3-8b23-7b6b60045d01 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.195382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.195382Z digest=sha256:e4ee2bd6a9261ec001174ae47f7dea2de489b140cfb0e0adcb012201fa7a374c

Observation 169abf66-44fd-43ea-a19e-a84eb2cf17fb · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.199801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.199801Z digest=sha256:541743565d7583acdff30e92726a66308fddb81f9d175742c5e9ab9ea340193d

Observation fb978c6a-5f8e-4476-9cb5-8f2be0c79f39 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.203629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.203629Z digest=sha256:3a4deaa2eabf6301da49070eb8e9b03e1395e4b3825177ab197483b54edc140b

Observation 6a3d78cb-6c88-43a7-8995-2200ec1fd3a1 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.206600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.206600Z digest=sha256:052a895e79dbbc24dec50c74098aaf596ab4f421fabc98908b86e520d7c76de1

Observation 9e8b2e84-e73c-4461-91c4-9bf00688c6d6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.210068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.210068Z digest=sha256:1d109aa6c3f9a6aa534764571b4686b0aad3665041935a4cfcc0bd7b7b7a525b

Observation 198ab2ef-01d2-41d7-923f-fb195d791173 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.214296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.214296Z digest=sha256:befd9b4d3a628addf5182d3aca70dc7d9233fbf681224f6f72d211456b8b2e89

Observation 790836ec-2a49-46b1-80d6-4ad9eb6be0a5 · outbound

This paper cites Proceedings of the 32nd ACM International Conference on Multimedia , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the 32nd ACM International Conference on Multimedia , pages=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.217495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.217495Z digest=sha256:00ec35acb9a449215d7c26336c18d2995701216b8992d6047b8b3fe9cb4a5518

Observation e5373747-d872-4110-aa85-a96428c7291d · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.221276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.221276Z digest=sha256:b06f7c1e142e00aa0b1604de3288481553e2dacda5cb3e65da6c69ee82f81b2c

Observation 0af7ddf2-e113-46d1-9b70-d2ab4e6bab68 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.225163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.225163Z digest=sha256:4e65780745423fa5635e74dc4237eb87bf0257da432ae5c5be1ab9092c8441d5

Observation a9b04ec1-e611-4fc4-9a3b-085206799bd3 · outbound

This paper cites International Conference on Machine Learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Conference on Machine Learning , pages=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.228795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.228795Z digest=sha256:e3ef2708023cf4226c51f1032419d80bb1a5ec1ffbe2c1b33931d0cb562a6dc3

Observation 7c113b82-daa2-47f6-aabb-a6533ce590bc · outbound

This paper cites 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.232127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.232127Z digest=sha256:091bf6afc7dd8afb595d1b6a0ac466de467fdaefef374c41ad2110e9c06dba00

Observation 166969f3-0ec7-4d50-8c18-99cf64e4e238 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.235430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.235430Z digest=sha256:719a9472d5010442cb5a8dedfdf2a642c6bfa63b31818a8f3bdde94bb41c21f2

Observation 631494c5-6dca-4ca1-a8b1-6655b275f04d · outbound

This paper cites Visual Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Visual Intelligence , volume=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.238598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.238598Z digest=sha256:ef28f6834816f00eea1022dcb41ccfcae59646dafb05c207562d4e801e4e85a1

Observation dd53a003-ce60-4e95-b197-78914fbff2e7 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.241956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.241956Z digest=sha256:22284f48c3f31bd384b113730f77e03e2b64cf482793552bac3f25e9b1ac0bb7

Observation a14fb8f5-6243-4fdc-85da-f6b0fde0383a · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.245336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.245336Z digest=sha256:f511d7c63aaaa842593c184edc2170bc2b4608c9211e4ec8b3029d87832b1708

Observation 24ddecf9-0999-43a8-b6dc-1a18525ae5f2 · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.248499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.248499Z digest=sha256:6e6e761a5bc6c6625222bfaea0dd5d93414e7ecd89bf1b6e06daff2385f72c2b

Observation 38806458-f364-4829-b4fb-fc5b09da3680 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.252010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.252010Z digest=sha256:507842323c82b1a4aaf004d3238deb41fe4d08d7dc98dfcc9b33a65e5208f6cf

Observation 47feca59-be9f-4838-aecf-0d40e502dc86 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.255311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.255311Z digest=sha256:730082a0d7457ec77d92b2ac2f2e3a85992e6f76134ff8c8e0cfdce51cbd63f7

Observation d0ff5c32-6006-4d1b-9c83-d82dc1017147 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DINOv2: Learning Robust Visual Features without Supervision

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.259063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.259063Z digest=sha256:c45757a7a303a03882c176ade650660ac7c0e513eac9cd6e9944fc215dab686b

Observation bbd3066d-cf2a-47a7-a5eb-dfda3484a6bf · outbound

This paper cites Proceedings of the AAAI conference on artificial intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI conference on artificial intelligence , volume=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.266450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.266450Z digest=sha256:e3d1686b2b6e9b91e5609fd8eb162a2fa630911316cdfccf85b374b91a617b9c

Observation ff44c456-9f22-4a7d-8893-38d52e44efd7 · outbound

This paper cites DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.269408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.269408Z digest=sha256:aa7ee6f55694f027e54f47db6de1ceeb32e394b67991f6348da829904d67f5e9

Observation b413ac14-a37f-46c9-8167-fba1bd896f1f · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.273874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.273874Z digest=sha256:67557d80249597ce6597331e043afffd45655f6c1c416e1054687a375edc7e0b

Observation b8c98ffb-b1ff-4c18-a79d-eebcbcaeb8b9 · outbound

This paper cites Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.277272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.277272Z digest=sha256:c7eb521622e1bf708b503b8b903160a367f0392d123556951c75d1daf6794a86

Observation f1431835-d342-4cc3-88c2-5a7d44c75e60 · outbound

This paper cites International Colloquium on Automata, Languages, and Programming , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Colloquium on Automata, Languages, and Programming , pages=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.281345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.281345Z digest=sha256:51ca001c4cc4ee418009793109cf49b8ee264c06a955a5cb16209aa8d6dbda8d

Observation ba2da6cd-5520-4f45-987b-2d330b2ec766 · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE international conference on computer vision , pages=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.285481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.285481Z digest=sha256:56b2adc7050c93184d3bc9a81e2efc68ef7bf13176c31dcfd036c72c6a2179c7

Observation 3033de1b-51b1-4341-b512-1827915139ba · outbound

This paper cites Tensor Fusion Network for Multimodal Sentiment Analysis.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Tensor Fusion Network for Multimodal Sentiment Analysis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.288993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.288993Z digest=sha256:fe30504d46d38ad0eb605f09d46a019fdf522e008171154ee113e9158d7edd84

Observation 93cf72aa-8317-4099-b5b8-1fd6a6d3c3ac · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.292484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.292484Z digest=sha256:5dc6155c24e088abddb63290b99fe6f1bd763de31b4e32b34a7363d0a1a68259

Observation 862e57a0-a390-4f67-8b7e-6b008c20bb70 · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.295725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.295725Z digest=sha256:bb8caaa4cd9db8ab27ca217c05ee53d8fc348e6cf344eea721bc515152490634

Observation d13c8183-9b69-418d-9ef0-3948308fe9d1 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Fast Transformer Decoding: One Write-Head is All You Need

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.299513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.299513Z digest=sha256:a9d97b8ae35f29e4504d755f7530aa8ef47f751a4a14d3d1145c4674b224e2a1

Observation b342f152-1b0d-45bf-8ac5-deeb7ca75eeb · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.303079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.303079Z digest=sha256:b4bc630ebfde470a7567a460305e3560243e9e2fdef277f6a4f5958a4bfeb09c

Observation aeb73f93-bd3a-49a8-88a1-8b5313b2fb04 · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.306273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.306273Z digest=sha256:a0883e4228940dca4136ef0d13f80fa027fcb8a7a2fee9801d222d2089bd32bf

Observation 58aa6810-536d-42ef-8104-bda0302ae315 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.309435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.309435Z digest=sha256:4e533900904597970aaad5b7d044170cfe88b03a9fc95f917d68435fbb5a513d

Observation 20c76a08-459a-4705-bf22-b9973526bb4c · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.313282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.313282Z digest=sha256:87304c859a34d357042dc3a898048273ad46d39f41fb9ea78bdc91c702e2f344

Observation 5c05fb9c-087a-4a51-a5f0-11564571a76e · outbound

This paper cites Scientific reports , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Scientific reports , volume=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.316532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.316532Z digest=sha256:f9df5b5d974d740d7a5699507394b88946aaa51e030bb05cf39e2b99f743c0dc

Observation 434536f2-3d05-492c-9ff9-c499cacfbf0f · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.320625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.320625Z digest=sha256:fded42a4293bb6fcc2aa45adba6bb942768d2d6725b8dcad662d4f4b5509b5f6

Observation 20aeb9fc-85c8-4089-b7b9-e16e7c68811d · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.324123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.324123Z digest=sha256:666b5a718870d9e4ab7c2c09c56ec9fcfdf8dc3f2a5cfddf0eb908a8e276b9b2

Observation 4f1fd045-68ac-451c-992e-7c6c08929aa8 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.327291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.327291Z digest=sha256:aa09b12078424a726692db2edb35d41ecaea9c167a161e7dfbfc3180ed92133d

Observation fe591112-dc85-4e93-b70f-bc762eb20c72 · outbound

This paper cites Retentive Network: A Successor to Transformer for Large Language Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Retentive Network: A Successor to Transformer for Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.330658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.330658Z digest=sha256:ee4cd8236c353c91e2bd1dab9894990dd339d2f952014f17ad8429331a79f030

Observation 834b29cc-6fe3-49c8-a74f-c1a1d8b1ec3d · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding RWKV: Reinventing RNNs for the Transformer Era

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.334068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.334068Z digest=sha256:6665b50a9b4edefcd8a72f5993a00e9cabac803785ae351dfd62f3773ca3d1f8

Observation 037c8e0f-fe59-4ff0-89cc-69a691be0dea · outbound

This paper cites A Systematic Analysis of Hybrid Linear Attention.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding A Systematic Analysis of Hybrid Linear Attention

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.338225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.338225Z digest=sha256:b363cc57bc1e2affd71453dfe449b431b5614d4e5de4f6027b313a220f5b1c6c

Observation b631e196-28cc-4f9c-b96d-fe9e23d8d46b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.342324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.342324Z digest=sha256:e863c502c74eb189506cf25a03e498cdddf6b11d3ae34dd48dea39cce4cb24e9

Observation 0be8a52b-9bda-4cc7-bec2-dc520ed3ac70 · outbound

This paper cites IEEE Transactions on Knowledge and Data Engineering , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding IEEE Transactions on Knowledge and Data Engineering , year=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.345488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.345488Z digest=sha256:e14531a192d019ef7c7657419a4376d19dbb08f18a0eff58fc276dc90cbcbf28

Observation ae418849-8255-46b7-b87d-7f14f0bec818 · outbound

This paper cites Learning deep representations by mutual information estimation and maximization.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Learning deep representations by mutual information estimation and maximization

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.348668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.348668Z digest=sha256:c2bee75b489ff12c4a7071d0cd11b1415366896e3f3374a6cff90ba34ba68560

Observation 173da82d-718b-401c-ba00-c9b11063f096 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.352825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.352825Z digest=sha256:99805eda9b7977731830b2952ac172aa5a553ef288068593b8e7bf00eefe4cf4

Observation c8d034c2-0382-4b57-95fc-bebd575fd027 · outbound

This paper cites Towards Achieving Perfect Multimodal Alignment.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Towards Achieving Perfect Multimodal Alignment

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T00:09:20.152184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T00:09:18.356102Z digest=sha256:fa2e7d2a67a6d32627e4018236f0ecf81332e9201c7e98e587e91717bc3a5b00

Observation 94bcdb0f-c482-477f-bbbd-02555ecee220 · outbound

This paper cites Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.359522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.359522Z digest=sha256:c9282ac59880fc2ef9321354162b4e3434013dd154fef17123ce3849cd2cad0f

Observation 6cddadf8-7db4-4e33-b6a6-4492a9084913 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding The Thirteenth International Conference on Learning Representations , year=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.362682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.362682Z digest=sha256:d9a39ce34e7544c2c66e8386214325fcdc0ef9b76638975adea97941ff493d5f

Observation d0c07504-d2c5-4e3e-8e0b-66e28e7418a5 · outbound

This paper cites IEEE Transactions on Visualization and Computer Graphics , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding IEEE Transactions on Visualization and Computer Graphics , year=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.366327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.366327Z digest=sha256:9fc8a52fa708d54fd22197c0da4b1a7a54068a7b3d55741986ea373845eb0b4e

Observation 294af071-c99c-47f8-9fba-cfcf71802ec3 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.370409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.370409Z digest=sha256:ac7fe8a6e8cdce83ccc66df2ad1789394ca18f354ef986fab27ab5aa6eb63230

Observation 76608d80-5443-43ae-a212-acd4d5b81c2b · outbound

This paper cites Enhancing Conceptual Understanding in Multimodal Contrastive Learning through Hard Negative Samples.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Enhancing Conceptual Understanding in Multimodal Contrastive Learning through Hard Negative Samples

Reference 69

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T00:09:20.132947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T00:09:18.375123Z digest=sha256:adfe17b1cb221c4c78bc5d9789c6ebb98dba7c1f75c237d20dcc1f7b57c58b07

Observation 3663de83-66cc-40f6-80e1-6fc00da5e1fa · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.378889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.378889Z digest=sha256:52f53655d55043c1a8918f2970fb5d75fae7ffc1ff2b85335cd55787f0d35aac

Observation 9a27dc3a-387a-424d-80b6-d1842df26b8f · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.382502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.382502Z digest=sha256:eca6437452e4f81161bc4340063788a258ff3b20ab1d7f431a9830503f63ed17

Observation 582bd7ee-ebb1-42f8-9fe3-c54da476fd65 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Representation Learning with Contrastive Predictive Coding

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.386507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.386507Z digest=sha256:688ff11d5733500712b4884fcf432cfa3b1927c8a5bdf8aa370bd8e5ca3b6d26

Observation 3f693ce0-b55d-4375-ace4-aff1f1285ad5 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.390945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.390945Z digest=sha256:bfddc2e61d372461094f16cb1c16858e3a00dec852f070bdac56df28e28634fd

Observation 807fff39-ebf6-415e-8287-d348b5d9300c · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.394896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.394896Z digest=sha256:2ba8bdc1618efb3e46aff09764afed7ede3c2ac37fa9617e3e5554263eaf5d66

Observation 84c5ce49-402f-46c9-a776-285a972c0e64 · outbound

This paper cites Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.398783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.398783Z digest=sha256:c4c95c66b85c080e4213a615d8884f6da41ffecd9198ecd5d666c4d67f737a72

Observation c6fa70a8-fc79-469b-963b-e9ce45a69b42 · outbound

This paper cites DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.403553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.403553Z digest=sha256:5f59f9883e2d37b5ef7b3920d0e2322d4ed7350e520d80016cffa63d981de422

Observation 61759e2c-5b3e-42ff-a09e-69345ff19fa6 · outbound

This paper cites Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.407130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.407130Z digest=sha256:0521dc323976d1b796543e22a377334af57bf24cce31899d32c4fcd8cb221e25

Observation f235c6f8-8702-4b7f-a42a-dd3fdcfcf2b1 · outbound

This paper cites arXiv preprint arXiv:2503.07465 , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding arXiv preprint arXiv:2503.07465 , year=

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.410782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.410782Z digest=sha256:ec8e8bb67d0e224300176c52300737a3f55637cbe937021b83cdc8f18087cbfa

Observation 757268e6-1c12-4301-ad2b-0de76dddf4d6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.414921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.414921Z digest=sha256:19dfda36efa922ed88b39875461514c36429738f90eb01901928bdc54b69b2b5

Observation f442bf30-05d6-41d0-ab8f-0d770b481288 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.418075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.418075Z digest=sha256:f9fb662a7b130668925523c7c8994bf696e2edd634f7e73c7feb674b8cd4d973

Observation 4a822874-dff6-4f9a-9e56-ec55b0e4d893 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.421565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.421565Z digest=sha256:0c2f3ea14418a4ae50be434f27bb01794a29ccf823d75053652af1f538b12142

Observation e908ccd2-a150-485d-bceb-743d546d6521 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.424416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.424416Z digest=sha256:287101ae826a4661bf7ec6de813a65765e7ba314305a4c71d5e0d71249e70c10

Observation 94186c65-14e1-4ef6-9934-bd93689fb394 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.427521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.427521Z digest=sha256:6abe548eadcac7253f00c6b3e12b7d62e61271490042f470c5a89acd5f9401cd

Observation b545fd16-3ec7-416e-9ce4-00c77ae285a1 · outbound

This paper cites Open-vocabulary Object Detection via Vision and Language Knowledge Distillation.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.431661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.431661Z digest=sha256:0953c8616d16bb604dcbfd823e210ed362161bedda65028f9315c52d64265cd0

Observation 3e262bb6-28bc-4f79-872e-2f0e029ba1dd · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.435447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.435447Z digest=sha256:eca52b3bd0c109f6f6961325c27ebd153afdca6a9f2527d74d6c0a83817a4198

Observation aa95a6b9-2d9d-4569-9f10-9f499d54cb9f · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.439386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.439386Z digest=sha256:83cbbdc56d0d18ae23fbf87676d443259e0d729e25080ed2d96bdc426fc202d7

Observation 16e0af0a-7474-46f1-97c5-a6ce637d7988 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.443344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.443344Z digest=sha256:a7f8b78044b48da917d9d80dedf6f42038a9a8b86e19bb6d16ffd497665827b8

Observation 58e84b30-1e5e-4bfc-82ff-4a46363cabb2 · outbound

This paper cites VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.447397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.447397Z digest=sha256:3b963e3ef44b487570b7a8b460c35428d70f9a33e925f0eb3b8792144d64eacd

Observation fa101d85-636c-47f0-9dc6-9f4878968d35 · outbound

This paper cites Reconstruction Alignment Improves Unified Multimodal Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Reconstruction Alignment Improves Unified Multimodal Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.451539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.451539Z digest=sha256:b1f03975c17568da240a85060db3827a0c5f16aa50edc91f6d9a13cd4d693e55

Observation c4606724-0164-4c2d-921f-7e9932cb4b89 · outbound

This paper cites Reconstructive Visual Instruction Tuning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Reconstructive Visual Instruction Tuning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.454827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.454827Z digest=sha256:8d7de8c52ad4591b0d8f356a946633848db5161eecb651fe627eed7077b068a1

Observation b9059908-295c-4d50-99aa-3e19fd988604 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.458692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.458692Z digest=sha256:62d3c222358c46988cff933fc9adcff9b0d9da3e194efbb01a6d3cbc6f32026a

Observation b044f2a9-f4b1-4c70-bb86-18d3710013ed · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.462406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.462406Z digest=sha256:f1ce6b451c3d534716a14cf5814e7cc5d99e5dc90bc510af2504666dc7d22a94

Observation 2305c43d-1a7b-475a-bb5b-92fcffeac25e · outbound

This paper cites AutoVP: An Automated Visual Prompting Framework and Benchmark.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding AutoVP: An Automated Visual Prompting Framework and Benchmark

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.465655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.465655Z digest=sha256:0e5daff077052ad4af31097263754d268032ea5f4027520c178510ac7107efae

Observation 2098a7e2-3c1e-42b0-ad3d-8915e61b50df · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.468974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.468974Z digest=sha256:94bd9576aa4c23ce08d39cf39542438826cd79641825159a16da577fd0a22d42

Observation 5a7ef476-59b4-47a5-a3ba-43babdb96744 · outbound

This paper cites Exploring Visual Prompts for Adapting Large-Scale Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Exploring Visual Prompts for Adapting Large-Scale Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.472945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.472945Z digest=sha256:f93268b44faf78b42b0947c9de0ebec57394fa37e3fd5aa6bbe1fd2f02b24b46

Observation 7b18dd3c-38af-4f9b-a832-10a6e057418e · outbound

This paper cites , author=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding , author=

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.477311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.477311Z digest=sha256:9c43a637d91ad45a301734422558b91770b6ebeca2ca2432e50dd61bf5ca8460

Observation 6b4b153e-f91d-494f-b8b1-5e528e43d0a3 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.480340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.480340Z digest=sha256:e6bf126c0b6f363cc7f17af7b5eb2a194525b285c9fada9e166378caad986dc6

Observation 35dfea57-56f9-4d8f-a039-9ece697abcb4 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.483633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.483633Z digest=sha256:032b879cdcaf352f95ebb62b5cffb5435c2a8f697ceae9c729c33af60b4e9ed2

Observation 431e6702-d31f-4fcf-be6d-7a75c5fa7469 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.487512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.487512Z digest=sha256:a621b0bc15bfd37fd2c57f44a900ee95830ffe57e4080b05d4de49ac69556631

Observation c7c9fe7d-7655-4784-8ee2-dc942b7961f6 · outbound

This paper cites P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.491464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.491464Z digest=sha256:7783423fb46f8b8df9f333bdd9665407e72f1fba5755100af1a571a25fe902d8

Observation e587b3c4-fedd-4f32-9535-011c1aad0ba1 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.495392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.495392Z digest=sha256:8c30ddc1b24aaeb729554937f5e842dd5134b2d30b680410836e72d8dea6e0a6

Pith citing papers

No inbound Pith citation observations are available.