Pith. sign in

Paper Citation Record · LEDGER

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding

As of 16 August 2026, this Paper Citation Record lists 100 of 287 outbound references and 0 inbound Pith citation observations for arXiv:2608.12748.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.12748 v1

Coverage vector

measured 100 of 287 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:09:18.495392Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 287 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved98
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d5ada3e7-5e40-4954-926e-02eec0f49532 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.131394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.131394Z digest=sha256:2e8c8bd049067aeeeced7461f94cdfe975cbb7c9fefc8d833e580d3be9a658cb

Observation a81588f9-4d07-4de5-8d0e-2e7ba330ccfe · outbound

This paper cites Applied Sciences , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Applied Sciences , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.135472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.135472Z digest=sha256:dce9fbbb77937445bc46d0832b4af0d7f94160c447a286ee4c3ccb9f4b910b67

Observation 7a1532a6-cb00-4695-8031-9a8beedf3cce · outbound

This paper cites Signal and Data Processing of Small Targets 1993 , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Signal and Data Processing of Small Targets 1993 , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.138865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.138865Z digest=sha256:5d1e4bd87a7682feb0a5c12aac69e19d5788c33a1e8f6b548df0fa222726d6d9

Observation 3febe7c0-0ae0-40c8-80fe-e7aed9825ef5 · outbound

This paper cites Sensors , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Sensors , volume=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.143478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.143478Z digest=sha256:c58b6715e6ee63b191d640eadf8097bf1504c5d81e90b4755d84edeb4bfc19f8

Observation f5a86670-90e2-4c0a-aba4-449af04cb6c5 · outbound

This paper cites Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , pages=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.147495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.147495Z digest=sha256:c1379cf11263c8c42b514713d71d8ac46d2b65bc958bdf88caf7349ce6d8f657

Observation cb5ff312-aa1b-4059-9149-a9c07845df58 · outbound

This paper cites Sensors , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Sensors , volume=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.150786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.150786Z digest=sha256:cdf10f8bf9fa98019a92f41551009f4eb8f86d78f423e555ec1bb035840a7940

Observation 6cac7bdf-ef22-4c44-96bf-66fdb7166eae · outbound

This paper cites 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.153828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.153828Z digest=sha256:8a985521b817a44fbe71197b310c562ceb439e13b9892c2280c27bd6232ffe6d

Observation c637ee4e-6a73-4ed3-acfd-8baccf73b7a7 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.158526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.158526Z digest=sha256:9f01f71df4446ac2d551ca6124a505ba25285686f17ad806f6e39f62716bac7c

Observation ee327802-dd29-482f-9599-43871f64980d · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.161713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.161713Z digest=sha256:758da4ab4a03744f37ad9cd31b5dc5af5e7a4d360f0be3e7f4305720def1a77b

Observation 6713d00b-68a6-407a-ac83-7f0e7df22efd · outbound

This paper cites International Journal of Computer Vision , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Journal of Computer Vision , volume=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.166091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.166091Z digest=sha256:81aca2a48d2123b41e859675789f3b68e889476fe03ada46802a3415932a0187

Observation 198ced30-915a-4898-b3d7-56b6f1096078 · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.169235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.169235Z digest=sha256:4805a8ebe307470a608ffbf82df1e8e7f74fa9e4741c8562d718f641f2717b71

Observation 1ef4a677-338f-4948-8ef1-88dd4e14e538 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.173688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.173688Z digest=sha256:d3e2652d0496814283d09470f1f969a542b13c86baf3e86b02649a091b129419

Observation f40c3e1e-bc8a-4f09-ae70-5f1976c6468d · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.176903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.176903Z digest=sha256:0ea34571d6e7700d2ecdd524136166604bccf3cf204ce2fbf6cb717787debec3

Observation a44725a2-b514-444d-b8f9-0221777dcbf7 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.180236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.180236Z digest=sha256:5f975a5a18853b10d07f23723ab385235fdfcc49ee5966805ae85dc76298f9a5

Observation 2384866c-174a-4bfc-b306-6b373233115a · outbound

This paper cites International Journal of Computer Vision , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Journal of Computer Vision , volume=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.183249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.183249Z digest=sha256:402d13836ac0fc9f5a70d76a9e49aef9af3495547ff4ca29695a53fff4409276

Observation 3150fbd9-0146-4dc5-9f1e-09ffd12aff8c · outbound

This paper cites Infrared Physics & Technology , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Infrared Physics & Technology , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.186095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.186095Z digest=sha256:d5e417dbd7def703dd7f78dd6833cbc9e8e13bc3bf9e03e10563d98595e3c579

Observation a41111da-8d77-4612-a8ce-e501f6f99bfb · outbound

This paper cites Kaur et al.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Kaur et al

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.189270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.189270Z digest=sha256:31d51ecf50ab90dfa6f7911a93e2919ca4412a7c8a6ea85c2560af85480b8abd

Observation ea7d1d10-5292-4302-826b-a1e991a320aa · outbound

This paper cites , author=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding , author=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.192080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.192080Z digest=sha256:1bd6de0e0d9b06b6b4ed363339247da4a95ef8f17ae3f2b9ab9811af278c8bc3

Observation 738f01d4-a2f8-47b3-8b23-7b6b60045d01 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.195382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.195382Z digest=sha256:e662e113ae912df6dd0ab9e4b62f2fae98c31e45700b861de75cc3db532f82bf

Observation 169abf66-44fd-43ea-a19e-a84eb2cf17fb · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.199801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.199801Z digest=sha256:1421cd041aaa0f9f2ab37b6f1821dac520f545927d6a3e00d740b19a83e84a8c

Observation fb978c6a-5f8e-4476-9cb5-8f2be0c79f39 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.203629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.203629Z digest=sha256:fc8175ce96a2f3994347c3e8204946f53296509909d9e61fda87e43488a7a715

Observation 6a3d78cb-6c88-43a7-8995-2200ec1fd3a1 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.206600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.206600Z digest=sha256:53b29a8fb433f57cee1f3b5bf6a22c62cb39560a19f4b044331ec53be0d5a1e3

Observation 9e8b2e84-e73c-4461-91c4-9bf00688c6d6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.210068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.210068Z digest=sha256:dfb0208425e594e207c8f09b0f20184bca354efa2f29961700c6a247abcbf19d

Observation 198ab2ef-01d2-41d7-923f-fb195d791173 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.214296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.214296Z digest=sha256:3866a7eb02126b0af0e4fc43e3469e1a5873bee19b7614adcc0091c1b2794de4

Observation 790836ec-2a49-46b1-80d6-4ad9eb6be0a5 · outbound

This paper cites Proceedings of the 32nd ACM International Conference on Multimedia , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the 32nd ACM International Conference on Multimedia , pages=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.217495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.217495Z digest=sha256:b16ee6d8890e9cead8fc15700c193550d43fab1bcb00a64408765ce48894de64

Observation e5373747-d872-4110-aa85-a96428c7291d · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.221276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.221276Z digest=sha256:16a0982fc1848fa0943768553ee5c842e75a6ca2e5aaa639745ed61ed2fbaefe

Observation 0af7ddf2-e113-46d1-9b70-d2ab4e6bab68 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.225163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.225163Z digest=sha256:91252fda658dca4a1777e66a01043e3ad7e2141b0c6428729cf6ed9854ed13dc

Observation a9b04ec1-e611-4fc4-9a3b-085206799bd3 · outbound

This paper cites International Conference on Machine Learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Conference on Machine Learning , pages=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.228795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.228795Z digest=sha256:911d5f010c72a633dcf95641a29ff4f321f9a1ce8972d4c378b3c8333ea3aa55

Observation 7c113b82-daa2-47f6-aabb-a6533ce590bc · outbound

This paper cites 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding 2024 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.232127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.232127Z digest=sha256:d5497931419d075595f636d7049e6fb59a1b776c5d7f570f312ea0ac1e322b4f

Observation 166969f3-0ec7-4d50-8c18-99cf64e4e238 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.235430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.235430Z digest=sha256:880c5f36a327f32a274433ac8e3405324df3bd7b53987b84d1300a36d90496c4

Observation 631494c5-6dca-4ca1-a8b1-6655b275f04d · outbound

This paper cites Visual Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Visual Intelligence , volume=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.238598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.238598Z digest=sha256:78eb0cfec695da60b683c1d0a74cea24813a5b4803358b4bff7d01dc122c88ce

Observation dd53a003-ce60-4e95-b197-78914fbff2e7 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.241956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.241956Z digest=sha256:7551605afc185fccf39d04eb2d92c5bbaac116253947db21eaa96beb67681309

Observation a14fb8f5-6243-4fdc-85da-f6b0fde0383a · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.245336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.245336Z digest=sha256:63f493553d77b2e3120a9bcd554cd551c9185f781e0d476659a5654bae50369f

Observation 24ddecf9-0999-43a8-b6dc-1a18525ae5f2 · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.248499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.248499Z digest=sha256:1450c76b1acb0faaad1aba1eb5a8f20845fd6243685389f55e8bdc4460f345ee

Observation 38806458-f364-4829-b4fb-fc5b09da3680 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.252010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.252010Z digest=sha256:8b266775fa6a26c1da95bdee4c3f93e66646d6c475b6b98ce2da78c53c6f0ea3

Observation 47feca59-be9f-4838-aecf-0d40e502dc86 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.255311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.255311Z digest=sha256:777e6c863539bbaddbb02969889dd51da728ac9761dcb04c6f80d033336fc023

Observation d0ff5c32-6006-4d1b-9c83-d82dc1017147 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DINOv2: Learning Robust Visual Features without Supervision

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.259063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.259063Z digest=sha256:be7dec89d89185bf09a9e817c659822700f8c23ded7cf722c8dc5f2253159d78

Observation bbd3066d-cf2a-47a7-a5eb-dfda3484a6bf · outbound

This paper cites Proceedings of the AAAI conference on artificial intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI conference on artificial intelligence , volume=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.266450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.266450Z digest=sha256:0e335ca57696b52690002574e37ff265640b749c4a91753561f73ed3fae88b23

Observation ff44c456-9f22-4a7d-8893-38d52e44efd7 · outbound

This paper cites DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.269408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.269408Z digest=sha256:9736ff01fac35ee037e451e90f0e9286fcfb9fd73b6c4c7b14700e112b45e705

Observation b413ac14-a37f-46c9-8167-fba1bd896f1f · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.273874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.273874Z digest=sha256:8b750343fa260b74894b161d767bfcd8ff2a78dbf43aae58494bae583183f175

Observation b8c98ffb-b1ff-4c18-a79d-eebcbcaeb8b9 · outbound

This paper cites Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.277272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.277272Z digest=sha256:96b13af9637a56a66c5d70f8346a267bed905a695a108bcdf570e5a72aec582f

Observation f1431835-d342-4cc3-88c2-5a7d44c75e60 · outbound

This paper cites International Colloquium on Automata, Languages, and Programming , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International Colloquium on Automata, Languages, and Programming , pages=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.281345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.281345Z digest=sha256:93edd73912d27da2f4767ecfea95952778887767df3256b38df7d7df3473daea

Observation ba2da6cd-5520-4f45-987b-2d330b2ec766 · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE international conference on computer vision , pages=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.285481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.285481Z digest=sha256:efd9afc8bfca5b6ae89c328aa1038543f3a9fe22324adf2a18c7bad99d60e4cc

Observation 3033de1b-51b1-4341-b512-1827915139ba · outbound

This paper cites Tensor Fusion Network for Multimodal Sentiment Analysis.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Tensor Fusion Network for Multimodal Sentiment Analysis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.288993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.288993Z digest=sha256:cd1a13a9341bae9b5c0827ace413a9c88d0cedbf4e24474bf9ed2b8f19be4318

Observation 93cf72aa-8317-4099-b5b8-1fd6a6d3c3ac · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.292484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.292484Z digest=sha256:70f1cdf90466bd11824cc9e6e5575d045049953ff4c45a657ff6cbf763282378

Observation 862e57a0-a390-4f67-8b7e-6b008c20bb70 · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.295725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.295725Z digest=sha256:ebe9695cea7d61798bdcc276b226a435578326de4acba7fd9395fc771725e84a

Observation d13c8183-9b69-418d-9ef0-3948308fe9d1 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Fast Transformer Decoding: One Write-Head is All You Need

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.299513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.299513Z digest=sha256:7d8fb112a485f78bee5d7600e43ad383565832edc0f31cacc19fcb552d1fe3f9

Observation b342f152-1b0d-45bf-8ac5-deeb7ca75eeb · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.303079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.303079Z digest=sha256:4d4a16c7eb8e55bb4fd642e8815953bf06644c5fd7607e5883b20be8a8f1d977

Observation aeb73f93-bd3a-49a8-88a1-8b5313b2fb04 · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.306273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.306273Z digest=sha256:125182525ea636b13771f770d5b07b6e8955a8563b29e71e1ff16c2ddcd5e375

Observation 58aa6810-536d-42ef-8104-bda0302ae315 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.309435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.309435Z digest=sha256:f0a49bebd243d5623e77afb033790bdde410d52e10e9a066bc3a77ddd6f1c4bc

Observation 20c76a08-459a-4705-bf22-b9973526bb4c · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.313282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.313282Z digest=sha256:fa2ddf4f9a8c6d08444b6868c092492ea2f9e232705d747c3e150f2b68655144

Observation 5c05fb9c-087a-4a51-a5f0-11564571a76e · outbound

This paper cites Scientific reports , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Scientific reports , volume=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.316532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.316532Z digest=sha256:16dbf29ed8e25491ddd25350ed9711ef4b7b1d48820a820114e549dbad58d2d9

Observation 434536f2-3d05-492c-9ff9-c499cacfbf0f · outbound

This paper cites Advances in neural information processing systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in neural information processing systems , volume=

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.320625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.320625Z digest=sha256:a6fc49e3a7f847ea6e92a300e83df6d3b7639597d0032c75ff1a1b772f0a2523

Observation 20aeb9fc-85c8-4089-b7b9-e16e7c68811d · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.324123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.324123Z digest=sha256:2f7d24e8dc568e6017d1fa4bc7aec646fb08eb8105d0319e0172c27d3682cbdf

Observation 4f1fd045-68ac-451c-992e-7c6c08929aa8 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.327291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.327291Z digest=sha256:65e63662629f96fe8eaed63530db38c865439adbaed9885c52d6e6731e88d3a9

Observation fe591112-dc85-4e93-b70f-bc762eb20c72 · outbound

This paper cites Retentive Network: A Successor to Transformer for Large Language Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Retentive Network: A Successor to Transformer for Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.330658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.330658Z digest=sha256:35b7e4c9aff0fb555c9a0af220abadfe62442fad9d4b27d02df11bbe3ba43a2d

Observation 834b29cc-6fe3-49c8-a74f-c1a1d8b1ec3d · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding RWKV: Reinventing RNNs for the Transformer Era

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.334068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.334068Z digest=sha256:08c1f2b1af386bb1ed4415c394c355688136054dfd369492e8afb902fa109e0c

Observation 037c8e0f-fe59-4ff0-89cc-69a691be0dea · outbound

This paper cites A Systematic Analysis of Hybrid Linear Attention.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding A Systematic Analysis of Hybrid Linear Attention

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.338225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.338225Z digest=sha256:f624d92ba76e4bc4fd67d43e9ff4de9790696c3fb008287fcdcdb65fc57644b8

Observation b631e196-28cc-4f9c-b96d-fe9e23d8d46b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.342324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.342324Z digest=sha256:df71ec1d8dd8423f7d7020b1d1ffcbadfdbcf050718fa7c3efc7a81783ef032a

Observation 0be8a52b-9bda-4cc7-bec2-dc520ed3ac70 · outbound

This paper cites IEEE Transactions on Knowledge and Data Engineering , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding IEEE Transactions on Knowledge and Data Engineering , year=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.345488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.345488Z digest=sha256:3d6f4c44f1a3fa3caa98f3aff2f2fd5e464c182dba503dddaeea5c17a6197d5b

Observation ae418849-8255-46b7-b87d-7f14f0bec818 · outbound

This paper cites Learning deep representations by mutual information estimation and maximization.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Learning deep representations by mutual information estimation and maximization

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.348668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.348668Z digest=sha256:e3371833a5cfc0185445b5677fee99fc4befea8aea6ba1adae42bc3b4d58a36c

Observation 173da82d-718b-401c-ba00-c9b11063f096 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.352825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.352825Z digest=sha256:cb0bfbd0c115238311d43caca5da75140b88b82a9b3b8143ba335cff4c3073c1

Observation c8d034c2-0382-4b57-95fc-bebd575fd027 · outbound

This paper cites Towards Achieving Perfect Multimodal Alignment.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Towards Achieving Perfect Multimodal Alignment

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T00:09:20.152184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T00:09:18.356102Z digest=sha256:e5be84e2b6a24b738a0c26a21ab3bce232fb7368e9c277a5eb066909876c9db2

Observation 94bcdb0f-c482-477f-bbbd-02555ecee220 · outbound

This paper cites Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.359522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.359522Z digest=sha256:9b80e121952d9925292c6fcca66dd3e5c225408d1f18856cab0ff3dfe9beb5e9

Observation 6cddadf8-7db4-4e33-b6a6-4492a9084913 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding The Thirteenth International Conference on Learning Representations , year=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.362682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.362682Z digest=sha256:24795444b120e3ea4b52a0bfa59cd0ca5b86e29d0edc1fec4944184b655f0b48

Observation d0c07504-d2c5-4e3e-8e0b-66e28e7418a5 · outbound

This paper cites IEEE Transactions on Visualization and Computer Graphics , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding IEEE Transactions on Visualization and Computer Graphics , year=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.366327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.366327Z digest=sha256:e4340634a16b11291161da180c8c8b9ffbdce533000bff4e08bc71775ba1f5be

Observation 294af071-c99c-47f8-9fba-cfcf71802ec3 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.370409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.370409Z digest=sha256:24896e5a5190151ae099cb4959fe5f8cb8376bba7ec051e8e41215f83421b464

Observation 76608d80-5443-43ae-a212-acd4d5b81c2b · outbound

This paper cites Enhancing Conceptual Understanding in Multimodal Contrastive Learning through Hard Negative Samples.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Enhancing Conceptual Understanding in Multimodal Contrastive Learning through Hard Negative Samples

Reference 69

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T00:09:20.132947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T00:09:18.375123Z digest=sha256:bf95a3b030db22b7f9f7c5fa6f411561072b6015c70ad65fd6dc1dbc1c268ac6

Observation 3663de83-66cc-40f6-80e1-6fc00da5e1fa · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.378889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.378889Z digest=sha256:37a43ffdcbb2e9037193caf9b53bfb08daa29d4aada90932d9b9886aa9cfa131

Observation 9a27dc3a-387a-424d-80b6-d1842df26b8f · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.382502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.382502Z digest=sha256:f9d053e25c434d1fbe7938394426ffdb8ade875e42f75c9c241f0e5b1e1a3cb0

Observation 582bd7ee-ebb1-42f8-9fe3-c54da476fd65 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Representation Learning with Contrastive Predictive Coding

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.386507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.386507Z digest=sha256:f19443ddedc9f0529641c4ebd59250d315b37cfdbfa3746d23d3b23bee586504

Observation 3f693ce0-b55d-4375-ace4-aff1f1285ad5 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.390945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.390945Z digest=sha256:edda4a88d0189ec89b779e73ea0da19d58304ae8c70a63037dce7f3984717b66

Observation 807fff39-ebf6-415e-8287-d348b5d9300c · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.394896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.394896Z digest=sha256:7d31b97ece99f82f03e825917e6b9886ec7f4c7c8bfd779315090b8885b3bcb9

Observation 84c5ce49-402f-46c9-a776-285a972c0e64 · outbound

This paper cites Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.398783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.398783Z digest=sha256:ea5487a5aad8bf53f0558972ad2276c99f3b4b1a26a06aec580ff50ebfab63e6

Observation c6fa70a8-fc79-469b-963b-e9ce45a69b42 · outbound

This paper cites DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.403553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.403553Z digest=sha256:5df0369c5227f896b3e0d3bf32807cd6a686bec8d002c20f606c0fe125248cde

Observation 61759e2c-5b3e-42ff-a09e-69345ff19fa6 · outbound

This paper cites Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.407130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.407130Z digest=sha256:f469e85d640a4d880d016145bd5f7908a71ad1da60fd5c66a9767eaa9c347c20

Observation f235c6f8-8702-4b7f-a42a-dd3fdcfcf2b1 · outbound

This paper cites arXiv preprint arXiv:2503.07465 , year=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding arXiv preprint arXiv:2503.07465 , year=

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.410782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.410782Z digest=sha256:b8c29f49a9da083da676b35d9a226f524c2749ec21d7ebc5afbb818ed470911e

Observation 757268e6-1c12-4301-ad2b-0de76dddf4d6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.414921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.414921Z digest=sha256:3c4499d22a343821eb705c9365d4864a456d4567c8b625f6cd1f63d7a1a5e877

Observation f442bf30-05d6-41d0-ab8f-0d770b481288 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Advances in Neural Information Processing Systems , volume=

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.418075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.418075Z digest=sha256:e05fdaab62713b6dee04ace4bbd9486e719f1629db339f6963fe1a008ada3ae2

Observation 4a822874-dff6-4f9a-9e56-ec55b0e4d893 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.421565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.421565Z digest=sha256:cd98ce72982e9be237c87bd58e9f2b7b69e576848c4c53feb312a17c05d92632

Observation e908ccd2-a150-485d-bceb-743d546d6521 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.424416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.424416Z digest=sha256:30b959d5e1903de76bdf69b05d349adc95f733c970ee29f380d88587596048f8

Observation 94186c65-14e1-4ef6-9934-bd93689fb394 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.427521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.427521Z digest=sha256:a059883a3205afb422c396a78c6c251d4b13ac51cc13828fde03f983b0351c02

Observation b545fd16-3ec7-416e-9ce4-00c77ae285a1 · outbound

This paper cites Open-vocabulary Object Detection via Vision and Language Knowledge Distillation.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.431661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.431661Z digest=sha256:8f5aa54c0ecf6577177285a9416d1cbd28cb94b3e0ff2fef39a4264bc406d08c

Observation 3e262bb6-28bc-4f79-872e-2f0e029ba1dd · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.435447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.435447Z digest=sha256:6d4a021804af35b86bfd82300068828b90eba751854e43fe0741e53fbb535281

Observation aa95a6b9-2d9d-4569-9f10-9f499d54cb9f · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.439386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.439386Z digest=sha256:f70e77ce7b04f3961ff92cd15c6610e75e725ad7bef104ebb38ce16021427060

Observation 16e0af0a-7474-46f1-97c5-a6ce637d7988 · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.443344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.443344Z digest=sha256:790e1b10dcc4235df4f7959588d2fe860caf557ce6c5ac43da6aa71cb24e8916

Observation 58e84b30-1e5e-4bfc-82ff-4a46363cabb2 · outbound

This paper cites VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.447397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.447397Z digest=sha256:eae4db57a7689c3e1ef2dac93eb5692908c879d853e704434073dbbaff2f0548

Observation fa101d85-636c-47f0-9dc6-9f4878968d35 · outbound

This paper cites Reconstruction Alignment Improves Unified Multimodal Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Reconstruction Alignment Improves Unified Multimodal Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.451539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.451539Z digest=sha256:c23f3c513aed719883291e3b216de490af66e3f872cb708feb86f27c8f5909e5

Observation c4606724-0164-4c2d-921f-7e9932cb4b89 · outbound

This paper cites Reconstructive Visual Instruction Tuning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Reconstructive Visual Instruction Tuning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.454827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.454827Z digest=sha256:7e0da9a0ca3f5101f95e7ffb81956910df5e71fc67962b801098b71eabe9b81e

Observation b9059908-295c-4d50-99aa-3e19fd988604 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.458692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.458692Z digest=sha256:46bb2a676c4bcce1f9d1cef91b96391cd703145f71b390cefce88378ada36da7

Observation b044f2a9-f4b1-4c70-bb86-18d3710013ed · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.462406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.462406Z digest=sha256:c8b02bc7a3320db32c9ed5e38dca2f38451a8b3c60a877ee537a6559842a9ebb

Observation 2305c43d-1a7b-475a-bb5b-92fcffeac25e · outbound

This paper cites AutoVP: An Automated Visual Prompting Framework and Benchmark.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding AutoVP: An Automated Visual Prompting Framework and Benchmark

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.465655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.465655Z digest=sha256:3c22b1ea1774bd9a53c3946021b7124872370262825ed3b9fb53a015dc064e1e

Observation 2098a7e2-3c1e-42b0-ad3d-8915e61b50df · outbound

This paper cites European conference on computer vision , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding European conference on computer vision , pages=

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.468974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.468974Z digest=sha256:f1640672395f14d43f4100f498c3fda15550c0db6246bfbaa050f4177d0065ab

Observation 5a7ef476-59b4-47a5-a3ba-43babdb96744 · outbound

This paper cites Exploring Visual Prompts for Adapting Large-Scale Models.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Exploring Visual Prompts for Adapting Large-Scale Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.472945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.472945Z digest=sha256:2e855e8cb1605ba2ee573235315039b9446b81826f2dc74f6ae07b98a1ae1256

Observation 7b18dd3c-38af-4f9b-a832-10a6e057418e · outbound

This paper cites , author=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding , author=

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.477311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.477311Z digest=sha256:3d40dbac0b7c1a8a227a2596fb1ba894c1c183b666267692cb4d6c37b62f3e44

Observation 6b4b153e-f91d-494f-b8b1-5e528e43d0a3 · outbound

This paper cites International conference on machine learning , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding International conference on machine learning , pages=

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.480340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.480340Z digest=sha256:caeb564712bee96d06ed1a358959fd202bf32b47c606d7cffa0345447f4addd1

Observation 35dfea57-56f9-4d8f-a039-9ece697abcb4 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.483633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.483633Z digest=sha256:83439a8b745cf489be416b4a3f4582fe0dcc783d19b4ee31ffcb90f6ffacfbb3

Observation 431e6702-d31f-4fcf-be6d-7a75c5fa7469 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.487512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.487512Z digest=sha256:4ae4d0497efe4886e615d75d40ace7c37c46bd1d4cef4789c0c3404309d42943

Observation c7c9fe7d-7655-4784-8ee2-dc942b7961f6 · outbound

This paper cites P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.491464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.491464Z digest=sha256:1ab702b3712ad0b0cb6342950fd0bf35bca76cc5097b3f011b558e21948c7262

Observation e587b3c4-fedd-4f32-9535-011c1aad0ba1 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-16T00:09:18.495392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:09:18.495392Z digest=sha256:2977a1281f33aa2c100e5239b294638ee299bb4ff49395cba43b633eb4af6029

Pith citing papers

No inbound Pith citation observations are available.