Pith. sign in

Paper Citation Record · LEDGER

EGM: Efficient Visual Grounding Language Models

As of 5 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2601.13633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.13633 v3

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T13:07:55.655699Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T23:38:51.488742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact6
  • verified fuzzy29
  • unresolved10
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch13

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a6173652-6af4-4e58-9664-5d89c0a99f58 · outbound

This paper cites GPT-4 Technical Report.

EGM: Efficient Visual Grounding Language Models GPT-4 Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.734420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:d76961abbaeda2e79946562f2999ea11160f2c1ee119822e053dead1279a8c9c

Observation 17d41e28-cfab-479a-b096-2881c783bb00 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 2

Resolution
parse uncertain
raw_fallback, observed 2026-05-16T13:10:57.606712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:07951849e67fbdc0e231d147d8f218581cf3111ab953d080028f873a33235fb1

Observation e3fee928-080c-4cc0-a3fa-6803df9e2f3c · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.571737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:907a4b335d25ac45813abb5606bb70ff800db603bc16d380a992bbfac1d11074

Observation 5202c00d-aabc-4b9c-8091-f04881ccb8cc · outbound

This paper cites Qwen2.5-VL Technical Report.

EGM: Efficient Visual Grounding Language Models Qwen2.5-VL Technical Report

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.749395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:ea7735990849c65f3dc8379222969aac1758bfbe074110b72bbcdf22441c7424

Observation bdb57217-bae4-4d4d-a551-d3b7f0b7a0f6 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.601763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:9d79dbb2504f36a70cfbd5d612f50774e565b69444ab8119ebb73b26d3be1f22

Observation 9374484c-766c-47a3-8528-315e8d0bfa21 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.583624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:8e265a2951a5c07ecfddbaa6d5417a464e5b214ada0555e07e5b8fc8454b846f

Observation f557b503-347e-4997-8c4d-7c825f3316bf · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.588840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:cfac85f4eb62c691ade4a3c4609ab3fd8268d7941e182fd0865be4765820f204

Observation 73c063f4-5930-464d-98eb-6750b130d49f · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

EGM: Efficient Visual Grounding Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.731528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:2c24e7b20899a1dad9c1491a8276dba99ac251471f75d175a6cc8c120c04adb9

Observation 9145da1d-2f48-4990-93b3-964499d31458 · outbound

This paper cites arXiv e-prints pp.

EGM: Efficient Visual Grounding Language Models arXiv e-prints pp

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.616917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:98d608f987294d00efaf748916a594fd24747a4e0b6254b9b766becd4245a246

Observation e871a1c0-4d5f-4179-afeb-a96f3ce206c4 · outbound

This paper cites arXiv e-prints pp.

EGM: Efficient Visual Grounding Language Models arXiv e-prints pp

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.596770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:a570aceb61d7561eeb6caae17a53fef1073ddaed9c3b227a3c46816ec4bf06e0

Observation 2033dc5d-a4ba-41ff-8e0d-01fa538263d6 · outbound

This paper cites Proceedings of the Ad- vances in Neural Information Processing Systems (NeurIPS) (2024).

EGM: Efficient Visual Grounding Language Models Proceedings of the Ad- vances in Neural Information Processing Systems (NeurIPS) (2024)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.568872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:8d7950fc2e59195a929f96c9b2911458f57de348122179d4b3c933f1f150b786

Observation ce7bc1b4-cd7e-4b0f-9a93-7db533d80b9a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

EGM: Efficient Visual Grounding Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.719054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c1083f513462ca0e5f334088b24dc0e3afa708739580b91a9e607c3e21e9d753

Observation ab5e14a9-a4d5-4f02-aad9-42cfe5919bcf · outbound

This paper cites TAO-Amodal: A Benchmark for Tracking Any Object Amodally.

EGM: Efficient Visual Grounding Language Models TAO-Amodal: A Benchmark for Tracking Any Object Amodally

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:10:56.714909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:5d4b181aa9794c6739d81f5892ae42d2049e490fbed5ac8834e92d277fec4391

Observation ea035e0a-1f92-4a7d-98e4-9ff029116b21 · outbound

This paper cites GPT-4o System Card.

EGM: Efficient Visual Grounding Language Models GPT-4o System Card

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.725695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:16cadb593b115b1d91dc75c503676b7c2ec2e5205196ec82c5cd851dfd91760e

Observation 3aba5604-d2d8-4a51-a8a3-0161ebcdfe5e · outbound

This paper cites Psychological Research88(2), 307–337 (2024) EGM 17.

EGM: Efficient Visual Grounding Language Models Psychological Research88(2), 307–337 (2024) EGM 17

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.614041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:259093d1f5d351c4715107320561ebbd797c445ef4cfb0815c5408effcb4ba37

Observation 4b004b11-4843-495d-bac8-1d1f29523eea · outbound

This paper cites In: Proceedings of the 2014 conference onempiricalmethodsinnaturallanguageprocessing(EMNLP).pp.787–798(2014).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the 2014 conference onempiricalmethodsinnaturallanguageprocessing(EMNLP).pp.787–798(2014)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.604327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:d38f3935923f9e7194762019d0a90fad8391f4b9d33d8e54b0103a76cb42be9a

Observation 27bf4157-3794-4325-9f68-d34c245c1a6a · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.610286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:da571617ec1dab6b55acef5c013b9c2839723a158a30d22e1d87883d21cffeb0

Observation 06dab8fa-903d-4831-8a6d-40dd5ef4b1a7 · outbound

This paper cites In: Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles (2023).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles (2023)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.611850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:dbf2d1a255f3a49417aca99b754ba505725c66e950d3d5d5cae0e67b1ea43842

Observation 69071cf1-ea2a-41f0-b636-c806130a04da · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

EGM: Efficient Visual Grounding Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.737621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:0ee638e579d700dddac6ba07c98e308016b06edc95d3008d0a2d9b3b21d2bd9c

Observation a0e41b03-23e4-44f4-9025-879840fce802 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the European Conference on Computer Vision (ECCV)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.619094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c486492ffecc126ba5735b03fb4b13a4690f041520d9945060b9d11ed58a6ab3

Observation a42441e3-423f-4393-8f34-56d2fa134908 · outbound

This paper cites In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.621580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:7ae4fd2a89da178dd38b0e3ca17a0415fa15a21b73fdb0480c1b5cc59f137b75

Observation cd6dc917-a95e-438b-a24c-b981569f3c02 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the European Conference on Computer Vision (ECCV)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.591576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:1a955ee3a072707e7ec451f39ad3b80705017545cf18e3e07c72f08a687b6b74

Observation fff80151-937a-4af2-bd96-e9fe936685d6 · outbound

This paper cites IEEE Transactions on Multimedia (MM) (2023).

EGM: Efficient Visual Grounding Language Models IEEE Transactions on Multimedia (MM) (2023)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.623745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:f5499b11754408c855469931f18795daeb6f3576a2f314620ccec3607e49bf03

Observation 1d09fc3b-caed-4dd3-8d68-f08324997b17 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.625998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:d1547689f2182f8f7b07f5430e66f480834129d09e45cb38e18fde7eba1e8cd7

Observation 8e0eb2c0-3a9e-48f4-b9aa-3da62be6e593 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.586577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:3a030b1e01761f7d56f00fe28a612680ebb0a11da9af2169e154924062cb15dd

Observation 16d83e56-4db8-439d-99a5-6d4e170db180 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.587743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:bb1204cc4423ebd227eb8cbf9141dcc43cf8672eee60b74b6d8d7a546891d1ea

Observation db26aa49-bdb0-4431-a281-a926174af312 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.593569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c0a1af2f5aa6b03bfa9c7a03c8397c3dad25114af73782e8c20764c3593905d0

Observation ed0eae98-da38-42de-be17-6a7d87ad7f8d · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.609352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:50376e528babea389e1eacb797201bd3c8c2b4d268651aab17a94c28b765b88a

Observation 20af6d65-1e89-472a-acd5-e58e16de1f38 · outbound

This paper cites In: Proceedings of the IEEE conference on computer vision and pattern recognition.

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE conference on computer vision and pattern recognition

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.591288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:e3cae7892818c30b569f16314d43bdebb9945e32ac410030c7c1feaed4b3bfca

Observation 788a646b-dc6b-4146-8462-0d85452d4fd6 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

EGM: Efficient Visual Grounding Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.706873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:9c83f719a3fd0816bfcd14c4567161c0da6e153e68af48b96a575e475204cf50

Observation b8b534b6-04cd-42bf-b771-754bf3d6eb09 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

EGM: Efficient Visual Grounding Language Models HybridFlow: A Flexible and Efficient RLHF Framework

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.740725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:55f82d5fb8d8127aaac4d4753cafe20225312e89c9d4573d59ede4d186ac1505

Observation a39eb763-a5eb-4593-9e5d-205bccb1c8a0 · outbound

This paper cites Gtpo and grpo-s: Token and sequence-level reward shaping with policy entropy.arXiv preprint arXiv:2508.04349.

EGM: Efficient Visual Grounding Language Models Gtpo and grpo-s: Token and sequence-level reward shaping with policy entropy.arXiv preprint arXiv:2508.04349

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:10:56.722822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:fe2517098ea1a47450a08db841594c9efc5071ed80fe22c00720a98488581bc5

Observation 80558958-e0dc-4f5d-848f-b8ac3e96cea6 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

EGM: Efficient Visual Grounding Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.728483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:db79244af10dc218d9259970526a94277f670d24c521438e335d87434f4bcbbf

Observation 650fd6b2-170b-4c43-ab65-78ef70da1454 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

EGM: Efficient Visual Grounding Language Models Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.746512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:0dc64321c47ac29c345203e4e202b91f6dcc56e9e237aeec13df530fc5ac14fe

Observation 803bb090-7124-412a-ab9f-4a258a95aebc · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

EGM: Efficient Visual Grounding Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.743529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:a699286549b3e4023108305d874ee652e43262826ca10cfa708e66a44fa76c2c

Observation 6c83cf21-5767-4c75-b8ee-c9aae01b99e8 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

EGM: Efficient Visual Grounding Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.702965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:bb96ce5b37d506ba91400de67f604c9291ed5694c843dee8b932b020c5a7774c

Observation 6a1c0a8c-c229-4107-baab-02f1846fb20c · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

EGM: Efficient Visual Grounding Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.689386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:abc572f09ee8fb64a026fa09c54d6a55af660f469e0ccfde3fbc92408d45d90c

Observation 0af3fd9f-14c7-4f89-8c2d-17653cc9f45c · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

EGM: Efficient Visual Grounding Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.692513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:32b4dcbb6617efb1180512cea4b883547d185fd66ad0cefcad1941d07975aeba

Observation 83fe5f8f-d0e0-4606-8235-a429a6b25ece · outbound

This paper cites Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images.

EGM: Efficient Visual Grounding Language Models Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T13:10:56.707028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:ed819ce792d9b5fce4068bfea1732f8df8945fcab6ae2e72b9f91ac46d00df66

Observation 8b77253e-ecc1-4f49-b04e-4492587c300c · outbound

This paper cites International Journal of Computer Vision (IJCV) (2025).

EGM: Efficient Visual Grounding Language Models International Journal of Computer Vision (IJCV) (2025)

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.602387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:0e58f5a724e4a2146d2f2d3773929dc2d1dc4039839ae092111e9026ec7353c1

Observation 2ff12b36-d25e-415f-ab4b-1d8e48e5c6ad · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024).

EGM: Efficient Visual Grounding Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.571957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:bf3e6de5932f15a1804893e43e1888af60e4f886bebde8a20f506bc726eb4379

Observation 50020c9f-6b63-4e53-82b8-2848c2951df9 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

EGM: Efficient Visual Grounding Language Models DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:10:56.699161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:5e083f32703c07629c01ec28eda86fadeb4b13bfbf76595cf8604e150507a9e3

Observation 37434e43-ec97-4df6-bc82-056928281abe · outbound

This paper cites Proceedings of the IEEE International Con- ference on Content-Based Multimedia Indexing (CBMI) (2025).

EGM: Efficient Visual Grounding Language Models Proceedings of the IEEE International Con- ference on Content-Based Multimedia Indexing (CBMI) (2025)

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.574891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:228ecfb48358f8b49288ae25430d70367ef762be36d3597babe4bec92d253ecc

Observation f473c917-b39a-41e4-b6e5-69eecd34e16f · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024) EGM 19.

EGM: Efficient Visual Grounding Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2024) EGM 19

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.577825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:1ed9378b48df54d1dc253f87dbc24d35eafcb8f99de45e632942e3452f1554d2

Observation 7ccb2ba9-ea63-445c-9ee6-61086bb3cab9 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

EGM: Efficient Visual Grounding Language Models In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.566424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:7bf3333b94d66fafd1b681eb539c5a85380a352f72cb39e93a5d870b8ce397f5

Observation baba472e-15bf-4990-85a9-1516a91fb7fa · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

EGM: Efficient Visual Grounding Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:10:56.685410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:de75accf26c942a858d6591b2b61c6056cc2093725eee8a2e301933886d581ca

Observation cd3f772e-aa3d-4404-a03d-94a822547520 · outbound

This paper cites {question}.

EGM: Efficient Visual Grounding Language Models {question}

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.637127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:1d0a488528a60d582714da1ff3f6681b1de052195504d76768f3080e4ec0bf25

Observation 6547ba3f-a7f1-41d0-a6a7-0e4923cb688d · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.639128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:7ff6b90128b549ff119398187fee11224fd0e428ab0806791930da5e895c3c46

Observation 8561e2fc-8969-4e6b-8df1-291603d19f4b · outbound

This paper cites the second/third/fourth xxx from left/right/top/bottom.

EGM: Efficient Visual Grounding Language Models the second/third/fourth xxx from left/right/top/bottom

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.563719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:a826acacda59486d10e615632983869e4f98dac4b7f070c7bab310aee27fae25

Observation 1a4f842c-8da4-4839-858b-6b76bc1b02dd · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.628153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:152083af44b22f386c850da6eb19a2d0c63e1289e999612f228134ce48dba76d

Observation 9b543d76-e20b-4fb0-ab13-39313e2e9c80 · outbound

This paper cites the man in yellow coat.

EGM: Efficient Visual Grounding Language Models the man in yellow coat

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.630277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:56654fe588f43c72d8a9f02f879676c9d1d44afe8ebf308b285178593473ab35

Observation 55f4ddcb-abfc-4e38-8cc2-385938014c2b · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.632503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:4e26c26e4d2567e08f0d0a622ddcf93e901747aaa7f8de069d7d32854a1add41

Observation bc65e258-d994-48fa-9b7f-79a6ae81e0da · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.607436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:0161deb36cf0d8646cc04bbb8565fb817c306adbcf9356c920d5e24da31e667d

Observation dfd2ddf3-682f-4c5c-acea-c1095bbb4296 · outbound

This paper cites YES" if the description uniquely and accurately identifies the TARGET region -.

EGM: Efficient Visual Grounding Language Models YES" if the description uniquely and accurately identifies the TARGET region -

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.634900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:a00957e8d6210d7d4ef21e9999b9d4dad3a61dad984670a1cde1f313049ca74f

Observation d165cc06-722f-4a60-9f3f-a43aa7955017 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.550405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:861c435db8b71bc9b3c60bec59a75b931490eb0c60870f96e08c112296b4a7a3

Observation eee4c9f9-958f-4392-a0cf-4f7d002a1a8d · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.552967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:80f02a520349111fba54a0786c7d7dcb176cb74b294230c7b551f8b368067b4f

Observation 57899d1d-061d-4301-93af-3c2633546d67 · outbound

This paper cites an unresolved cited work.

EGM: Efficient Visual Grounding Language Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-05-16T13:10:57.540048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:73222c7eaee00e56489e7c76d70db77aa50f14664f252cd0905c4e1e7aba7066

Observation 2021a286-609b-48c4-8ad6-d0ee04cf3821 · outbound

This paper cites slightly.

EGM: Efficient Visual Grounding Language Models slightly

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.542633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:ec7f9064db85b3f170acbb1a914e4fdf868bcf17bca9999439b7f2fe8901b844

Observation 4c51ddc9-50be-4398-a342-23277cb24cd0 · outbound

This paper cites sofa against the wall.

EGM: Efficient Visual Grounding Language Models sofa against the wall

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T13:10:57.537439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:07:55.655699Z digest=sha256:c3023b34cc5971f051c91005e7b477ad4f78e1f50b780bd0b24670b50ffea5ce

Pith citing papers

Observation 69f25467-88b4-4b13-bac1-282fa739d11b · inbound

Reasoning-Guided Part-Level Visual Grounding via Reinforcement Learning cites this paper.

Reasoning-Guided Part-Level Visual Grounding via Reinforcement Learning EGM: Efficient Visual Grounding Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T23:38:51.488742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:38:51.488742Z digest=sha256:282118e1672abe54358d254bf2c3cab271ca4aac5e9bccc1aece1fb9b8067800