Pith. sign in

Paper Citation Record · LEDGER

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models

As of 18 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2505.21465.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21465 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:33:58.406188Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:09:38.885464Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T22:09:39.028658Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 31bf5866-4760-4335-a224-51805862b122 · outbound

This paper cites GPT-4 Technical Report.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:52.786550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:52.786550Z digest=sha256:b7683289c3045ced0500d60a04a67cf1de807807ce6f9f97632e992638bc7dd1

Observation 65df0608-8b64-422c-ac3a-5a0cc8a74c1e · outbound

This paper cites A Survey of Multimodal Large Language Model from A Data-centric Perspective.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models A Survey of Multimodal Large Language Model from A Data-centric Perspective

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:52.897102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:52.897102Z digest=sha256:aaadfb12285f8314746f0895aa4f3d98cf02887d59455254d499e154d28f81c5

Observation fc049b0d-8b0a-4c77-a607-63bbce204588 · outbound

This paper cites Round and Round We Go! What makes Rotary Positional Encodings useful?.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Round and Round We Go! What makes Rotary Positional Encodings useful?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:52.995504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:52.995504Z digest=sha256:bc50ad54fe013a9f95b5e60b8db76c21dd90d51294c31ea739e530d519aaaabb

Observation 021a6a3b-4a9c-49b5-96ff-7773d022ab89 · outbound

This paper cites InternLM2 Technical Report.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models InternLM2 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:53.142625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:53.142625Z digest=sha256:bfdf7c5d5d1746addae0d01158d14dda848d706edd635a50749171ebfd9f71d2

Observation 961d16dc-35d7-43a3-b59f-4eed278cd3b9 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:01.204957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:53.242770Z digest=sha256:317cbf56aed85fbfad091c4da29176da9d06bbfe207da6d0fb76324c378ce6a9

Observation 174bece8-caae-4494-8294-58f95ae0b1b0 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:53.326198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:53.326198Z digest=sha256:72d96b0d5d3b5df81003c734fa45953250651de99f480b3264ff9eaa396db6a4

Observation 5cb000da-6469-444d-9edb-ffa4d03a386a · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:53.419678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:53.419678Z digest=sha256:d621f9a3b41930f2a61e89a8a931d76886c6153a26ec212fdf4fc35d4acd49bb

Observation 7ff211dd-c979-46bd-af70-b6c8130a3f72 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 8

Resolution
verified exact
doi, observed 2026-08-07T13:33:58.682183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:53.554561Z digest=sha256:48601cfd985f43cb691fd0818fbda4ca0a4ff626a1391bb7ead42ac469f03002

Observation 22035ef3-ee65-4c95-a16f-ec251a2eaf41 · outbound

This paper cites Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:33:59.270505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:53.715127Z digest=sha256:45975876d946e526b10bc46b501f61a822d671356a384831f0d9c526fab6597d

Observation 794782d4-e9a5-40f2-8394-4d03513b5c12 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:53.834178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:53.834178Z digest=sha256:2dca5840b41d6b70e9835af0375c6885b17372d46c8301ebbac96bec96332523

Observation 2b52ef69-0d0e-47b8-b6c0-63a838a325d0 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:01.069569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:53.989655Z digest=sha256:35020743f04c946dd4d85c2d887d78160a4b6046979ee8a7a59316160b424a86

Observation f4daa6b2-6f6b-455a-8c49-6dec666b2e9c · outbound

This paper cites NVLM: Open Frontier-Class Multimodal LLMs.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models NVLM: Open Frontier-Class Multimodal LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.127078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.127078Z digest=sha256:c20959a34d73df10d4821ce1852d0161543666ba3f2bfe635f6a577af7cf1679

Observation 90e0e22b-6ecf-4464-8d9b-cd5f8643d0c1 · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.254409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.254409Z digest=sha256:86ceb6fa269eda3c7a2980792613f8ab2c1a43bab5f2f1a14e1f394f2a2da38b

Observation 53af61fd-b775-44fa-8ba1-1056032c9cb8 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.385147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.385147Z digest=sha256:d0515caa8de06c2f789a69a2c636e4ba544391dc76008ed6b5ccec272f060c55

Observation d0d9a0b4-b9ff-4e5a-b609-662b45996130 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.509152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.509152Z digest=sha256:422bfd5ef48b4889ec503b75eaa0577b06cd571325ff099f4a6ec148bd9524a9

Observation 126ea3e5-301e-4a5a-a734-aa477467628f · outbound

This paper cites V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models V2PE: Improving Multimodal Long-Context Capability of Vision-Language Models with Variable Visual Position Encoding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.675516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.675516Z digest=sha256:ca6739bd2dc38c329e46ee8c18b61d71ef494f784c49ed0f8670ee9615bd019f

Observation 9868fcb6-494a-47c1-9ca2-d2b203075d4c · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.796540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.796540Z digest=sha256:7c08973bd6951e74a4448d509023eec554ceccaf3a9e4d8dc013e345f5ce92bb

Observation 50dcb0ac-92a5-49ec-a5a1-fa2b4735cbe7 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:54.890272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:54.890272Z digest=sha256:a8edd6f0dc851fd4e1827a8386eaf4e010c67df22bdad04e7af4d7649405dc80

Observation b8834c94-9199-485b-9a0a-00b0b974100f · outbound

This paper cites DeBERTa: Decoding-enhanced BERT with Disentangled Attention.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models DeBERTa: Decoding-enhanced BERT with Disentangled Attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:55.005545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:55.005545Z digest=sha256:f90adf57295c0c5b036b346d873c3e4463cc30d28b3cb16b3d553097bead425e

Observation 0d5c2ac6-77ed-4dba-9adb-c0bfd758da57 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:00.882532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:55.157273Z digest=sha256:19eb4b44764b636fd49ae1f6dda4d0ee0f26b2a35fb46be375d28fe0ef072cf6

Observation 8ebc8c25-c5e4-4cd8-9ae5-9f2d8400e533 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:55.226798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:55.226798Z digest=sha256:6f37649f4462374515107e10453a4989a39f4499e732ac5c52fa18a1bd03ec46

Observation 1232fd8c-ece2-40af-b83d-2d200ab8e011 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:55.328954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:55.328954Z digest=sha256:5733d673c18b8a109f98adee41edee9b736d1cd38e7b78caf89b95d3c0d0150f

Observation e18ee29b-d0d0-4e2d-994f-e6dce1cf75d7 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:00.716905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:55.428535Z digest=sha256:50ea73f92af94c4f7ac0a44899780754e9d6a594d7ee459e4d9478fce3556e7b

Observation 6953a941-df0d-4d58-be17-63306f8c0ef2 · outbound

This paper cites DeepSeek-V3 Technical Report.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:55.564186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:55.564186Z digest=sha256:d9f687e5402dbedbf482cd56d185b269e06a67e6e386ba70a52f2bb17af2020d

Observation 4994c686-e20f-4cb0-b24e-967bbfac862f · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:55.651336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:55.651336Z digest=sha256:cb361e6bd66f93a89a67f0eb1bcdcc2ba200f3845ad92c40977804eab9930490

Observation c992774b-f4cf-410d-9d96-80ca3fb70229 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:00.513308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:55.739456Z digest=sha256:abb4c2134753b754791a0537184d0891371ec66208fde32cf6870c16934d2a20

Observation 22a5d2cb-145c-43e4-8c35-8f6c0604747f · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:00.269860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:55.862527Z digest=sha256:8c2701ef5837aa836fd4111b318da86ca6e82b632489aff14e6e54fbfbf31a1f

Observation 1d0f732c-c7f3-4754-9dc9-fb37f3374e46 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:34:00.104746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:55.952443Z digest=sha256:eb5165d0a897f4a0d801015e6ade64fcb7ce98213b592eb1aa98b3db1a8da459

Observation 9725a4c1-a970-4d44-bbc9-282a7e4deba2 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.040259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.040259Z digest=sha256:f5a2703211b080a2d90977743d734a592b2c9227c87416bfa71f7db2ea58e88c

Observation 8a9d3f84-1937-4239-94cd-d89cce996e5c · outbound

This paper cites NVILA: Efficient Frontier Visual Language Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models NVILA: Efficient Frontier Visual Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.166173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.166173Z digest=sha256:12877820d99d45a82b616394200d33ca32b176d0cd7e851c0627ed6055d0824c

Observation f012f968-a984-4c1c-9cdd-b8780337ffcf · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.275413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.275413Z digest=sha256:33c7f78d144ebee199aae9e881c7f97c6dcfec1fa8d9a894857cdd015afaeb2a

Observation 7ff1ff27-44de-442f-9c80-195928c4474d · outbound

This paper cites Base of RoPE Bounds Context Length.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Base of RoPE Bounds Context Length

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.359481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.359481Z digest=sha256:34f2fb549f473d37235a66d6b81e2730ae2cc6117c90c000efa5c43124ea21af

Observation 73151cc9-449e-4f9f-8947-f5943b8cd198 · outbound

This paper cites Pointer Sentinel Mixture Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Pointer Sentinel Mixture Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.458542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.458542Z digest=sha256:2d95ffe18389f5f761b3096894477a8f146738823b8f93a217e15deb2ad2c255

Observation 5bb636ed-8ee6-41e9-bf7e-aa92eeab8708 · outbound

This paper cites Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.579316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.579316Z digest=sha256:f12f65edd60b1ea4340f3f2161b687b14efdbd5b80724e89b9ba28ab4f59a595

Observation 76bd9e0c-de80-42c6-abdd-ec2d9898b530 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.683044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.683044Z digest=sha256:24bbc6d2dbaca540c1729040326a5f64a1f7e578c89079d37233f31e3cba83b0

Observation 74b635af-12a3-4abc-afa5-1ba2b08c12eb · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.796996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.796996Z digest=sha256:39c1914620042b0822b18cb79ed550f7463c516f4cb6265045a1e40569305d92

Observation 98a31dd4-2a28-4754-b24a-6bb8921f0b82 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:33:59.901224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:56.875340Z digest=sha256:a0aee0af4b24b45b71c423a76d501a02c1b041e0c703717e53bd87d8e0764fcf

Observation c09f3dba-9868-4a6d-9c0e-151ee9e37123 · outbound

This paper cites Self-Attention with Relative Position Representations.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Self-Attention with Relative Position Representations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:56.986384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:56.986384Z digest=sha256:29977c039bb61c07bc5dd310109af3b6e3eacdf44af179577fa48fbf299a264a

Observation b8d2cb3e-66ec-4823-bd9d-b18205b1182f · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.100720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.100720Z digest=sha256:61f5b7c3b068f693f16fb4b6ce120f313177aa6f843dc4384450d190010ce4da

Observation 915f3e5c-98f6-4ac8-b3b6-e2750c6180a4 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.210002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.210002Z digest=sha256:4a7866e36c7b6b94d0b818d53d3ea1d1d3de4802a14c9140486fa23d6569d0f0

Observation ca325593-33c2-4ad6-a7fe-b8674347ae2c · outbound

This paper cites Encoding word order in complex embeddings.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Encoding word order in complex embeddings

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.328427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.328427Z digest=sha256:d244fb0f9db5255d43ea546bf6ed108d5dc94874f0f00dd8844d59ff2bc79eaf

Observation 7ddb8249-8866-4779-a1bc-eb51bde00052 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.417393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.417393Z digest=sha256:dee794c6e9486ea4d159fdb0b60d8c1bd77021dc0f9e7317be9571db971a1f91

Observation 42562388-ac79-4dbe-8a31-3e06cbd2b6c3 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.535327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.535327Z digest=sha256:e1f9018d024e721877c78ee0282e208b8dc0e73696343ede7665a2cf64dbfb51

Observation 797e8bae-a8b1-4e21-8e8b-7ff760c8b926 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:33:59.666576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:33:57.644912Z digest=sha256:3dcba34c03faa177ab73c1a124c795d7ba6e343e5a7cdbc616519ce04bd1cfeb

Observation 63015977-dede-45fa-8d58-f80bfd172548 · outbound

This paper cites Qwen2.5 Technical Report.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Qwen2.5 Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.728305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.728305Z digest=sha256:2fa7bb5df13bd98c9496248adc4081b90b2cfffaea77b9ce3f941ba9c05b9b47

Observation 0ce4d2f0-3b53-42a2-bb52-07096d722b8e · outbound

This paper cites A Survey on Multimodal Large Language Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models A Survey on Multimodal Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.813742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.813742Z digest=sha256:33cd31881194f0b2e4262367fc098ab3c3ad5d755097ca9f1a19839227cc76a2

Observation e7efaf5c-7ae2-4df8-be04-c6982fc11e52 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.904413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.904413Z digest=sha256:fa6999e7d8be21a208f26612b771ae776137778c51f87a2651c44320a6339580

Observation 8ecc0117-d171-4197-a078-51dff60f6ff4 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:57.971570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:57.971570Z digest=sha256:d37f5cd83ba9eb96275681a367802e9235e706be12088b856e2cdd25529b18c8

Observation b4a2bfdf-d0be-4e4b-bf84-6059ee7d55e2 · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:58.047950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:58.047950Z digest=sha256:5afbda89a1420f2fe590ac8b78cceae6e06b1471e617aef26e59779fdf20be09

Observation bd964784-c1d4-40fd-bb41-cc1c275e4637 · outbound

This paper cites an unresolved cited work.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:58.127521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:58.127521Z digest=sha256:9262d8d3e5690fbc460ce8f10d465da31ba00409bca1f68e9880df4fe1b8694d

Observation 281c2a70-5c47-4125-9504-d3bffeb7b64b · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:58.230319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:58.230319Z digest=sha256:5e85ac6feaa5fd60ace7eafb32ee4632679d959b0508ac16eded990fffb72f3a

Observation faff2384-d844-4751-aad3-a997392df084 · outbound

This paper cites online" 'onlinestring :=.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models online" 'onlinestring :=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:58.305535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:58.305535Z digest=sha256:81f6134590047a130231b2d8cee90e35f14eebbd794f51fbafa92ebf7d599a2b

Observation b9a86f81-c671-4285-822f-361704ed29fc · outbound

This paper cites write newline.

ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models write newline

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:33:58.406188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:33:58.406188Z digest=sha256:242261f897d9d678677b4021db04f9f941d9a1bbb33c0fd54326e5c38031231d

Pith citing papers

Observation 92ac655a-3001-48ca-8190-6244a8bddd14 · inbound

Not All Visual Tokens Are Equally Safe to Remove:Consequence-Sensitive Visual Token Compression cites this paper.

Not All Visual Tokens Are Equally Safe to Remove:Consequence-Sensitive Visual Token Compression ID-Align: RoPE-Conscious Position Remapping for Dynamic High-Resolution Adaptation in Vision-Language Models

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-08-11T22:09:39.037088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T22:09:38.885464Z digest=sha256:c4a1a24748f3bdb52445d3e2fb271346bbb2d5276f0d028bd863aabc9a7e334a