Pith. sign in

Paper Citation Record · LEDGER

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?

As of 19 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.10415.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10415 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:36:57.791208Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 31c65fd3-40c8-4b8d-858a-08d5bda3d075 · outbound

This paper cites online" 'onlinestring :=.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.842851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.842851Z digest=sha256:e276b3ab42bbe794a9b324489ad3bb896ad43d55491713baea4e7039cae0d431

Observation b5e898a2-cd89-4029-b420-8c86610052cf · outbound

This paper cites write newline.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.910445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.910445Z digest=sha256:f3f63b71955fb4ba878c9e47bf9c4c09c03a8af1eeab54740cf6f919c552c213

Observation 90da2a9b-2ad0-4a37-81c4-fcba53e4aa8d · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.968362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.968362Z digest=sha256:cb4951e8b33d108c49022bcf632aa318c48fbe453575d6caeffe5c78de8a4646

Observation aa144632-b19d-4ee2-892e-12eaceb40209 · outbound

This paper cites GPT-4 Technical Report.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? GPT-4 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.005935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.005935Z digest=sha256:e8a2c0cbccf671d950ed5761cd50e65e8a947f78219755b9e35ae53b3b5cbee9

Observation d384cc06-f01b-4501-9b84-265165211a64 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.093486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.093486Z digest=sha256:32b968a3364629d3b6d69dd2ca826fd3ffb95b8b5bac8d820362a46fd928c504

Observation 26f63b4d-6496-40e8-924d-82df3d97b04a · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T04:36:57.849183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.141937Z digest=sha256:e3602ce315fecf91c7073873f338fa1903eccbd0de970ec089f9a42e285940d9

Observation a4ed913e-e0ba-4396-8743-f4378946d707 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.306506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.185800Z digest=sha256:567445e5efbe5cc571cb8f07e454133c4f86a84e793c7c8f83ff875fbc808257

Observation 19f9dbbe-8ed9-411c-bb04-0d4f57fc7fb6 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.228030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.228030Z digest=sha256:28702c3ff7cdc69e382ae8220c6df0192771349236d5e790eab356556c267759

Observation fb374256-0dd1-4b4d-a3ca-c07e0744ebe3 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.290223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.267735Z digest=sha256:9c677242aa9cb793c4cfba3da447fb7184c068c2dc4c309c32492b6cfb9ae5ff

Observation d4690e8f-a6ff-4b89-af78-e6dd86db56b8 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.303597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.303597Z digest=sha256:e453983a6f0a6d689f3fa00b71df1b9520107796eb0fa9dc63bc9465de6fd9a5

Observation 7166f125-6432-4892-9b2b-cf08503f23c5 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.352974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.352974Z digest=sha256:f6b292178aa27ff3248209d6d86544396f38d4f2dd05e695cc63ddc2685581cc

Observation 3d591363-7b61-4c01-b648-91b2e7daa879 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.418189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.418189Z digest=sha256:750486a5197a52fb0e927d9646ed813dcb058216800a45c75840db5c8bf99a7f

Observation 9a505477-76a1-41a3-bd06-4a82bd59794f · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.274288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.484181Z digest=sha256:225a0574ebcf3634025c7606d44644ddea5cc17ff5069351f9bb001374dd690b

Observation f6ca50f9-4f86-4913-9064-40a506cc1344 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.265367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.552826Z digest=sha256:1e8ac1bb09413854edf5be0d8700893ae2961639218998dd6bb791f63dfee9b8

Observation 554cc7f2-cfbf-4f4f-a6e6-a6711c45b3ac · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.586973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.586973Z digest=sha256:13e305d125930da0a7c641cb8cb8cc9b12d68a429516c2ac91b6c5a7320dd405

Observation 851cadbf-3692-4203-9ab4-74d172aada91 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.251177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.623691Z digest=sha256:f9d4a0b978831f0e81751b0044523d5bb1b7968ec56c861dab99d50a8a5df6eb

Observation 6357a1b2-4a06-4e72-ae96-19caa511c722 · outbound

This paper cites Ku, Qian Liu, and Wenhu Chen.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Ku, Qian Liu, and Wenhu Chen

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:36:58.241963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.689422Z digest=sha256:7275634eff4e0b9261da8a1d2db6ba58eff582d681c702b872927e81891f300d

Observation 61e8e627-bce3-4b03-b763-7090f19bff57 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.756183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.756183Z digest=sha256:8919af9817264a33eaee60a2ab9f1528dcf73b28df6b8222c87f1cb722ba5c77

Observation ee18bd61-5898-473d-b604-74072ed4d97b · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.800127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.800127Z digest=sha256:d863d6fd6d291503d3f9cb791e71b81839b1075da69457264fbf87502ff9ea73

Observation 9cbdcc6f-eed7-49c7-b3ef-b58c46889f77 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.871215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.871215Z digest=sha256:0b5323a623fee92acb93ba48c299e3367b9ef86c0d06c8106314628718af3123

Observation aec1934d-55f9-4823-b709-84fa00b75b98 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.903250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.903250Z digest=sha256:318af8b1f9435c7c8bff165b6b2135f4dee953079577eab3c9d0185044520785

Observation b3b66b49-a8e3-45ae-99c4-5c84ace87e7b · outbound

This paper cites A Survey on Benchmarks of Multimodal Large Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? A Survey on Benchmarks of Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.960671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.960671Z digest=sha256:4486d645b09f52177b33b9d241895f791d4988cd7546f6b858dd793d77f07495

Observation 38f2f181-a075-4453-b6d4-6c11922694ef · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.221435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.035431Z digest=sha256:1227038d2214fdc4a2fded67444e0cde9d2a5a45cf2d51fea28211ab95e74059

Observation 84c2d7c2-8b62-4e08-8ebc-ea82b2ad1476 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.212845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.083977Z digest=sha256:02820138ba2199abceef9b504a37a23a27ef5c34b4c0843f04c7333358e6245f

Observation b565f4d8-0151-4476-b5da-228d5d773e7e · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.203854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.178525Z digest=sha256:45e2f6a34a994df245c66b1a6488fd9fc789f625b627d180d4142522a8feb82c

Observation 2382cdb4-0975-4350-b76c-b6060a36dc1b · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.195026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.223584Z digest=sha256:8b4b4ae2fb832b6d06fb02eb98c34b2534fb487a59e7ff2414d2b10b1ee69474

Observation a3a0c202-9922-420b-9427-584464a29986 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.184983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.312487Z digest=sha256:42fafb3b0a4a438a2c5405d4aeb65d468a849a3e8f98c8f962a9cc13b6730272

Observation 2a5ac443-47a1-4cc5-9b78-d33bc962200b · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.175717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.355871Z digest=sha256:acb998f34f4e7e0c1abeb2ea19398bc356d2735d6d8670b1f642b4379a583ede

Observation 9f40534c-5146-4bc6-bf49-8f471341bfa3 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.165973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.406025Z digest=sha256:298204c9467cfd8bf8e80060d1d2a9f910b69f366d9c78dbba525bcfe970d7e0

Observation a779e404-8c5e-4802-935e-36120f0f9bea · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.156334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.493936Z digest=sha256:a20041626b4f9ef5a12ab24dfa2ef9e7d108756d18b28b0b8f870d0604bf3791

Observation cfa1177c-7f38-4a46-a26e-1c46c67e3bdd · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.146711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.529318Z digest=sha256:de309a515d9ca52b79ab78dc3f7d74ed4c29f8e9e83b8957c232dcd9a06f5514

Observation 6f3ec60e-cef3-4ab6-ac1b-2abc939dcfe2 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.614358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.614358Z digest=sha256:f27eaf869e24ecd46afef38c8e1d52325365766f558e1e204d3239d7cf5a0e8d

Observation d5b9b921-0af2-411c-99ba-ff972055fb86 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.130607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.660424Z digest=sha256:a9c0c8d6f50b7c6a34fc3ec97a8988ef5fcbc3016dd4be310457cec3887f7d37

Observation fb57719b-1494-4abc-8166-e03811368756 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.121927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.724102Z digest=sha256:3ecc909a38a0537fc98463e804e92594cfcdb2ecdb44e93ca9057937c070267c

Observation 315654a7-0260-4135-a85d-a0d4189a2a95 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.113028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.745924Z digest=sha256:1fd278796db54587e79365b18e8f0249bf8f6a92234946e2954a550e27f21cf6

Observation 75cecedd-f28b-458e-8c73-a1db0c65008d · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.103796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.749772Z digest=sha256:6a0fd396096ce45cc04fa864171513682451200a2c27bbce0cb48a82860c7fdb

Observation 7e92028f-d1cf-455d-b143-4a5abba33883 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaMA: Open and Efficient Foundation Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.752458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.752458Z digest=sha256:e67167fb36c1a6192ef3061aa32db4452e2edea866b80d333549e2442ffc75bf

Observation dca4467a-b6e6-4873-a438-5a406bf2eba0 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.094986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.755306Z digest=sha256:d07854999ad43fb7259fe7f56b79c61cd021c4964c04ee0607f51a2bafbf460c

Observation 7a6459ba-220b-4ae5-9fed-9da86d444cbf · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.758726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.758726Z digest=sha256:bdea2e9022fe211c93dba3a20705ae8221f6659969ecddab35f98eac8dec9031

Observation 1d00e8dc-26b0-4a19-b456-da1fac3709de · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.761786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.761786Z digest=sha256:e7f40305a58bed72018b043181e960b15e82b9ad2ac39b1e4b0c3c27f9d690e6

Observation a7b68da8-5e05-4361-96a0-b382231b5f3a · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.764923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.764923Z digest=sha256:283fd2979008a4993b7eadb774790eda08946bcacf8bc978d99ad4a130bea768

Observation 60289f6a-c095-4d0f-8598-7e7a8d5c3156 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.768267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.768267Z digest=sha256:38a9c93bbe3acabfcb8c9069d290fa433507e1781f8e117029ce4ed80a479e91

Observation ea391491-2233-462b-99df-43c0d61d2dd6 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.078067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.771470Z digest=sha256:3552c9ca9d844cbea15593f30978e9f589b492da0a5ecff01a4077225fee95f1

Observation ae19ed72-386f-4706-9dd8-fc22b2680059 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.774458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.774458Z digest=sha256:4799674a94601c02e5ee491cf51f1d7435604c16d62c4c0c78e6457cd0bbebac

Observation 48c77cfe-290f-4270-adb6-2b1adcf3f0d6 · outbound

This paper cites Long Context Transfer from Language to Vision.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Long Context Transfer from Language to Vision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.778449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.778449Z digest=sha256:04017e3f2c1d262e0fc7ea125c6def761e8d5524d7f27f283cf99dc3f979afe5

Observation 6e1a7f37-a7c4-4c2d-9e09-8187eea79bce · outbound

This paper cites Weinberger, and Yoav Artzi.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Weinberger, and Yoav Artzi

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.782456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.782456Z digest=sha256:4e70a259c456ea4d0698961bf37c3156fed1c8b1728158d41b29853b1070d47f

Observation ac596079-51f6-453a-9ae1-83a32905c794 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.785475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.785475Z digest=sha256:a1226e8f7e7ace6d2441e6e969e0ec2f00a31cb3b8dc9d03172b7f1ea80f57c7

Observation 7632a9aa-e197-41da-a670-197147fc38e5 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.050574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.788190Z digest=sha256:ca4800d96f505b5a3d326da0a9d8454af97026c9889186002fe5239edf5c2e5d

Observation a3890a0f-9ccd-4788-a921-54571984658d · outbound

This paper cites Zwaan and Gabriel A.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Zwaan and Gabriel A

Reference 49

Resolution
verified exact
raw_fallback, observed 2026-08-07T04:36:57.954724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.791208Z digest=sha256:f4c5f70b92814b7723fd11bcf66664917192c43d74ee762393dd7aa40c7dcf75

Pith citing papers

No inbound Pith citation observations are available.