Pith. sign in

Paper Citation Record · LEDGER

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering

As of 18 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2608.01660.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01660 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:08:38.986063Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

75 of 75 outbound references displayed

  • verified exact1
  • verified fuzzy28
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e07367e-ad2f-4286-8965-4b457e004f4d · outbound

This paper cites European Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering European Conference on Computer Vision , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.715296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.137288Z digest=sha256:ee11486d822c5dd6b6614db19dde6f33dd785a2deb9c1f1d19886d84cf5afe15

Observation f80d940b-fb5f-4d60-b764-3c00b82a9c30 · outbound

This paper cites 2023 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2023 , eprint=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.236299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.236299Z digest=sha256:eabdb4d80776ea5c223d6ad9e272b36c16e7b8fdff2261860e12970c0f486348

Observation f1d5d609-1817-430d-b93d-2502aa81b6c0 · outbound

This paper cites 2025 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.699847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.240854Z digest=sha256:24833d25bc485129f7c85e9a405ac662888e85b2a29c698829e491f926ed796d

Observation 4c8f8d32-5713-481d-8617-51a6e8426dcb · outbound

This paper cites MovieChat+: Question-Aware Sparse Memory for Long Video Question Answering , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering MovieChat+: Question-Aware Sparse Memory for Long Video Question Answering , year=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.688778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.244355Z digest=sha256:9d73fd171eda0d7aa6c3645d54d361937c5b7759689f5b56c4291917bebd616a

Observation 30e998dd-7757-403b-a696-8cd1c87e8de0 · outbound

This paper cites 2026 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2026 , eprint=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.556579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.248378Z digest=sha256:e145ddc546fdb45ecf3e64949559692da462f494a6504c95897476953affc1ac

Observation f87b3c8b-b499-480b-986b-fe8454df5881 · outbound

This paper cites Proceedings of the 21st annual international ACM SIGIR conference on Research and development in information retrieval , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 21st annual international ACM SIGIR conference on Research and development in information retrieval , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.257284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.257284Z digest=sha256:974ab99241c7ff24558d7737713100838194a22b48ddcfcf981026ade6ed3584

Observation 731a362f-c9ea-41f9-a974-8dfd949df099 · outbound

This paper cites Classification Problem Solving.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Classification Problem Solving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.369466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.369466Z digest=sha256:ca1b7d89052df127f9646df46bb201c82ce0a4bd98d2b6438294da2e15902a62

Observation 7fb517ed-fd23-488a-a795-59bd633fbe48 · outbound

This paper cites , title =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering , title =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.450358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.450358Z digest=sha256:e99c21105d5d21fc7d753a9dc95ea81de10d4cf67707ca33037f02f235e73c8c

Observation bcdd3945-3806-4e5c-8143-6df7a5e58a88 · outbound

This paper cites New Ways to Make Microcircuits Smaller---Duplicate Entry.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering New Ways to Make Microcircuits Smaller---Duplicate Entry

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.454244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.454244Z digest=sha256:7c9741290fb2fe6e09b0204c146ac28077c5ff770bcc53bbb86c2d932e08ec61

Observation 647134e6-4879-4caf-8f01-584f35045f15 · outbound

This paper cites Clancey and Glenn Rennels , abstract =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Clancey and Glenn Rennels , abstract =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.469414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.469414Z digest=sha256:c5a8d874f45d3f0ca54f794281f1ceb77330430a8ae07a8d793f78a065f6eae3

Observation 03e553fe-709a-4736-9ed4-5f3d6906e7b2 · outbound

This paper cites and Rennels, Glenn R.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering and Rennels, Glenn R

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.518806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.518806Z digest=sha256:e1aee584aabd59cdb62c7bf5834f819cc0fb0e19f33871bca5bda7c5f28ceb8a

Observation 4eb3de7b-adc4-43bf-bfea-b75d9ec7ff45 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.516542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.537659Z digest=sha256:f361339d62b35f0a5c26bd32cfe77078575ce5205b03fac08acd573a5f95ed69

Observation 08d5c6c0-416a-457c-a5bd-ce61436bbdd4 · outbound

This paper cites Poligon: A System for Parallel Problem Solving.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Poligon: A System for Parallel Problem Solving

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.541546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.541546Z digest=sha256:e97c6286c7332446dfee899e8f9500f07c7f1dd079883efd073c03d8cb570d9a

Observation 6398792e-5697-45d6-b6db-a6228abd5a73 · outbound

This paper cites Transfer of Rule-Based Expertise through a Tutorial Dialogue.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Transfer of Rule-Based Expertise through a Tutorial Dialogue

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.545441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.545441Z digest=sha256:f1bf4c234fa96fcd8f505c1eab8003fd8c8dceb1235d7dba2f97770d82677f0e

Observation bbcc461a-59a8-48bd-bf56-dc6f6ae55936 · outbound

This paper cites The Engineering of Qualitative Models.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering The Engineering of Qualitative Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.549013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.549013Z digest=sha256:f3cd787725fb1cc98bcd68047250f75e19f7cd065d60a861dec172ddb3c1d358

Observation 80a3d2a8-78a4-4a89-9fa7-e6cf7c4a5ebe · outbound

This paper cites 2023 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2023 , eprint=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.552405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.552405Z digest=sha256:4eb44cbc912f5ca0080293f110afdd314dfbf6f9541ba1ec4a2f929994acc9d7

Observation 0b77ec0d-8f7f-48de-8bd0-2a1077c22345 · outbound

This paper cites Pluto: The 'Other' Red Planet.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Pluto: The 'Other' Red Planet

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.640023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.640023Z digest=sha256:bf1d0e52a4401887871d413006184df7dc92bcd9bc8c54f3783bd4fd2f51535f

Observation 190f8748-9f09-47f7-81e2-b74a79508286 · outbound

This paper cites Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.428207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.756939Z digest=sha256:6912a9a948a2447bb31d893caea865bf964c03c8ac8dcbcad6891bec3d0dd84e

Observation 8ece195d-8ce3-4ee6-82ad-c45dc7a5738f · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.417185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.792374Z digest=sha256:8b860a754029d8fe8627d3e3eb131be67e08c24c02d9bf1c25580eba94dd600b

Observation f6b9d767-4596-4ceb-9ac4-9f106a591760 · outbound

This paper cites Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding , url =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding , url =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.287918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.799515Z digest=sha256:6ebe4e1d8ae3ace0329ff7f079c95f7a53501c128e3fd82961fec4d9aab9f7f7

Observation 5ff192d5-0538-4ac6-bea1-732cd5b13155 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.277701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.802855Z digest=sha256:f8ba304846d11171bc01ba4bd660004e43370a586fa4db9d5053bfb8f55766b3

Observation bbccdefa-fb6a-4da3-af5c-978829d5e4b2 · outbound

This paper cites 2025 IEEE/CVF International Conference on Computer Vision (ICCV) , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 IEEE/CVF International Conference on Computer Vision (ICCV) , year=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.174178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.806918Z digest=sha256:3f36de88867ae915de83ba59de9bb643e6baa48195ac3c8886b2d2d83c28049d

Observation 69f5769f-4d62-429f-b526-2fdffa94d299 · outbound

This paper cites European Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering European Conference on Computer Vision , pages=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.810421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.810421Z digest=sha256:38548e864e5079a946118bc21d73bd19e0dba5765b5f62dbe2903b546d58adea

Observation bd063ae0-a4e3-4387-ba09-77952f165a2f · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.055167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.900572Z digest=sha256:3a0670321c82aacf5c79ee82b0228654f892c793b9dac08264fbd9b264f1b55c

Observation d3dc87d4-d1b2-40a7-b602-0dd1753bf505 · outbound

This paper cites 2026 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2026 , eprint=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.043524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.949122Z digest=sha256:22d0e94cd04f05d1e0e174de73b05850d5a6daec4527f9f941cd821eb9952bc0

Observation 2b627819-3db4-45ef-b362-27c4c5a7b65a · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.926114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.953951Z digest=sha256:18100af63c41e9a528590b64a2564ed5c5f1c898347ee908d81e8234ee54f133

Observation cbe1ffc0-9a39-46c0-bac1-c287eff57ef1 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Findings , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Findings , month =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.867359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.957665Z digest=sha256:511a93575cbbdcab4ea63811709656d7b0a4cc5c0217931f37510c968bea0ba1

Observation a472e074-83f3-4c82-bca6-5a29eb6d07bf · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.857178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.960941Z digest=sha256:4a844cc14ff0b8cae414e9dba8bfced06d6f90189f3ab85b40b325f16f752b95

Observation a876bf8e-81d5-4f2f-a4b4-c4fa558bb782 · outbound

This paper cites Proceedings of the 2023 conference on empirical methods in natural language processing: system demonstrations , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 2023 conference on empirical methods in natural language processing: system demonstrations , pages=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.769058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.964397Z digest=sha256:be34c1a25f60365345d5c9bbecaeec367882e6a0a5370f5a6c3ee571f78c30d8

Observation d1136dfb-e2ca-467b-9357-a2659b9d3e15 · outbound

This paper cites Proceedings of the 2024 conference on empirical methods in natural language processing , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 2024 conference on empirical methods in natural language processing , pages=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.010486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.010486Z digest=sha256:3d455bf3edc54f4eb182ffa1e0f31298f9e865b41cc1250874e53f74b4f9b263

Observation 755be9eb-1221-46d8-83ac-6ae17c683be9 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.050499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.050499Z digest=sha256:a158fee338c1ed1ce888cd3cdc45bde1344ba8603d3a45792892b7b91323f8eb

Observation 4de4fe1b-bc01-438f-a749-415504653898 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.053955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.053955Z digest=sha256:acd31799610af5b608e03421119e999db6a48a2b8251e6602bcf1ca365e12b26

Observation c9de6a90-ad9f-4198-a57f-bb2a5ec12917 · outbound

This paper cites International Conference on Learning Representations , volume=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering International Conference on Learning Representations , volume=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.651943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.057129Z digest=sha256:dccbe49cc69439b3c4797e66b382e330147063981b5d3584c6ee1ecbd1c645d3

Observation 02a3a939-f765-4fe5-8d5d-badca8281be6 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.641702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.061226Z digest=sha256:e2e9685441173c62b40601b34b1fae133c78db07f99a71ef9ebbae56a7ed315f

Observation a95c83d8-96f4-429a-922a-26b4f8f5e33d · outbound

This paper cites European Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering European Conference on Computer Vision , pages=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.066004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.066004Z digest=sha256:16650c70635fb410092a7347a054df0ccde4532a9d0e1671a3f650fd8dec0d35

Observation e457d4a3-4834-4943-a18f-918fa659dc98 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.103769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.103769Z digest=sha256:a7949102f0aa4832901f1b38435a5e7a1ee8e276cae3d0046920d23ed9e3fcdf

Observation 29b91e4b-5a26-4ecf-a580-ac19761e7bb9 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.619177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.108522Z digest=sha256:a3f13665ac4f5cacf9a93244a637623a15dc44b72ed4a939bddf463b0463a38a

Observation 32ad0c3f-bf94-46c8-8c1f-78e37955dde7 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.554916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.112045Z digest=sha256:78205492bcd4caa39906d155b3bb77484e4a7aada6f57cd84e083a9a0eb5927f

Observation 8c46febb-1ea0-4954-9982-edb5a21eda24 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.423055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.116085Z digest=sha256:d2ccce34f67846ca79e9a3366804e9865ce1046ea73ea0a76db69e44a20019fb

Observation a6128d49-3434-457b-a0f0-1eff0c755f34 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.412791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.120013Z digest=sha256:0eff3f09176e8923938e3dbaec75238d03de16a17f5ea8e24671387f6c0fa791

Observation f622ae6d-5a07-41d3-a504-1e57066c6e82 · outbound

This paper cites Video Summarization with Long Short-Term Memory.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Video Summarization with Long Short-Term Memory

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.402464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.171862Z digest=sha256:5bef74c0dcfe2c0e415f531aa243e1acdbc05563c9aef4f67705feb505b298cb

Observation 1b339b60-a989-495d-bdd3-7c3639b04889 · outbound

This paper cites DSNet: A Flexible Detect-to-Summarize Network for Video Summarization , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering DSNet: A Flexible Detect-to-Summarize Network for Video Summarization , year=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.362731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.216920Z digest=sha256:e238b35baf5557d5584dfda3afedb90060fd20ad055c460e31c906c16b6ecdbc

Observation 4280af97-6077-4899-9eb9-a2c3e2428c90 · outbound

This paper cites IEEE Transactions on Multimedia , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering IEEE Transactions on Multimedia , year=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.236814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.220440Z digest=sha256:cd9b09d4b8f8637e48cee76f01e0dfae2c626b7d4c92a50e6d168056b0dba4c6

Observation a02cea44-4834-4011-b474-6dbaab7c855e · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Advances in Neural Information Processing Systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.224302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.224302Z digest=sha256:8cd7b65e2b3aa47e4a60909e63767088e11642b7bfa637954d64b590a1049bc9

Observation e4813424-0ceb-42cd-95ab-a72c4bbdb84a · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.291234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.291234Z digest=sha256:e9bc166a3e75a3aa9e7393dc16c0edeb94ac04d8c21c1ea5ed9e7e9070efe6d3

Observation 487034f9-8c7f-4ede-bad8-b5b941a54730 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.095211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.320191Z digest=sha256:102c74927b426084c301d803233021f63db9c7d2ca3143745d84054b81340414

Observation 044124bc-e69f-46e8-b97b-010ee995a53d · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.323888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.323888Z digest=sha256:040f356d0f97ce27e37024814d58f5ae188d3bea6aa06fde5828efd847e6d8b2

Observation baf2a62f-6102-4d3e-b124-538c6fcc3707 · outbound

This paper cites and Han, Rilyn and Fei-Fei, Li and Xie, Saining , title =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering and Han, Rilyn and Fei-Fei, Li and Xie, Saining , title =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.327293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.327293Z digest=sha256:a72626a8029d7f56fd2556fbbfc1fc810574b40afcf0a0e9e0ba88f16fc49187

Observation 0875204e-35ff-42b8-89fd-1d1bc0801623 · outbound

This paper cites 2025 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 , eprint=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.006036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.331591Z digest=sha256:e6e7d30643794d69106b8f910aa0ba37481197eaf2e2ffb8997a4a1f58ca7c5a

Observation 628055f0-75db-4fb1-92d2-77cb86298db7 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.370879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.370879Z digest=sha256:eb8b742d95a5e257ee9a232cdd38f335c7bcc46d9e8ab0a4c248b7703392bbb1

Observation 6ae87d8b-7a8c-4712-a36c-41c5d8e70a2e · outbound

This paper cites 2025 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 , eprint=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.420696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.420696Z digest=sha256:4478f1a5da7109dcd9f3dde6a9738f7a83340ac9f4bb28d5a026006c89b3fc4e

Observation 095ad355-d31e-42e6-8e46-9a1ac0e03a70 · outbound

This paper cites Qwen2.5-VL Technical Report.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Qwen2.5-VL Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.425523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.425523Z digest=sha256:637b8d64c53161401297b88495aedc7110a9f097202eae99b1ffef6346697091

Observation 4ec6e2ab-553a-45dc-aef0-18518d1d3ca6 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.430210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.430210Z digest=sha256:48d56da20fd41f8b27b53ec37b71305c88237beb8c6d7d2a6e4a5bfe54559695

Observation 6b71a19b-365f-48a4-9de3-854f69669728 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.942750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.432835Z digest=sha256:e3fed351446a8f4f1d2caaea3ecebd28068ced8e80816074698fb9fa1896d873

Observation fb6d0876-1523-4d31-b80d-414336e16598 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.859460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.498126Z digest=sha256:69734b0a387e0a5f4540e1ada8ce7d058799da26bb2a2519adf6091920f7d766

Observation 5eae232a-d8a5-48bc-9ebf-1a15e34654de · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.501680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.501680Z digest=sha256:1db41dafa37bfaef55b9bf30f0dd1f06d542337f5f4ce3eafee925b366ab6baf

Observation b382ae23-13fc-4e43-b1bb-82ef6d08c681 · outbound

This paper cites Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.506143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.506143Z digest=sha256:ceff9fcebdbec680e88b17069586004eb7ae3a6a05bf1d4a1c0a880b18f4e4c8

Observation d12aa015-26e6-4572-b517-efb079a0cbdb · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.843117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.510086Z digest=sha256:383d10489fa5938651e52bd068cd345b10ac081fa29fda40df67bbad188cb741

Observation 381cf954-3567-46fb-8117-99ec00fdf079 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering LLaVA-OneVision: Easy Visual Task Transfer

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.513658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.513658Z digest=sha256:1c78080c9eb08ea66cb1357c3db754776520fa22cdb1652e17975f95e90af082

Observation e2e61142-1020-484c-b388-4ae9353500c8 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.517383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.517383Z digest=sha256:5657a714392c138e6b8f9af628d43de05f7e34dd5ce4824414b0168fb6170860

Observation 21a33fce-14f1-4418-9798-8c8ba66c098f · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.751712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.618037Z digest=sha256:c2db39cff4e557670e1ac9984edd8921bb2583f1e730cb9cd19ca398cc9e2693

Observation 36d64ae1-cb5a-4418-aca5-3f300da8c1d0 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.741430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.621710Z digest=sha256:38a1d1a714e6fa0eee347b9ecf8b04ef2ab8636d7a22547d3846c378c7b22dad

Observation fbb3f12b-3416-48e5-962a-50cfee2e3f3a · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.673024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.625207Z digest=sha256:1b4825ce3db152a4e76a78cf859b5eefcb543671975ed831a4f2b1ae70816c90

Observation d24ab680-ee46-440c-8749-12b49c7398c2 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.662222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.630154Z digest=sha256:46babde2a650220bb82fad2d54b1165cca4512cdeb3a92833a1b292935438490

Observation 0ae16a87-02bb-40df-bd00-4ed4458b3905 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.652235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.634381Z digest=sha256:21244a78cba54ddee31ed1127b0c416c7aa81a856ffe731053bbd32fda4e8582

Observation 4fcb77a8-dc9d-4c0f-9cdc-768e84906422 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.600497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.744451Z digest=sha256:fd3e2b0a8997677c301cd01a16e53d5218df0926fb9849ce641e0318198851b1

Observation 88791888-b073-4630-b756-7d033096f2b4 · outbound

This paper cites Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.813415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.813415Z digest=sha256:f73748a640221464af212e9a6cad5097397296b2b33ae7195ead3f139f257142

Observation 8677b137-e454-49a3-8fc4-0ba484fe82ad · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.590616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.818920Z digest=sha256:abae3731d0a698de7b00792980d6192d6d4e76f4209eb6d7fe47959e06f6e1de

Observation a6bff458-a54b-4c55-bb7d-d4cb1917c3c3 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.553321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.822766Z digest=sha256:f8fbc83908c290aa65c5fbc3fcf4dc55370752f7f352d240a4191e1a805a42ce

Observation c171ed28-6573-459d-8bb4-8da2e75989bb · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.827139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.827139Z digest=sha256:d72d12e318dec218004f176590bcab666fc6dd37b08b75178348abf28fe971cd

Observation e70afab0-02a9-4ba5-bbf9-a8de4aaf6eae · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.491721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.906682Z digest=sha256:9b68ccd7122a091fa5509f3b1a7ec934ec6d387256991528a404f53df1006cc0

Observation 90499b80-f861-415c-bcd5-1beff92be612 · outbound

This paper cites MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-15T15:08:39.183218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.973063Z digest=sha256:bbbe572099461776ec38129955beda512eb1659795cd897d7f5f8b9eda9b3b46

Observation 29d4acbd-38bb-4148-8f5b-cd22878cb88c · outbound

This paper cites C.; Adeli, E.; Li, F.-F.; Wu, J.; and Li, M.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering C.; Adeli, E.; Li, F.-F.; Wu, J.; and Li, M

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:39.395109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.978117Z digest=sha256:827699452d3da3b707e4893d3457264c0ecc40affa38a78e21676b61fadcc849

Observation b4d4bf7c-f593-4ed1-9dac-15e6d61e0d9c · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Sigmoid Loss for Language Image Pre-Training

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.981773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.981773Z digest=sha256:e2e37a00e54cac0b4bccd063056677896c81e8a453f1ca61055c0dcc654dd067

Observation 4c9df4ee-33a4-48cc-8ebb-15aa8ab5d755 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.986063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.986063Z digest=sha256:2e564f5609b520656ae7f498bb592bc6920f908c872cc425f76c97d8a313030c

Pith citing papers

No inbound Pith citation observations are available.