Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T12:54:56.474300Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 0 inbound Pith citation observations for arXiv:2502.07411.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T12:54:56.474300Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
87 of 87 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 20a166fa-0211-4747-8913-f525b65029fa · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70b606bd-fcfd-4e71-ba3c-c5b514aa235a · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Where did i leave my keys? - episodic-memory-based question answering on egocentric videos
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12c78ca8-94ce-4a70-bc4c-046db810186f · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Scene text visual question answering
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 701d60ec-cf0d-4114-9199-9c1eea17b542 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Nougat: Neural Optical Understanding for Academic Documents
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b7ebe36-a71a-44df-b74c-9fb4f39b937c · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering InternLM2 Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a3ad532-9edd-4f91-a7da-72196dccc641 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering ShareGPT4Video: Improving Video Understanding and Generation with Better Captions
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adb75ff9-37fb-458b-a65c-3d2cbbcc3648 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9661cac-6181-4b7a-bee4-f2a91c45d520 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b68f2c0-dcc7-4251-8541-cfa9be868bbd · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b98d8e6-29cf-4455-bb5c-51f67345d832 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Egothink: Evalu- ating first-person perspective thinking capability of vision- language models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdef0dc8-631b-43ab-86e0-740dde11ec80 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Grounded question-answering in long egocentric videos
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 060f0cb2-c727-4ea5-a003-85cc2ce94246 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Egovqa-an egocentric video question answer- ing benchmark dataset
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 320ac191-1522-4b26-8062-0dac98c0c302 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a1f8428-c2b7-4097-9d5c-e43294e33dfe · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Ego4d: Around the world in 3,000 hours of egocentric video
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd64bf27-6376-4f56-853d-a0148fe9ab1b · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Context-aware graph inference with knowledge distillation for visual dialog.IEEE TPAMI, 44(10):6056–6073, 2021
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9cf3228b-03da-46f6-970d-c11f31eb9879 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Vizwiz grand challenge: Answering visual questions from blind people
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a8936b8-d246-4859-9ec9-d5c652b389c6 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering GoMatching: A Simple Baseline for Video Text Spotting via Long and Short Term Matching
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b749407b-50fe-48b0-a978-ce8c3d8f9694 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering CogVLM2: Visual Language Models for Image and Video Understanding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9b38a20-4d6d-4bd2-b78b-d9538c0a23b8 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Understanding video scenes through text: Insights from text-based video question answering
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e950c31-aac4-469b-89d8-369b270623c6 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Egotaskqa: Understanding human tasks in egocentric videos
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation de1cdeb5-a53e-4964-9007-ac5b3d5ba440 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Llava-next: Stronger llms supercharge multimodal capa- bilities in the wild, 2024
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7ddea211-36e8-41b0-a85b-d97653ced7d2 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering PP-OCRv3: More Attempts for the Improvement of Ultra Lightweight OCR System
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc3bc08a-889d-45db-ab84-ad27feba1338 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Flex- attention for efficient high-resolution vision-language mod- els
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e573457e-0ffe-4b46-9ccb-2873f6e870fd · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Mvbench: A comprehensive multi-modal video understand- ing benchmark
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ad03128f-81c0-439f-b859-7bbc660cd8ab · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Invariant grounding for video question answering
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11d2956c-034d-4817-8ebd-446f2ec62029 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Transformer-empowered invariant grounding for video question answering
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 832a7826-8ac0-4a40-9743-48d429db36a9 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Vila: On pre-training for visual language models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 35ec54d9-b5ae-410c-a685-7a899fc47138 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Egocentric video-language pretraining
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation afa12dbe-bca5-4a39-8c19-f12ce12055ac · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3f9cd5da-9089-41e9-9dc0-b25fbb69f521 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Ocrbench: On the hidden mystery of ocr in large multimodal models, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 479d4a04-a8e8-44c7-89c6-c9552dd96ba4 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 021e9583-38fd-4bca-9973-fa36d7b69e7d · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47d7ef9f-b74f-4650-85d7-479402dc0a49 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Egoschema: A diagnostic benchmark for very long- form video language understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e4a0a78-fd94-4af5-9336-0400dd24e2b4 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Docvqa: A dataset for vqa on document images
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 078de082-67ba-4d79-abcc-6bc671999581 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Infographicvqa
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78603dc3-316d-4193-9e7b-c32dfcbcb91e · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Ocr-vqa: Visual question answering by reading text in images
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee178bbd-26bf-4bf4-ab7a-c2d1842cf9e8 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Gpt-4o system card
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d88a5540-81ff-4582-ba0e-db44de5c68fd · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Gpt-4o mini: advancing cost-efficient intelligence
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b8c4d99-595b-4aa8-b55d-507fefc2139a · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Egovlpv2: Egocentric video-language pre-training with fusion in the backbone
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1a549ecd-6650-4be7-97ab-26169dfee7ff · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Learn- ing transferable visual models from natural language super- vision
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b40979d0-1bb6-450f-af94-f324be730f5a · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79a9a262-70b9-4325-b797-1de2079f524e · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd87f972-824e-4eb2-943b-4e0dad9ca5a9 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Annotating objects and relations in user- generated videos
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ad865063-2f7f-46ff-8da0-c522eb334193 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering GLU Variants Improve Transformer
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f35c8e7-ee3a-4a8a-b152-e8ba7621aba7 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Towards vqa models that can read
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f75d1ff-79d6-49ef-ad9d-e34462632914 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering EVA-CLIP: Improved Training Techniques for CLIP at Scale
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 322010f4-f1b3-4cb8-b534-b2a38c6f2186 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Gemini: A Family of Highly Capable Multimodal Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c5febe1-94b6-4db5-9200-00f1c50ca31b · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Reading between the lanes: Text videoqa on the road
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ac6caca4-760a-4449-ae7f-b41e840b3571 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 505a1408-9d54-4ed5-a34d-ef627ddcb886 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0dce2cd-fd45-4e5c-8037-27e7602e7d54 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering CogVLM: Visual Expert for Pretrained Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce44bf7e-1841-465e-b4d6-66bba7d0d2bb · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering On the general value of ev- idence, and bilingual scene-text visual question answering
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c6960a3-522f-49e1-924d-5beab171956a · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Assistq: Affordance-centric question-driven task completion for ego- centric assistant
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 21504019-c8a1-4b2a-b024-8bcaaa3c39ae · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Next-qa: Next phase of question-answering to explaining temporal actions
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bc512d6f-164a-4ff4-91ac-0044c19ddbab · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering VideoQA in the Era of LLMs: An Empirical Study
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de7ca57f-0f03-4635-ba80-185cf6a168b2 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Deconfounded video moment retrieval with causal intervention
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 579d69eb-bbc1-4f7b-b01e-272581b454eb · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Video moment retrieval with cross-modal neural architecture search
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dcb06e1c-a964-4ef3-bab3-768a016de2e9 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Robust video question answer- ing via contrastive cross-modality representation learning
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 028fa900-a965-4130-b355-e1a364996573 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb601c86-bfa8-4d66-b3ba-f2c41546f759 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9437a4c9-1fae-4824-a387-cdecb0997592 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Sigmoid loss for language image pre-training
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 743af68e-86a0-4e98-a5c4-6662b4f6faf3 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Multi-factor adaptive vision selec- tion for egocentric video question answering
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03acbac3-de00-4c32-926f-a2c60e18e971 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c964b9ad-5945-4630-94c1-15f72c0224b5 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Llava- next: A strong zero-shot video understanding model, 2024
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6490b45e-29a4-45bb-a171-635ebdb7bd6c · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Diffusion-based blind text image super-resolution
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8c8b5f18-f5f1-4841-8f7e-5927283508ab · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Beyond LLaVA-HD: Diving into High-Resolution Large Multimodal Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dd917b1-93d0-4ce9-94f5-f25fc4a8ae42 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Towards video text visual question answering: Benchmark and baseline
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72714bec-23ec-4e34-a8fa-b3ab4e74f246 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Exploring sparse spatial relation in graph inference for text- based vqa
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 68071ac0-de17-4ee3-9072-9ca247484a51 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Scene-Text Grounding for Text-Based Video Question Answering
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cc3978e-f96d-4cba-a5aa-c51c18902e0e · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering I” should be used appropriately. Requirement 5: The questions should be of moderate length. When announcing the question please label each question as “Question 1, 2, 3: {question}
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e2e47122-0bb7-4a7c-b3bd-dfa58fb80fa8 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering For example:
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b9872bd-9eab-43fe-9ccb-b8c93b47c37c · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 458701d6-fef6-4b9e-9a3f-844702db5b7d · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 54dec6c6-efc5-4252-97c2-4b18cbce608f · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering For example:
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 82e9b652-e3b4-40f4-aa59-b4cecc0acba3 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a1ed87b2-e23a-473e-a0a2-10f9877f776c · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a97fc98a-d790-4ae9-96b9-b64303857f51 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a2249fe-e1cd-4213-91f8-40bdbd6c6474 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering For example:
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5b4ca62c-4072-42f2-acf9-3f8579a0c459 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 22df1ca5-ea0d-4615-beff-1780dbe5dafa · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c703cccc-3d52-47c6-aa33-797bd13546f4 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c8d3008-3003-4077-b3ff-77a4a12865ad · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering For example:
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86832389-b959-4be6-add2-26981f4cbd4b · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 34ead312-430b-4a56-b89c-d83a5f1eb0a1 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a77d3a1-a090-4eed-89fd-57360faa4222 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering For example:
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13fb9799-69ef-463f-8e69-265da5d13af4 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unresolved cited work
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c69b76a0-2013-4a16-bcb6-6290905136b1 · outbound
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering Unanswerable
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.