Pith. sign in

Paper Citation Record · LEDGER

MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.07177.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.07177 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:54:56.363353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T09:27:44.030871Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5dd15c1c-9360-4df2-991e-28a84b51cac1 · inbound

Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces cites this paper.

Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:27:44.034466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T09:27:43.919941Z digest=sha256:058e89c2c08aa02cf5c9475cd77d8e4cb32711803a617418b7aab74dc59649ae

Observation eb601c86-bfa8-4d66-b3ba-f2c41546f759 · inbound

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering cites this paper.

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T12:54:56.363353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:54:56.363353Z digest=sha256:b3dacd007e59dd662af5f38dcd40f274de9fdb035fe1ec2ac484b0b28125cf68

Observation 70d551c0-45c3-408b-9516-84314d43b4e0 · inbound

The Repeated-Stimulus Confound in Electroencephalography cites this paper.

The Repeated-Stimulus Confound in Electroencephalography MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T10:08:00.839120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:08:00.839120Z digest=sha256:40945fd2f27a02d99338b3b428153c91bb665a9acd1a1d58d91d6c9920255095

Observation 51dae55d-cb6b-4bb9-964f-d6d64d461878 · inbound

EgoSound: Benchmarking Sound Understanding in Egocentric Videos cites this paper.

EgoSound: Benchmarking Sound Understanding in Egocentric Videos MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:46:42.523208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T21:44:47.636912Z digest=sha256:7ec4a7a13e58d3f22fac4ebd77d136b95a3162309b6122ebe7aaa9b3c5c22f84

Observation 1579ba76-ac98-4aba-943d-addc31974519 · inbound

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next cites this paper.

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T05:50:27.273604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:50:27.273604Z digest=sha256:b50a74bc3b504cd929898620ab9453cf607050815ca69e64e08d9d6394e8a18e

Observation 02537c44-2174-47a2-9657-393e8f3df8ca · inbound

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks cites this paper.

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:06.630701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T17:16:31.820718Z digest=sha256:99605847c196e3f8c406b6bb179687a32a0374c0adff10bf005228a76802fb58

Observation f6c91d06-c0bd-4ca6-9c1a-427c92d69b0a · inbound

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks cites this paper.

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-04T05:19:56.331819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:19:56.331819Z digest=sha256:a8565ab978fdecb8eb1a42d350ee9df07a67de299e99700a870340a5c9e79a8b

Observation 36593094-7e0d-4fb6-bf5c-bedde56fc7a7 · inbound

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning cites this paper.

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:48:23.168692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T14:48:12.091885Z digest=sha256:fe5d10d89db3673cb07c3d8d614ab6dfef32fcd031e6669bfa31d85b05209fd4

Observation 82c2559e-64ed-4b5a-a851-f257df6fb9c5 · inbound

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models cites this paper.

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:38:12.494220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T10:36:20.388169Z digest=sha256:e3f337ed54e5bfcfaa5fe792ead3ab3b11179905165aa5219216c0091f13cb1c

Observation 9a8cce54-98b9-488c-ad2a-3b209436c2fe · inbound

Reinforcing Egocentric Spatial Perception in Multimodal Large Language Models via Ego Scene Augmentation cites this paper.

Reinforcing Egocentric Spatial Perception in Multimodal Large Language Models via Ego Scene Augmentation MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T01:59:17.934979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:59:17.934979Z digest=sha256:eba5ae2fef26bdddcd00670f47291670accbb293f9b37b28bcef4bfd97563a66

Observation 687e6591-edcf-4b2a-b4e5-32cabf53b413 · inbound

Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents cites this paper.

Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T22:44:01.350884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:44:01.350884Z digest=sha256:221bf04c06b336ea98899b0f62ac8b73f07f62373a8b0cf208d35351f1eb86e3