Pith. sign in

Paper Citation Record · LEDGER

MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.07177.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.07177 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:54:56.363353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T09:27:44.030871Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5dd15c1c-9360-4df2-991e-28a84b51cac1 · inbound

Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces cites this paper.

Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:27:44.034466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T09:27:43.919941Z digest=sha256:4cb322668c744fd4d7100eb2aa54a2fb263848956664d1c038a5d27f59855a5d

Observation eb601c86-bfa8-4d66-b3ba-f2c41546f759 · inbound

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering cites this paper.

EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T12:54:56.363353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:54:56.363353Z digest=sha256:d074589bfec265fe53fcef8ff4b820cd26adf766967d1db05a511943fedfbda3

Observation 70d551c0-45c3-408b-9516-84314d43b4e0 · inbound

The Repeated-Stimulus Confound in Electroencephalography cites this paper.

The Repeated-Stimulus Confound in Electroencephalography MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T10:08:00.839120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:08:00.839120Z digest=sha256:be3c788ace4ac509f16179c77b5a5177e0488c785c60f54db4dd3db255d5a0b3

Observation 51dae55d-cb6b-4bb9-964f-d6d64d461878 · inbound

EgoSound: Benchmarking Sound Understanding in Egocentric Videos cites this paper.

EgoSound: Benchmarking Sound Understanding in Egocentric Videos MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:46:42.523208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T21:44:47.636912Z digest=sha256:2f2d7ac1ca7e3022b68ebb6ff8fbd10b74433cfaee68d08c94938e041aaf5b4a

Observation 1579ba76-ac98-4aba-943d-addc31974519 · inbound

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next cites this paper.

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T05:50:27.273604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:50:27.273604Z digest=sha256:dd4a34f79953b3f0b1ba0f758eeff326eb404a6ca1725a6122f8534a95b17769

Observation 02537c44-2174-47a2-9657-393e8f3df8ca · inbound

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks cites this paper.

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:06.630701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T17:16:31.820718Z digest=sha256:f6e726b5e9a24888b3b89bc5492b7a2576a19a2fb06dddba11056378e4eb0610

Observation f6c91d06-c0bd-4ca6-9c1a-427c92d69b0a · inbound

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks cites this paper.

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-04T05:19:56.331819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:19:56.331819Z digest=sha256:fe00b51b15417ab801d2c351c1fffa349c1f3d3f1bc1ea3d26267db46c8224b2

Observation 36593094-7e0d-4fb6-bf5c-bedde56fc7a7 · inbound

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning cites this paper.

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:48:23.168692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T14:48:12.091885Z digest=sha256:502c347c314e9621edbc56ef4ed2a21663b6fb076623b4b7e827ab8c34ce7566

Observation 82c2559e-64ed-4b5a-a851-f257df6fb9c5 · inbound

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models cites this paper.

Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:38:12.494220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:36:20.388169Z digest=sha256:b4d373050de0ce6106cfaba25712e1eb8e776f3a0ffa7f35cea1d8ec27b9b698

Observation 9a8cce54-98b9-488c-ad2a-3b209436c2fe · inbound

Reinforcing Egocentric Spatial Perception in Multimodal Large Language Models via Ego Scene Augmentation cites this paper.

Reinforcing Egocentric Spatial Perception in Multimodal Large Language Models via Ego Scene Augmentation MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T01:59:17.934979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:59:17.934979Z digest=sha256:fb028867a1353ec81a3d555368065f095a3535271838406e812b5e047c5f1ad0

Observation 687e6591-edcf-4b2a-b4e5-32cabf53b413 · inbound

Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents cites this paper.

Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T22:44:01.350884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:44:01.350884Z digest=sha256:5bc88b1572990fc3472b0470ceadc3da4f87c6ad3d7dc282421ecfed48959f16