Pith. sign in

Paper Citation Record · LEDGER

EgoVLM: Policy Optimization for Egocentric Video Understanding

As of 18 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 7 inbound Pith citation observations for arXiv:2506.03097.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03097 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:14:29.581142Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:26:46.165458Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T10:41:29.637290Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9cf0adb5-e990-427e-b9ba-fcc2aabeae8c · outbound

This paper cites an unresolved cited work.

EgoVLM: Policy Optimization for Egocentric Video Understanding Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:14:31.912954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:26.484755Z digest=sha256:7539900e10e07bfbf53885d6208b4983cbc977e3e1aa500ffc50bfdfbd56e8d1

Observation 7c62262a-f97b-4a42-a49e-e7c8c4d7f46d · outbound

This paper cites GPT-4 Technical Report.

EgoVLM: Policy Optimization for Egocentric Video Understanding GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:26.585399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:26.585399Z digest=sha256:0dfbe5715eae965b796f28440bb66ac2c18f5cac566c2a3edd25ae36adccefd8

Observation 268c3634-4b6b-4f1d-8dc7-fdbe3453239c · outbound

This paper cites Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs.

EgoVLM: Policy Optimization for Egocentric Video Understanding Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:26.665985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:26.665985Z digest=sha256:5fa362ff0b873561575b5c6b9d1771a1311ad4bb2374404f8e6a47df488c120a

Observation 6e54bb2d-fa9d-4c34-aa24-d369516496b3 · outbound

This paper cites Qwen2.5-vl technical report, 2025.

EgoVLM: Policy Optimization for Egocentric Video Understanding Qwen2.5-vl technical report, 2025

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:26.747337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:26.747337Z digest=sha256:58a28d524d7373d4b2d542b7aaf69b00b1cb148512a24a74546300231aa2e90d

Observation 829fa18c-862f-4e6b-8a1a-caaeada02d0f · outbound

This paper cites EgoPlan-Bench: Benchmarking Multimodal Large Language Models for Human-Level Planning.

EgoVLM: Policy Optimization for Egocentric Video Understanding EgoPlan-Bench: Benchmarking Multimodal Large Language Models for Human-Level Planning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:26.803299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:26.803299Z digest=sha256:dcfdae308e84b9d4e232528c519464509132447fd24d792df18d0a8396f314b4

Observation 466cd78f-6508-472d-b095-c69c619c2e8e · outbound

This paper cites VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI.

EgoVLM: Policy Optimization for Egocentric Video Understanding VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:26.902583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:26.902583Z digest=sha256:390eaab4316986b23a7031c9bf6bc7283390e535163b0aa104a6ff8f7c75eed7

Observation 6531419f-4afd-4eed-8a25-94de004f008c · outbound

This paper cites an unresolved cited work.

EgoVLM: Policy Optimization for Egocentric Video Understanding Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:14:31.779038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:26.966208Z digest=sha256:b04866fe82858892c00952132a75d9774712b3be66b88462c4849819a3d8dc19

Observation 7281fecc-e61b-40e9-aba9-dc314e630fbf · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

EgoVLM: Policy Optimization for Egocentric Video Understanding ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:27.043868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:27.043868Z digest=sha256:ef42c3719c8ed286a89b8cf3b226269a1acf1ff5e3b169889303ae3917d971ed

Observation fd5018df-37a8-459b-bd86-54df58edf779 · outbound

This paper cites GPT-4o System Card.

EgoVLM: Policy Optimization for Egocentric Video Understanding GPT-4o System Card

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:27.136926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:27.136926Z digest=sha256:03b1077bbe503cb865d465ae5da59d6e9caf882ee8fee391dc670790070fb877

Observation 8e1f2ebc-5652-4ef5-9a09-f499df2d6a3d · outbound

This paper cites Mvbench: A comprehensive multi- modal video understanding benchmark, 2024.

EgoVLM: Policy Optimization for Egocentric Video Understanding Mvbench: A comprehensive multi- modal video understanding benchmark, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:31.650068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:27.251656Z digest=sha256:902c9b59c1728cd32a51de8b73912473bdc824bad80b8407594087d19a5f509b

Observation 9d089f33-2fa1-42a4-b067-705125e34387 · outbound

This paper cites Dual-Difficulty Curriculum Learning for Direct Preference Optimization.

EgoVLM: Policy Optimization for Egocentric Video Understanding Dual-Difficulty Curriculum Learning for Direct Preference Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:27.351094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:27.351094Z digest=sha256:9a067f4b64f73428f915395cbf79b37cd5200af25d736063f65007a74abd4118

Observation 94afe2fa-f34c-4ffc-9b8f-dce191566858 · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries.

EgoVLM: Policy Optimization for Egocentric Video Understanding ROUGE: A package for automatic evaluation of summaries

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:31.558170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:27.482076Z digest=sha256:bd8377e9ba818b8c4bcb06aeb95f2b5a5f48ba19a234f7d23c6c1d5f0d7aa3c0

Observation 3094ddd2-5790-45dc-8488-ba0939c6f1c1 · outbound

This paper cites Microsoft coco: Common objects in context.

EgoVLM: Policy Optimization for Egocentric Video Understanding Microsoft coco: Common objects in context

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:27.604637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:27.604637Z digest=sha256:6673ed4e77110f852c3b38983255b40424825e768a0435f2cf35409865a4dea5

Observation 1f733e8e-e71c-4a73-b8e3-a4d59c859591 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

EgoVLM: Policy Optimization for Egocentric Video Understanding Understanding R1-Zero-Like Training: A Critical Perspective

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:27.701631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:27.701631Z digest=sha256:4c5711f675f87a3685d569f8a59208ce5e9efb2298a18743d05a367221325f2a

Observation 52c4be18-2d30-49e9-ba1d-c1880b371685 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

EgoVLM: Policy Optimization for Egocentric Video Understanding Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:27.794624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:27.794624Z digest=sha256:1c684a3ea4783d0072129913b0c9507e3c6a682572d3b655ae8abedc7815ecaf

Observation 32e78173-37f0-4384-9aa4-b0f083b520eb · outbound

This paper cites Reasoning models can be effective without thinking, 2025.

EgoVLM: Policy Optimization for Egocentric Video Understanding Reasoning models can be effective without thinking, 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:31.418970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:27.897137Z digest=sha256:2a228a09ac65b2a5408a1b19694d774da9d242506d997fbe545ec9b629367159

Observation a91b1c91-5fe9-4fc6-8fb2-4495f31efc8e · outbound

This paper cites Egoschema: A diagnostic benchmark for very long- form video language understanding.

EgoVLM: Policy Optimization for Egocentric Video Understanding Egoschema: A diagnostic benchmark for very long- form video language understanding

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:31.299087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:27.989782Z digest=sha256:1ae8fcf8c3e53b0e9f025420873ecacad80eca9d3b49ae5c6547266a537f55c9

Observation dd2855f9-5bca-45ea-969e-eb567c9996a8 · outbound

This paper cites Training language models to follow instructions with human feedback.

EgoVLM: Policy Optimization for Egocentric Video Understanding Training language models to follow instructions with human feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.107807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.107807Z digest=sha256:accd040ca39eafbb33a25877dbaa15cd05b2270cf14efdf930ccb2cab852f42a

Observation 08c17e3c-3a5d-403c-b440-72646bd5320c · outbound

This paper cites Medvlm-r1: Incentivizing medical reasoning ca- pability of vision-language models (vlms) via reinforcement learning, 2025.

EgoVLM: Policy Optimization for Egocentric Video Understanding Medvlm-r1: Incentivizing medical reasoning ca- pability of vision-language models (vlms) via reinforcement learning, 2025

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:31.178831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:28.186843Z digest=sha256:a15a0bafc2f224f118d97c2c1cfa19f663380997dd232c4274bfe89dcc0fe435

Observation 298fb0c7-5478-40fc-ad63-b781ba56dcf8 · outbound

This paper cites Learning transferable visual models from natural language supervision.

EgoVLM: Policy Optimization for Egocentric Video Understanding Learning transferable visual models from natural language supervision

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:31.041234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:28.261259Z digest=sha256:78c70095d3834b7ebfab2867a3d169fbc07defed6305825a537d5441cd532575

Observation 7742fe99-1aa1-48bf-9fde-6792242a6b3b · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

EgoVLM: Policy Optimization for Egocentric Video Understanding Direct preference optimization: Your language model is secretly a reward model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.361137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.361137Z digest=sha256:2887a998ff230e19080b39e9b6b0be3a20bc85efd8d103dc39dc661208a414ab

Observation f9c58009-197e-4dfe-8e46-e8a0dc3e55c6 · outbound

This paper cites Proximal Policy Optimization Algorithms.

EgoVLM: Policy Optimization for Egocentric Video Understanding Proximal Policy Optimization Algorithms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.448038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.448038Z digest=sha256:1bff12da9db646247f454769be7af920766992e48fa98e026dc9796547a6f728

Observation e43a7e88-60ed-4972-a6b7-ff2fe5ebbd17 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

EgoVLM: Policy Optimization for Egocentric Video Understanding DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.572354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.572354Z digest=sha256:05af5402bac3ad156af8930a143a07e3dc5e9cd6e287cb6278e2bc449c42d3c1

Observation 42623fe7-8624-4ba8-aad3-ed55fcf989cf · outbound

This paper cites an unresolved cited work.

EgoVLM: Policy Optimization for Egocentric Video Understanding Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:14:30.870663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:28.653266Z digest=sha256:e0beb97069dd07ff3930156964c8b899fda1d0a8fbbccf901b4002cf0b5b5701

Observation 071969d7-a674-4a0c-ae25-d3cf448472a5 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

EgoVLM: Policy Optimization for Egocentric Video Understanding VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.744187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.744187Z digest=sha256:06baa38ae59c8b47bc1e80b3e2292a7b794a369eadad87d6947d0766be434de4

Observation 3badcadc-13ad-476b-9356-b973be507cc3 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

EgoVLM: Policy Optimization for Egocentric Video Understanding Gemini: A Family of Highly Capable Multimodal Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.861844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.861844Z digest=sha256:04fbf6c77100b85170cf7b15f1a9e4a78cdbc9c90fa45e5fde121ab283bd9b14

Observation e7ae7185-273b-42c0-a295-50a794bcd589 · outbound

This paper cites Internvideo2.5: Empowering video mllms with long and rich context modeling, 2025.

EgoVLM: Policy Optimization for Egocentric Video Understanding Internvideo2.5: Empowering video mllms with long and rich context modeling, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:28.924627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:28.924627Z digest=sha256:5ff6dc0abc346680db2b531cbde8368916aab2fecb1c8073f7b9ccbf5e35529d

Observation 350801aa-6cc5-4d4c-a2e3-37346c54cf80 · outbound

This paper cites Wizardlm 2, 2024.

EgoVLM: Policy Optimization for Egocentric Video Understanding Wizardlm 2, 2024

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:30.711275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:29.032858Z digest=sha256:2d86b179d573b1afa3faa428ca05f5e823c29a2918c9660483b1220781f6ffc4

Observation 7825ea91-ba7e-42fb-a506-0c2806d66248 · outbound

This paper cites ST-Think: How Multimodal Large Language Models Reason About 4D Worlds from Ego-Centric Videos.

EgoVLM: Policy Optimization for Egocentric Video Understanding ST-Think: How Multimodal Large Language Models Reason About 4D Worlds from Ego-Centric Videos

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:29.112728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:29.112728Z digest=sha256:f454c24fa2ab616a3ba6ee80c88de8a84838b88f537b742646c743199daa2f2a

Observation 819dae82-f2af-4964-b142-c92a2946656e · outbound

This paper cites Egolife: Towards egocentric life assistant.

EgoVLM: Policy Optimization for Egocentric Video Understanding Egolife: Towards egocentric life assistant

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:30.571100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:29.229419Z digest=sha256:1b2a20f3d8dae0e984cc14f0804dbc9b6a7265d130efd0c710a8ec87389488ef

Observation f48211bc-4251-41c3-bd38-c7686456d4e3 · outbound

This paper cites Mm-ego: Towards build- ing egocentric multimodal llms.

EgoVLM: Policy Optimization for Egocentric Video Understanding Mm-ego: Towards build- ing egocentric multimodal llms

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:30.450637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:29.289694Z digest=sha256:c5102bce1f0298cdfd071e134f72a3a84834ac16b9b3a7073d569810cb83f839

Observation 2a23480f-3d85-4f22-964e-c9c8e56192d7 · outbound

This paper cites Video instruction tuning with synthetic data, 2024.

EgoVLM: Policy Optimization for Egocentric Video Understanding Video instruction tuning with synthetic data, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:29.397148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:29.397148Z digest=sha256:cf822ac238dd562f0dfff2c64c2b92337d2e2e21818e6853397699d81a0044c9

Observation 6781ea2e-d1f7-4430-a462-1821dc6db063 · outbound

This paper cites Llamafac- tory: Unified efficient fine-tuning of 100+ language mod- els.

EgoVLM: Policy Optimization for Egocentric Video Understanding Llamafac- tory: Unified efficient fine-tuning of 100+ language mod- els

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:30.310972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:29.444276Z digest=sha256:34f75223d9ad87a21f87d2364c431f0cd7a6f4d3628d91a16ca5a12dcf0c29af

Observation c3d95f10-d7c8-4aaa-b7c9-c88f519f0f44 · outbound

This paper cites R1-zero’s ”aha moment” in visual reasoning on a 2b non-sft model, 2025.

EgoVLM: Policy Optimization for Egocentric Video Understanding R1-zero’s ”aha moment” in visual reasoning on a 2b non-sft model, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:30.161229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:29.544152Z digest=sha256:200c9598e3b3380ed3a3a445d50cccd51017f554a9c45b1bcd0a94d5cdf6e130

Observation da53d852-81dc-4541-ae7c-4a469c6a3fa0 · outbound

This paper cites Egotextvqa: Towards egocentric scene-text aware video question answering, 2025.

EgoVLM: Policy Optimization for Egocentric Video Understanding Egotextvqa: Towards egocentric scene-text aware video question answering, 2025

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:14:30.027639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:14:29.581142Z digest=sha256:63ba9c244d368721a0b88e839974cbb729ca6a3b25e9b8b20d74519033c70f8d

Pith citing papers

Observation 004476fc-40b1-40e2-bccc-5727d2f6faaa · inbound

EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos cites this paper.

EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T17:26:46.165458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:26:46.165458Z digest=sha256:8e397ede2ab6424da9b1619f77b700065348c0ffc9beba9d1e8b2cce6b719556

Observation 358e3010-f86b-4f49-a0c8-0dcd3bb9fe7c · inbound

EgoVITA: Learning to Plan and Verify for Egocentric Video Reasoning cites this paper.

EgoVITA: Learning to Plan and Verify for Egocentric Video Reasoning EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T20:52:51.049418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:52:51.049418Z digest=sha256:daaa98e9871693f244b86c94415b3c578cbb7dbf000ab79653e385c41cd8300c

Observation db8a76be-cd62-4481-9ca2-4f1014065a3b · inbound

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next cites this paper.

EgoIntent: A Pre-Outcome Micro-Step Benchmark for Understanding What, Why, and Next EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T05:50:27.250841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:50:27.250841Z digest=sha256:b6ad6dcd3c38a2e6af462c2eef673b07ecb94392fed387d65b894805d83f2c88

Observation 1fdb557d-46fb-422e-b3b0-ef98a0901133 · inbound

Robot Learning from Human Videos: A Survey cites this paper.

Robot Learning from Human Videos: A Survey EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:41:29.640044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T04:55:44.273643Z digest=sha256:072f2772cb9c29f7bdb8b64798b7232a440a30fd482af507412c3d28b88fd899

Observation 6dbe324b-7bf7-4845-adb5-069be76a7df9 · inbound

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks cites this paper.

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:06.581439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T17:16:31.820718Z digest=sha256:2cce6857ad26ab518d078a990cdd752535823a7cb5dac10ad995c612bcddcf9e

Observation 2ed5ebf4-0079-4353-a6bd-f0edb17c0757 · inbound

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks cites this paper.

Pro$^2$Assist: Continuous Step-aware Proactive Assistance with Multi-modal Egocentric Perception for Long-horizon Procedural Tasks EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T05:19:56.255880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:19:56.255880Z digest=sha256:514d3b876fb97b615ebadc4fda2c4c37cb4edb4e8bd73effa5239ff835bcef11

Observation aa9dc3c7-0bba-4a07-8cad-434d76a7a218 · inbound

Towards Context-Aware Clinical Motion Understanding in Daily Living at Home: Freezing of Gait Detection with Egocentric Vision cites this paper.

Towards Context-Aware Clinical Motion Understanding in Daily Living at Home: Freezing of Gait Detection with Egocentric Vision EgoVLM: Policy Optimization for Egocentric Video Understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T14:44:14.335124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:44:14.335124Z digest=sha256:f861a548a6d97be34bc455cbfbf0f344e2e8be983b8e6ec9f72fef5515a0aa9e