Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:31:42.599141Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 138 outbound references and 0 inbound Pith citation observations for arXiv:2506.07886.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:31:42.599141Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 138 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 15894f6c-633e-4832-806d-3e43a6a26cda · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a8f9399-7488-4ef2-a179-2678ae232f49 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Phi-3 technical report: A highly capable language model locally on your phone
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf5137c5-86e3-4333-9e18-23dcab4e0c56 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Cosmos World Foundation Model Platform for Physical AI
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f63069d3-06bb-4d13-8e40-4a4c0cd3b781 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Flamingo: a visual language model for few-shot learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fad7658-0c0b-409e-886d-56e0c4d08177 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Scenescript: Reconstructing scenes with an autoregressive structured language model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0617dc0d-0e34-4703-96c7-c1db19e69ccb · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Newcombe, and Vasileios Balntas
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc8ce495-3b8b-4b19-bc39-a6519066f386 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining MultiMAE: Multi-modal multi-task masked autoencoders
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 449cfd9e-580e-4616-901f-ab5e8a91cdd1 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining 4M-21: An any-to-any vision model for tens of tasks and modalities
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13c506a7-8635-4d99-9918-5431ab0404ce · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Lending a hand: Detecting hands and recognizing activities in complex egocentric interactions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f68b6e4d-f5d9-4ffe-b548-1bef9b959f7b · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8f5cdf0-83c4-44f1-8530-eb19bc33e420 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Stable video diffusion: Scaling latent video diffusion models to large datasets, 2023
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49ff6499-ee5d-430f-8f77-31886228d558 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Align your latents: High-resolution video synthe- sis with latent diffusion models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 541dd731-2077-4ef5-b5a6-00c8efb30f53 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Scene coordinate reconstruction: Posing of image collections via incremental learning of a relocalizer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a852ea0f-f485-4732-8c3c-2aa14db15dee · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Genie: Generative interactive environments
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c1d5162-9dbf-4d9a-b756-670d26569b51 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining End-to-end object detection with transformers
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f9630f1-da1a-4b4d-9abd-2f82a4159002 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Emerging properties in self-supervised vision transformers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15fb6131-0fbc-44c4-a196-1853f71245d8 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models, 2024
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e90f27-0f2f-4cdd-98fd-c557c5db6b1b · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Fleet, and Geoffrey Hinton
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8afff46f-d8e4-473d-8a61-84d53bd87593 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Control- a-video: Controllable text-to-video diffusion models with motion prior and reward feedback learning, 2024
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b910072-2163-4309-89dc-77bd1f85f945 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Scaling egocentric vision: The epic- kitchens dataset
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00a2618b-7fce-4b42-9bec-9b401e93b5e4 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d95ac89-690b-435f-af8c-9526af2a8c0b · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Structure and Content-Guided Video Synthesis with Diffusion Mod- els
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34ffb50e-9f66-460b-8ce2-7ad3e4a53694 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Black, and Otmar Hilliges
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dfc6897-3adf-4384-893a-18b9fd95211d · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining HOLD: Category-agnostic 3d reconstruction of interacting hands and objects from video
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 505417fb-eae6-4fa0-b2de-e1c0b6e9f9d2 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining VIOLET : End-to-End Video-Language Transformers with Masked Visual-token Modeling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a2aeec-0259-4acf-8ae5-04c17ce47536 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining First-person hand action bench- mark with rgb-d videos and 3d hand pose annotations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fd4f5f4-7520-4d04-a576-e2df90ba89ff · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Imagebind: One embedding space to bind them all
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e5a3b5b-3cb4-4816-a31b-6ad45b96c684 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Accurate, large minibatch sgd: Training imagenet in 1 hour, 2018
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0801895d-a503-4cb6-a12f-0c36348f3d69 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Ego4d: Around the World in 3,000 Hours of Egocentric Video
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcf61676-1751-4004-80ab-d9232c3a7592 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Ego-exo4d: Understanding skilled human activity from first- and third-person perspec- tives
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59fcde70-c33c-4575-9deb-69eab36e6483 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Animatediff: Animate your personalized text- to-image diffusion models without specific tuning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9af1b819-83cf-4904-80fc-840d4704c4aa · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining World Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c1575ea-1ee2-4ed0-8020-9e970c0fd46a · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Girshick
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfecc510-d9c4-4fa7-8696-949fad01d9f0 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Classifier-free diffusion guidance
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c087ee39-de85-40d5-ab00-c371152241bb · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Video diffu- sion models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a4af980-0b14-40ca-abf4-c16d7b4263c4 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining The curious case of neural text degeneration
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c459429-91fc-4ea3-a3cc-f48d486d3cf5 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Cogvideo: Large-scale pretraining for text-to- video generation via transformers
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e6a5851-42d7-4baf-be4c-4db19fd1e816 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Gritsenko, Jasmijn Bastings, Ben Poole, Rianne van den Berg, and Tim Salimans
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ff59b2a-d994-4462-9651-ffcf2ec61f47 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Ross, and Alireza Fathi
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d63a9227-59ab-40b8-b469-aae67ab99370 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Predicting gaze in egocentric video by learning task- dependent attention transition
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a3c13d0-fb71-46ce-b967-3ad57c9ea5f2 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining GPT-4o System Card
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 400f5c80-0591-4a7f-a8ab-ffd96cbb14b8 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Epic-fusion: Audio-visual temporal binding for egocentric action recognition
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04454c8a-dc38-4169-a86b-a8039f528166 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Re- purposing diffusion-based image generators for monocular depth estimation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c185bded-22cc-4cc7-bf12-3d5c9ceec635 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Video depth without video models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36c03888-f351-40d8-834e-23e0d4fd5506 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Text2Video-Zero: Text- to-Image Diffusion Models are Zero-Shot Video Generators
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 306b8dd8-70c4-4016-8e9a-b90231c1cee4 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Segment anything
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85cfafe2-386a-4e0b-a8b3-b6856a206452 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Harmsen, and Neil Houlsby
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56202ac3-0576-4504-b221-2e677d77f744 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining VideoPoet: A large language model for zero-shot video generation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aaa8fa9-d631-4558-9fcf-3e60bf0ae668 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining H2o: Two hands manipulating objects for first person interaction recognition
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4caa288a-432f-4f65-affa-3db175476b05 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a82b8a1d-6305-4cfa-93db-c0ffa7b3405a · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Lisa: Reasoning segmenta- tion via large language model
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31d3e8ba-650b-4fb4-992f-ca9afb0f6ad1 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Egogen: An egocentric synthetic data generator
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4bd1b8c-c516-46d8-a784-fb266f7e1a4f · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining VideoChat: Chat-Centric Video Understanding
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aec46962-59cc-45c6-acbc-8457da029965 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Megasam: Accurate, fast and robust structure and motion from casual dynamic videos
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45d097af-9a90-49d3-91e0-60c25371e42e · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Video-LLaV A: Learning united visual representation by alignment before projection
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e263881-3e9a-4794-86af-d218551e21ae · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Cross-view exocentric to egocentric video synthesis
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a27f75-5b00-4edf-8daa-383a5ad7a7c2 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b3c69ab-e05a-4fbb-b93b-29013773e9f4 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Exocentric-to-egocentric video gener- ation
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f7ff5ef-cb26-4e70-a7fc-dc3429d9c4d3 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Li, Ying Shan, and Ge Li
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fee22a39-4c72-4ac1-8728-4142f468e909 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Hoi4d: A 4d egocentric dataset for category-level human-object interaction
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baa0c2cf-52c5-41ae-99fe-26b4bc392390 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Unresolved cited work
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9064648b-65c1-4b38-a48f-fdd5d5acbc60 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining A convnet for the 2020s
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95d2052-f35f-4041-8469-e7085267b8e4 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Decoupled weight de- cay regularization
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d35445cd-14d2-407e-a5e3-3c65dc60ab61 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Unified-io 2: Scaling autoregressive mul- timodal models with vision, language, audio, and action
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ddd5a67-06cf-42cf-968e-3ddfee63a802 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining UNIFIED-IO: A uni- fied model for vision, language, and multi-modal tasks
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a87fcc94-61c6-4fb1-a7f3-9790ca147d68 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Align3r: Aligned monocular depth estimation for dynamic videos
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e196fdd4-630d-480a-8b97-1abbf6d199b6 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Dream Machine
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85040bb0-72f9-4587-923d-e6ba6321b337 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Videofusion: Decomposed diffusion models for high-quality video generation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b49d3759-826d-4058-912b-f96bc2c8be84 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Aria Everyday Activities Dataset
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6e9ab21-5033-431c-80fe-6db7083e494f · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Nymeria: A Massive Collection of Multimodal Egocentric Daily Motion in the Wild
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd07221-c02f-4ec3-a438-49767fe5dfa2 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Video-chatgpt: Towards detailed video understanding via large vision and language models
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9977c314-da17-4ddf-b118-483416f1e3f9 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Mm1: Methods, analysis & insights from multimodal llm pre-training, 2024
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9deb7a6-65ed-4c3a-b7db-84c72c36390f · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Project Aria Glasses
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8ba322c7-646c-418d-9622-261ee4a5c97b · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Transformers are Sample-Efficient World Models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dc57102-cd1d-463d-b21a-8dcf0ada7d48 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining HoloLens 2
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c137608e-1b3b-4edf-bd0f-d879c2a3dbf5 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining 4M: Massively multimodal masked modeling
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b4420ba5-ac62-4b89-b5cb-6046902b1220 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining AssemblyHands: towards egocentric activity understanding via 3d hand pose esti- mation
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c79c1e4d-4e9a-4db0-aae0-0eff5aeb866a · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Video generation models as world simula- tors
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 374a7076-4d4a-4983-9a51-fd3c31220660 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fb761d3-f991-4650-a924-fac389397dcf · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Aria digital twin: A new benchmark dataset for egocentric 3d machine perception
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b0563564-1aa4-491d-a10b-6305dffedad5 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Movie Gen: A Cast of Media Foundation Models
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a104cc6-49f3-486a-a341-1b1e0f6f3a57 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Learn- ing transferable visual models from natural language super- vision
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43e157ec-730b-452b-8b52-c5821f2d7f52 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64ae6b91-6d2e-4268-bbab-dff2f7364b2d · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining High-Resolution Im- age Synthesis with Latent Diffusion Models
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3f70a23-1293-4c39-9ff0-f360abc10974 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Gen-3 Alpha
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66bddc3a-7e9c-4f03-bd41-b88c485106d3 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Lamar: Bench- marking localization and mapping for augmented reality
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07d1e775-b7d6-4c92-94dc-a1b1a51ed7b1 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Sener, D
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee2d64bc-3bc5-478e-af14-5ff57d27f4ce · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Make-a-video: Text-to-video generation without text-video data
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2706b4d7-8665-4f5e-9dc2-06dd67743634 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining The Replica Dataset: A Digital Replica of Indoor Spaces
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c84926b5-1768-451c-b056-ae70ede2ec88 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Part, Ioannis Papaioannou, Arash Eshghi, Ioannis Konstas, and Oliver Lemon
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5e558a0-b2ed-4d3d-bf55-4fb269ab28b3 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Emu: Generative pretraining in multimodality
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ba77c0d-e496-42e5-a6c2-5147d5e6eb40 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Gemini: A Family of Highly Capable Multimodal Models
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 522fd6c4-b46a-4fc0-b37e-a072ab425b44 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Kling ai video generator
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8559f55-245d-4006-a0d8-af30007aedce · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining DROID-SLAM: Deep vi- sual SLAM for monocular, stereo, and RGB-d cameras
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2fa6a850-a295-4b97-a6bc-d0ddc6ddd58f · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining VideoMAE: Masked autoencoders are data-efficient learn- ers for self-supervised video pre-training
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a8dbe85-a2d5-4cb7-a498-9d7c352ce84d · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Towards accurate generative models of video: A new metric & challenges, 2019
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e2f2197-f1ed-4ec3-9ad4-e1979f876097 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Diffusion Models Are Real-Time Game Engines
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e673542-0d8e-4206-bd37-c9d6f93d6215 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Neural discrete representation learn- ing
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fc44bf5-10be-4363-8f68-d9a1423b43b9 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Attention is all you need
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f383cfda-5eff-479b-b21c-46daadaed112 · outbound
EgoM2P: Egocentric Multimodal Multitask Pretraining Phenaki: Variable length video generation from open do- main textual descriptions
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.