Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T07:16:29.588452Z
Paper Citation Record · LEDGER
As of 2 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2602.05638.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T07:16:29.588452Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:30:57.042487Z
A source-named dated measurement, never combined with another source.
Source: cited_works
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 65c8424e-6828-4ea6-8cf6-2bab48d67022 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos DINOv2: Learning Robust Visual Features without Supervision
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation c7eaac3f-bcd2-4ab7-9889-f3bd8f9f0e80 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Masked autoencoders are scalable vision learners
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5a94e565-6873-4479-8a3f-d1166deb91bb · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Masked autoencoders as spatiotemporal learners
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation a8c9586c-5198-4534-9ffa-53d9cfbe7d17 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Endovit: pretraining vision transformers on a large collection of endoscopic images
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation f0317426-8f8a-4b0f-8a0b-9561d3a8b46c · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Foundation model for endoscopy video analysis via large- scale self-supervised pre-train
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation bf41add1-0ecf-4883-a5fa-aa24103cf21a · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos General surgery vision transformer: A video pre-trained foundation model for general surgery
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 87a0759a-26b3-4e95-bce9-2fc11c56f802 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation d34de7fd-0b84-4b29-b33b-c8fcc05c2a96 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Videomae v2: Scaling video masked autoencoders with dual masking
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 098ed1b4-6051-48cc-9cfb-d45625b74eef · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Dissecting self-supervised learning methods for surgical computer vision
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 71913d2d-5ac9-4783-a5dc-4b4654ef6414 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Endonet: a deep architecture for recognition tasks on laparoscopic videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 100195a1-baa4-47bc-a1c2-67a51dad23bf · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Pitvis-2023 challenge: Workflow recognition in videos of endoscopic pituitary surgery
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation b2615ac1-326c-4387-9d6a-beaf33b18a48 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Egosurgery-phase: a dataset of surgical phase recognition from egocentric open surgery videos
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation b991f193-b445-45b3-8239-99d6d20e37d5 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Revisiting Feature Prediction for Learning Visual Representations from Video
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation b5c27669-8051-4f81-9fca-a33b4e8e9d6e · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 1f629bb9-cf2f-445b-a19a-b81cd415a272 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Bootstrap your own latent: A new approach to self-supervised learn- ing
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 4e21a4e5-bfbb-4062-9c7e-20bf4afa6a3a · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Internvideo2: Scaling video foundation models for multimodal video understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 7214fd01-073e-495c-ba87-88203e7ec15b · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Internvideo-next: Towards general video foundation models without video-text supervision
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 8d73de36-f6d3-4ab2-bc36-7e6bf6778573 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Emerging properties in self-supervised vision transformers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 1021f588-4279-4938-b861-5e2df567ecaf · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos DINOv3
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 4e6d5886-4dd6-4273-a38c-99dd91c43f7c · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Gastronet-5m: A multicenter dataset for developing foundation models in gastrointestinal endoscopy
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 1c07827e-2596-4bfa-be7d-42d423e629cc · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Self-supervised learning for endoscopic video analysis
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ebc9fd4d-172e-470d-8857-de06490ac087 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Endomamba: an efficient founda- tion model for endoscopic videos via hierarchical pre-training
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 6fd0697f-7fca-40d5-a3fa-587ab0e66800 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Scaling up self-supervised learning for improved surgical foundation models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation a0969ef8-c80e-4c83-8510-7030bec8b518 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Learn- ing multi-modal representations by watching hundreds of surgical video lectures
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ae41b1ce-250c-4a0b-86ad-af55f87cc6f9 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos The TUM LapChole dataset for the M2CAI 2016 workflow challenge
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 881fa467-b602-416e-b807-a77cf03bdad8 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Rendezvous: Attention mechanisms for the recognition of surgical action triplets in endoscopic videos
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 7ee17f12-9e50-4697-b874-622379a6e908 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Autolaparo: A new dataset of integrated multi-tasks for image-guided surgical automation in laparoscopic hysterectomy
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 91c4640e-79da-496f-a89b-21038eaea380 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Surgical workflow recognition and blocking effectiveness detection in laparoscopic liver resection with pringle maneuver
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation bd78fb99-e91a-4db8-b9b3-a2cb15f72130 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Ophnet: A large-scale video benchmark for ophthalmic surgical workflow understanding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation f4bda7a8-62b1-4c61-89a2-f4f9cae6c375 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Analyzing surgical technique in diverse open surgical videos with multitask machine learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation bcaee95c-e5e9-48f3-97a5-39cdcec28c13 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos A dataset and benchmarks for segmentation and recognition of gestures in robotic surgery
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 416044c4-e24f-4216-8ad4-2198076a5ebd · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Aixsuture: vision-based assessment of open suturing skills
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 7f67398d-310d-4c54-844f-8d8438a43f0a · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Video retrieval in laparoscopic video recordings with dynamic content descriptors
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0e5c71cd-c19b-4398-98b6-8ffae8dba498 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Contrastive transformer- based multiple instance learning for weakly supervised polyp frame detection
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation f98b61b0-9374-4ff1-a737-679468ca8c54 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Wm-dova maps for accurate polyp highlighting in colonoscopy: Validation vs. saliency maps from physicians
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 2cc3a194-d676-4fae-88de-d27188eeae67 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Implicit domain adaptation with conditional generative adversarial networks for depth prediction in endoscopy
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 7df0091d-7d66-432f-a3d1-bdbd0e6f5b4f · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Colonoscopy 3d video dataset with paired depth from 2d-3d registration
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 7a94bfda-bcd3-46bc-a550-8a6d427f0949 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Cataracts: Challenge on automatic tool annotation for cataract surgery
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation eef2e08a-40b3-4bf1-b948-f20326cd3f64 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Challenges in multi-centric generalization: phase and step recog- nition in roux-en-y gastric bypass surgery
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 2d9333d2-f622-4027-8bdd-c18428c6e1b9 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Copesd: A multi- level surgical motion dataset for training large vision-language models to co-pilot endoscopic submucosal dissection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation fe24adf8-d2a4-435a-b1e2-1636d1db2711 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Towards holistic surgical scene understanding
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 00a9411f-b337-447d-82ab-0ab0a47cb83b · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Kvasir-seg: A segmented polyp dataset
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 10bfce8b-2270-4931-b1ae-c7ab493c2d0e · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos A benchmark for endoluminal scene segmentation of colonoscopy images
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0d5a1ad6-0bd0-4022-9370-4de51db83397 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Towards automatic polyp detection with a polyp appearance model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 59a4ce96-7c01-40eb-8b5b-5b2c83f43849 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Toward embedded detection of polyps in wce images for early diagnosis of colorectal cancer
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 91b31e61-a7de-4ceb-8c61-a2b21be828ba · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Pranet: Parallel reverse attention network for polyp segmentation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 4d2cb680-6211-4d5a-b3bc-b180c933ea31 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Uacanet: Uncertainty augmented context attention for polyp segmentation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 243a59d5-e604-4e35-950e-9cf958af9d13 · outbound
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos Pranet-v2: Dual-supervised reverse attention for medical image segmentation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0ba885f5-863f-45d2-9ca0-d3b792d444ab · inbound
LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.