Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:1706.04261.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T17:16:52.452286Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T17:09:58.596248Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ee1876de-1579-4769-9d4e-a8cb6fecf1d4 · inbound
VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents The "something something" video database for learning and evaluating visual common sense
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5fdac18d-6351-4019-a772-6945646dee21 · inbound
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models The "something something" video database for learning and evaluating visual common sense
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7183458b-6492-4297-9937-aaba07f7b4e8 · inbound
Co-Evolving Latent Action World Models The "something something" video database for learning and evaluating visual common sense
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 994d8252-2e83-4c78-809d-7b71f84a16e5 · inbound
Interpreting Video Representations with Spatio-Temporal Sparse Autoencoders The "something something" video database for learning and evaluating visual common sense
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09f9666b-78fd-4c09-840e-50e786f135d5 · inbound
From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data The "something something" video database for learning and evaluating visual common sense
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1ee09952-dde3-45bb-9494-20e2eeb9c0ff · inbound
HumanNet: Scaling Human-centric Video Learning to One Million Hours The "something something" video database for learning and evaluating visual common sense
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a2e9053-f973-4bd5-92a1-d09c21bfe8f3 · inbound
World Action Models: The Next Frontier in Embodied AI The "something something" video database for learning and evaluating visual common sense
Reference 176
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 934f1c32-356b-44da-9f54-8e717575702e · inbound
DiLA: Disentangled Latent Action World Models The "something something" video database for learning and evaluating visual common sense
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d794f375-357b-4a20-b4b0-ab59df96ca84 · inbound
Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model The "something something" video database for learning and evaluating visual common sense
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f0b5406e-cecb-4ae6-b460-96405cd53248 · inbound
PEIRA: Learning Predictive Encoders through Inter-View Regressor Alignment The "something something" video database for learning and evaluating visual common sense
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1173a47c-6d92-4b71-9b1e-f9cfbef8840d · inbound
MJEPA: A Simple and Scalable Joint-Embedding Predictive Architecture for Audio-Visual Learning The "something something" video database for learning and evaluating visual common sense
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 02496fb4-be63-4472-adb3-5ee9d78284ca · inbound
iFLYTEK-Embodied-Omni Technical Report The "something something" video database for learning and evaluating visual common sense
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 037bd44e-3e4d-4bf9-9449-51853ebdf1f8 · inbound
Learning to Detect Cross-Modal Negation: An Analysis of Latent Representations and an Attention-Based Solution The "something something" video database for learning and evaluating visual common sense
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.