Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T14:15:57.272572Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:1908.03477.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T14:15:57.272572Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 09347ffc-8b47-4a24-89f0-7dcf42b6192e · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 70e071cc-fe06-4f05-8010-266a3834f327 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Re-ID done right: towards good practices for person re-identification
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7526b53-9e57-488c-8b35-236a261a4bca · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings NetVLAD: CNN architecture for weakly supervised place recognition
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eb52af3d-3229-4db1-b3f3-9c81d452bc61 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings An empirical study and analysis of generalized zero- shot learning for object recognition in the wild
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 40f1f8b8-8e51-435a-8b4e-ce75457b613a · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond triplet loss: a deep quadruplet network for person re-identification
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0b862c3b-5658-46a5-81f0-01806def26c8 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Scaling egocentric vision: The epic-kitchens dataset
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c4095747-5b8f-4326-9d83-56a92a40de35 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Predict- ing visual features from text for image and video caption re- trieval
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 86de2abf-eb5e-4198-8626-0b4c4a2c1b66 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Dual Encoding for Zero-Example Video Retrieval
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 91628ae4-1ca5-4ff3-a5a9-f518899f9046 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improving image-sentence embeddings using large weakly annotated photo collections
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 24c39c6a-2d93-41d4-ab83-5847f64df17e · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep image retrieval: Learning global representations for image search
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 437a4b4c-65a0-4334-91a8-3b7703c905ea · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8356357a-0ab1-4ed0-a846-89dbbf0fb2a7 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Something Something
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 48d26ae9-a8df-4a56-a7a8-807789c60ed5 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Ava: A video dataset of spatio-temporally localized atomic visual actions
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ce2de66e-540f-4bbb-87b6-0d4b12c9a203 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b1399349-65c3-4448-af8d-9686b1960d0c · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 47c7d464-2874-4da8-b38a-bda26323ec06 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings In Defense of the Triplet Loss for Person Re-Identification
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce0550f-0af1-45af-bb9b-8b78364b9faa · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep metric learning using triplet network
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 46f38826-3731-4783-a1ae-56bcc6ed91d4 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unifying visual-semantic embeddings with multimodal neu- ral language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6d57a355-c19c-4736-9c72-0db32e089aa0 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning two-branch neural networks for image-text match- ing tasks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 165f5939-6426-4545-b699-ce487380717f · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings On the effectiveness of task granularity for transfer learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 611c895f-318a-438d-82f6-b8eef42c219e · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learnable pooling with Context Gating for video classification
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad39cbdd-ded8-4ef3-a79b-62d059664eaf · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning a Text-Video Embedding from Incomplete and Heterogeneous Data
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aac8fecd-fc60-415d-9bb5-d576ca70a05c · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e5e29eb-be5f-4f25-9112-4a3b710ce7a2 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint embedding with multimodal cues for cross-modal video-text retrieval
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bcaa3663-4b73-45ec-8b97-3e3a89210ea3 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint representations of videos and sentences with web image search
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 23fb74cc-65c8-4b6a-88d3-d2106a05bdee · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Enhancing video summarization via vision-language embed- ding
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9e4c1524-2906-440f-998e-d771ed3b57aa · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2aa1b1a3-a3bd-4aa1-b917-fb7cae7cb855 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Recognizing fine-grained and composite ac- tivities using hand-centric features and script data
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b72f6f25-cffc-4b26-8a4f-4af865fa33aa · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Facenet: A unified embedding for face recognition and clus- tering
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9967b099-b0d6-4ea4-8209-3ee0cbe3c6e2 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Higher-order Network for Action Recognition
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a2eab173-ae32-4ee7-b040-75ac41588ab3 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 888645c1-88be-4105-b163-bea0dd21e166 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improved deep metric learning with multi- class n-pair loss objective
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ffd6b78b-0223-4de0-8d70-c4c8d8e10659 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Cross modal embeddings for video and audio retrieval
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8555b7f9-6dfc-4b21-8d2a-0122fb6050c0 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning Language-Visual Embedding for Movie Understanding with Natural-Language
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df1b3d13-f745-42f3-b12f-0d53586ee99e · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learn- ing fine-grained image similarity with deep ranking
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 939fc6d5-a252-4f04-9e2f-fbf3a1fbdf57 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning deep structure-preserving image-text embeddings
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ebca65a9-0e2d-4061-8e1c-0fc8167dc29a · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Temporal segment networks: Towards good practices for deep action recogni- tion
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1bb59219-bbbf-433e-aa3d-aacdf9ac1e7b · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning visual actions using multiple verb-only labels
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e00a62a1-00b5-42b8-8970-f9002b97fd6a · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Msr-vtt: A large video description dataset for bridging video and language
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 51108899-bda2-4313-823d-5bb99511e90c · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Jointly modeling deep video and compositional text to bridge vision and language in a unified framework
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ecc408dc-a83e-4a8f-8726-c74941c7f6b9 · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings A joint se- quence fusion model for video question answering and re- trieval
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 78072827-6038-4ace-9a0b-2f39734e026c · outbound
Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Zero-shot learning via semantic similarity embedding
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
No inbound Pith citation observations are available.