Pith. sign in

Paper Citation Record · LEDGER

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings

As of 16 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:1908.03477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.03477 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:15:57.272572Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy30
  • unresolved9
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09347ffc-8b47-4a24-89f0-7dcf42b6192e · outbound

This paper cites an unresolved cited work.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:15:58.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.015401Z digest=sha256:a71146dfccb5627e48b3eb32d29bfaba4ea954e39c0053d499dc7f6acfa2e921

Observation 70e071cc-fe06-4f05-8010-266a3834f327 · outbound

This paper cites Re-ID done right: towards good practices for person re-identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Re-ID done right: towards good practices for person re-identification

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.023091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.023091Z digest=sha256:77045c83c0bd9e568e7c276782963da5b193c4b50e43a9cf87c61491d24a0c4e

Observation a7526b53-9e57-488c-8b35-236a261a4bca · outbound

This paper cites NetVLAD: CNN architecture for weakly supervised place recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings NetVLAD: CNN architecture for weakly supervised place recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.187494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.029900Z digest=sha256:c38eb644dc45c472d6fded8d38ec89a1e010b753d26e72e067b4b501fc06001f

Observation eb52af3d-3229-4db1-b3f3-9c81d452bc61 · outbound

This paper cites An empirical study and analysis of generalized zero- shot learning for object recognition in the wild.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings An empirical study and analysis of generalized zero- shot learning for object recognition in the wild

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.160419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.036174Z digest=sha256:6a8563c373db16a392ca569bef5c41129a105aa98260b413c95c28ed757e6cea

Observation 40f1f8b8-8e51-435a-8b4e-ce75457b613a · outbound

This paper cites Beyond triplet loss: a deep quadruplet network for person re-identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond triplet loss: a deep quadruplet network for person re-identification

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.139875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.043008Z digest=sha256:468f7faee6e1e0133f7aa6f92296ad5b3c6022f96f52ec862d42ee12852be0e1

Observation 0b862c3b-5658-46a5-81f0-01806def26c8 · outbound

This paper cites Scaling egocentric vision: The epic-kitchens dataset.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Scaling egocentric vision: The epic-kitchens dataset

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.120942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.055743Z digest=sha256:f34c945376af5eb8b5140cfc0484b226f4beb1b375a6ebd65c49c4ad382b7930

Observation c4095747-5b8f-4326-9d83-56a92a40de35 · outbound

This paper cites Predict- ing visual features from text for image and video caption re- trieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Predict- ing visual features from text for image and video caption re- trieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.103558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.061105Z digest=sha256:d79b704c5746e356c72bfae5494890898ad9ba4910adf32ebbddfc23f9970e64

Observation 86de2abf-eb5e-4198-8626-0b4c4a2c1b66 · outbound

This paper cites Dual Encoding for Zero-Example Video Retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Dual Encoding for Zero-Example Video Retrieval

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:15:57.473760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.065902Z digest=sha256:5af71e17c74e520bf2ee352519baf35f83d75c3cba36a8cf27d8d88a0a855fa8

Observation 91628ae4-1ca5-4ff3-a5a9-f518899f9046 · outbound

This paper cites Improving image-sentence embeddings using large weakly annotated photo collections.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improving image-sentence embeddings using large weakly annotated photo collections

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.086177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.071958Z digest=sha256:0c01e3eedd221ab2a80cd2b6697fce2c7cf5463888d4a395e3200e47f9f8db20

Observation 24c39c6a-2d93-41d4-ab83-5847f64df17e · outbound

This paper cites Deep image retrieval: Learning global representations for image search.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep image retrieval: Learning global representations for image search

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.065837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.077021Z digest=sha256:ef0aa0ba24e2b64713f2c9fbd408f543facc68025f879582f6f1e6df18c16729

Observation 437a4b4c-65a0-4334-91a8-3b7703c905ea · outbound

This paper cites Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Beyond instance-level im- age retrieval: Leveraging captions to learn a global visual representation for semantic retrieval

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.046935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.082890Z digest=sha256:c9958467f6137f59c2eebea193a6ffa477fb92e1be85a45db1ff1c3cef2dcf09

Observation 8356357a-0ab1-4ed0-a846-89dbbf0fb2a7 · outbound

This paper cites Something Something.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Something Something

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.028366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.088103Z digest=sha256:583104437052cf2a6c375e3148eb670d98d917f8daee5b206ce067e0ab08f08d

Observation 48d26ae9-a8df-4a56-a7a8-807789c60ed5 · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Ava: A video dataset of spatio-temporally localized atomic visual actions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:58.009196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.093921Z digest=sha256:49e34d6ea1b9818d4e7bc0f7729d60dce681cca3fa265e2037b1d03f09f21253

Observation ce2de66e-540f-4bbb-87b6-0d4b12c9a203 · outbound

This paper cites Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.985015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.099943Z digest=sha256:e7c3d0629c87b29d2f81c95a8c14aeefe039111a1df8fde155ea4124561d1620

Observation b1399349-65c3-4448-af8d-9686b1960d0c · outbound

This paper cites an unresolved cited work.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:15:57.966058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.106075Z digest=sha256:29160201121058b9f7ac1de47f48e4eca4847943f2e7ab76fc41429f955dffd3

Observation 47c7d464-2874-4da8-b38a-bda26323ec06 · outbound

This paper cites In Defense of the Triplet Loss for Person Re-Identification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings In Defense of the Triplet Loss for Person Re-Identification

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.111220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.111220Z digest=sha256:79b99b7556ce5a5f3eea0547d4f85fb6118727ad2ae31c4e634739770c2e0899

Observation cce0550f-0af1-45af-bb9b-8b78364b9faa · outbound

This paper cites Deep metric learning using triplet network.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Deep metric learning using triplet network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.945593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.116903Z digest=sha256:a642770e9d8c1bd13364511cd835043eb3d5cff90e26284511e3ac2cb63646e4

Observation 46f38826-3731-4783-a1ae-56bcc6ed91d4 · outbound

This paper cites Unifying visual-semantic embeddings with multimodal neu- ral language models.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Unifying visual-semantic embeddings with multimodal neu- ral language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.927132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.122863Z digest=sha256:9ecff00e38ecaae093c371d973ad761212f659f72751733f397a00bee1563d53

Observation 6d57a355-c19c-4736-9c72-0db32e089aa0 · outbound

This paper cites Learning two-branch neural networks for image-text match- ing tasks.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning two-branch neural networks for image-text match- ing tasks

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.908471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.127947Z digest=sha256:84861c5975766af937800d48957b0e805c86b5453aa809ca2cce468c8c0e1d85

Observation 165f5939-6426-4545-b699-ce487380717f · outbound

This paper cites On the effectiveness of task granularity for transfer learning.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings On the effectiveness of task granularity for transfer learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.133392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.133392Z digest=sha256:00627d920bf39c9a47be98b5c7e5e21c2065c971c61d7ccead908582ef91e673

Observation 611c895f-318a-438d-82f6-b8eef42c219e · outbound

This paper cites Learnable pooling with Context Gating for video classification.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learnable pooling with Context Gating for video classification

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.139487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.139487Z digest=sha256:aa90f20069e8b906dca57d1e4ad8aa0cf7340f6a0e775300275efb900d75587c

Observation ad39cbdd-ded8-4ef3-a79b-62d059664eaf · outbound

This paper cites Learning a Text-Video Embedding from Incomplete and Heterogeneous Data.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning a Text-Video Embedding from Incomplete and Heterogeneous Data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.144801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.144801Z digest=sha256:5c3c0703d18734d86b1dcac4c4720e7961505faa1727b27c464c71163b021035

Observation aac8fecd-fc60-415d-9bb5-d576ca70a05c · outbound

This paper cites HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.150310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.150310Z digest=sha256:76622851e15abcffa77441c564ab72976e9aab0742265fddf7823130d7f4a2db

Observation 8e5e29eb-be5f-4f25-9112-4a3b710ce7a2 · outbound

This paper cites Learning joint embedding with multimodal cues for cross-modal video-text retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint embedding with multimodal cues for cross-modal video-text retrieval

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.878011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.157125Z digest=sha256:784ebb131e5733e5d00073191d1d2b60eb3b3b92967308ed0dc363932d22fcfc

Observation bcaa3663-4b73-45ec-8b97-3e3a89210ea3 · outbound

This paper cites Learning joint representations of videos and sentences with web image search.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning joint representations of videos and sentences with web image search

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.859538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.162683Z digest=sha256:afbed51bc20e7ce50050bf6595b156a3a06bb5f8cf3bfb8289868ec0b9365965

Observation 23fb74cc-65c8-4b6a-88d3-d2106a05bdee · outbound

This paper cites Enhancing video summarization via vision-language embed- ding.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Enhancing video summarization via vision-language embed- ding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.839909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.168392Z digest=sha256:c0a421a76c70e742abc0892b9afb27b0e60e3199479ede7eb3518c6366384e05

Observation 9e4c1524-2906-440f-998e-d771ed3b57aa · outbound

This paper cites CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings CNN image retrieval learns from BoW: Unsupervised fine-tuning with hard examples

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.818744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.174858Z digest=sha256:0714eecf5367c610b8b8afca78ae210fafa4407278eb59f8a6ec0038f42011b1

Observation 2aa1b1a3-a3bd-4aa1-b917-fb7cae7cb855 · outbound

This paper cites Recognizing fine-grained and composite ac- tivities using hand-centric features and script data.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Recognizing fine-grained and composite ac- tivities using hand-centric features and script data

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.797187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.181674Z digest=sha256:d9298e7acd6f22f1bf7e9b80d209d2f337811e8157cca3a4585de907810095a8

Observation b72f6f25-cffc-4b26-8a4f-4af865fa33aa · outbound

This paper cites Facenet: A unified embedding for face recognition and clus- tering.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Facenet: A unified embedding for face recognition and clus- tering

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.773385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.190379Z digest=sha256:a13ec30dbd717c448ec0ef66849b7bd20909bea2d1fa8534c0e7181b4f1ef49b

Observation 9967b099-b0d6-4ea4-8209-3ee0cbe3c6e2 · outbound

This paper cites Higher-order Network for Action Recognition.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Higher-order Network for Action Recognition

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:15:57.347997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.195289Z digest=sha256:3393e80f47051d58965eb5455fb6e371e8c1dcf5f1ec0504488c4246f737b914

Observation a2eab173-ae32-4ee7-b040-75ac41588ab3 · outbound

This paper cites Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Sigurdsson, G ¨ul Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.745079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.201647Z digest=sha256:b81290f34a1922ec4c80064fddd0e69e610a84b9b16ce04a3280e86374abf186

Observation 888645c1-88be-4105-b163-bea0dd21e166 · outbound

This paper cites Improved deep metric learning with multi- class n-pair loss objective.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Improved deep metric learning with multi- class n-pair loss objective

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.726059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.207509Z digest=sha256:0bf6be0cf99151b128ccd4464ecec65325b20e02190fbaaec47a5f7a18bfbf6c

Observation ffd6b78b-0223-4de0-8d70-c4c8d8e10659 · outbound

This paper cites Cross modal embeddings for video and audio retrieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Cross modal embeddings for video and audio retrieval

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.705322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.213105Z digest=sha256:ebc20c5bf9d7168f8024fe7486b20be0942f6862792365bdf2ddf5a6e3f93ee0

Observation 8555b7f9-6dfc-4b21-8d2a-0122fb6050c0 · outbound

This paper cites Learning Language-Visual Embedding for Movie Understanding with Natural-Language.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning Language-Visual Embedding for Movie Understanding with Natural-Language

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T14:15:57.218372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:15:57.218372Z digest=sha256:b69108ca4bfbde242bba2ef13b6a29bb7eb0f79d0f042745b0ef06137363f378

Observation df1b3d13-f745-42f3-b12f-0d53586ee99e · outbound

This paper cites Learn- ing fine-grained image similarity with deep ranking.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learn- ing fine-grained image similarity with deep ranking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.683785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.224034Z digest=sha256:a7e214efed335e19cc316ac1450ea270b6c63f07fbf3c61bb58e155390f3e3f1

Observation 939fc6d5-a252-4f04-9e2f-fbf3a1fbdf57 · outbound

This paper cites Learning deep structure-preserving image-text embeddings.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning deep structure-preserving image-text embeddings

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.653864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.230640Z digest=sha256:6228d2892eb8a567dfe7800d65c51e4676eebc13288ab8bfa3cfe7aa4b706d77

Observation ebca65a9-0e2d-4061-8e1c-0fc8167dc29a · outbound

This paper cites Temporal segment networks: Towards good practices for deep action recogni- tion.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Temporal segment networks: Towards good practices for deep action recogni- tion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.634381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.237032Z digest=sha256:b9a2088bb7c1712c38eb2015c4260c27da202aaa97251ac88df7c122cefdf34b

Observation 1bb59219-bbbf-433e-aa3d-aacdf9ac1e7b · outbound

This paper cites Learning visual actions using multiple verb-only labels.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Learning visual actions using multiple verb-only labels

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.611464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.243318Z digest=sha256:e203a643ec1462418fc2c99e8152141653fdd9b10ad874bbf2e177c512864c0f

Observation e00a62a1-00b5-42b8-8970-f9002b97fd6a · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Msr-vtt: A large video description dataset for bridging video and language

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.591146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.249711Z digest=sha256:64c4378f6cb4b5e5e2d8f91df3a831044dfc6f156a8f61d951d7ea85310a82f6

Observation 51108899-bda2-4313-823d-5bb99511e90c · outbound

This paper cites Jointly modeling deep video and compositional text to bridge vision and language in a unified framework.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Jointly modeling deep video and compositional text to bridge vision and language in a unified framework

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.556663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.255506Z digest=sha256:a6629949ac3de71492567c28263fd1b0cfa5a0ceceac4267c852e9416e0c3c2a

Observation ecc408dc-a83e-4a8f-8726-c74941c7f6b9 · outbound

This paper cites A joint se- quence fusion model for video question answering and re- trieval.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings A joint se- quence fusion model for video question answering and re- trieval

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:15:57.533791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.262176Z digest=sha256:84fb87f31d7eea39e3a9100a358ef77294657f3501fd832a25e8dfb04d1c68bb

Observation 78072827-6038-4ace-9a0b-2f39734e026c · outbound

This paper cites Zero-shot learning via semantic similarity embedding.

Fine-Grained Action Retrieval Through Multiple Parts-of-Speech Embeddings Zero-shot learning via semantic similarity embedding

Reference 42

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T14:15:57.515796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:15:57.272572Z digest=sha256:fbfefac853faba38fc75d8a2e9323c39c3934a5f9533624967aa13eee4563d90

Pith citing papers

No inbound Pith citation observations are available.