Pith. sign in

Paper Citation Record · LEDGER

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living

As of 19 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2502.03459.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03459 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:44:58.910397Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a09c5ffe-d7c5-4a20-a2c9-ce6ba00e2cae · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.739680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.739680Z digest=sha256:08eb1fb918b0c9f204f160714162dad9b3bc3cf9c73d0952d68e4fce378021c4

Observation 4f6807ed-6db9-4ec3-98c4-febcecead612 · outbound

This paper cites write newline.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.743392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.743392Z digest=sha256:5dd5504a3c235783164819df3f3a2e33a276ceb1ad1beed1f5382c748738c5f4

Observation cc067b34-46b2-4595-bee6-e9419a29f7f9 · outbound

This paper cites GPT-4 Technical Report.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.747112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.747112Z digest=sha256:2fbc646b78753824e77d6ed1a74b27f17dfc47fcce87235bcf7af6725c780479

Observation f80afc5e-07e6-468e-982f-0eec120e20ca · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.384190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.750548Z digest=sha256:482df452ac7da9518b270b8721acc980e9e72df813065145fb3ce1dc84a1fd38

Observation 45365206-6f49-4d6a-b69d-8c516c560bfb · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.375481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.753636Z digest=sha256:e119cb6dbef35dc0181666dcb38a0aed0004c0b9297f817f6f1a9e07753151f3

Observation f7a4e3ea-b264-435d-b50f-aba1b6f51a87 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.367040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.757401Z digest=sha256:694710ce631145705c66fe315c8982a00e0655994b293cdfa925c21dc873a6b3

Observation b1657822-4607-476a-9942-8927d45b8dff · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.358290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.760473Z digest=sha256:48983d6fcf8ff24916962abdfdf84a3f88f7fa1aab38e0b66bd16494aab9731a

Observation 74124eed-04fb-4b42-aace-8b2990a08336 · outbound

This paper cites E.; Stoica, I.; and Xing, E.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living E.; Stoica, I.; and Xing, E

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.763620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.763620Z digest=sha256:ab68428cd147004d44471ff95fcf54c8fca9f7cdd3e7ea2982e2292f2f897091

Observation a46e8080-61a8-43a2-a237-a1b6a75d6267 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.344668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.766847Z digest=sha256:2a62a19e34a8064c3201267744431219498b86db4461dd5260dc0e6ec5bd5b76

Observation 84ed8ecd-1680-455f-8062-2924e9e12524 · outbound

This paper cites Vision Transformers Need Registers.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Vision Transformers Need Registers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.770111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.770111Z digest=sha256:ca8d511fd56eee54459e12982663fe5a53a8d2431a375dff6114a944731386e0

Observation 9768034e-e4d3-4f33-86a9-e67abcf99ef3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.335786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.774535Z digest=sha256:fd3a9dfc6f674690bbbc24e9b5c904c3664c28c86fe464bb786716fa0336b485

Observation 6c824462-e257-48e1-9d20-41056c32535c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.327639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.777364Z digest=sha256:8f8ebd824ca2137463b658ee765049922faa520be76270790514da842d97542b

Observation 4862d20d-3eb7-45d6-a138-0dc844aa3e0f · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.319320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.780311Z digest=sha256:323e41e2fddf53183687eb7184343143418dae9a86ee76a17726aba8e5a15b21

Observation bc61fd4b-297a-4c8a-b3d0-f1e4cc34aef3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.310476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.783021Z digest=sha256:c3c36b13c84543cb25d774bccf130db6dd38c39e1139fbea59b7c14af4d6f235

Observation 1b8533ee-f3aa-4fc6-a219-9df01b1f972b · outbound

This paper cites K.; Sun, Y.; Patel, P.; and Black, M.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living K.; Sun, Y.; Patel, P.; and Black, M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.301511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.785815Z digest=sha256:08fb0a771a7c756775acbf9b7c8bb3bae5168cc66eb2622024303b026e688e13

Observation 4801c7b1-12b4-448e-a577-15719e8fa9f9 · outbound

This paper cites V.; Joulin, A.; and Misra, I.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living V.; Joulin, A.; and Misra, I

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.788512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.788512Z digest=sha256:bfe85bcd09d0937e3f1e6449c41f4b44720cd98c984c089e5f9687f026a2cc39

Observation 25f595f6-74c1-43a4-b478-16ed7a7bf540 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.287126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.791341Z digest=sha256:b9cbba4a1de7f0cc96f8946a86628b5696c24e59d1605c32b7fd59fbfd2a9a02

Observation 8fcb828b-0ccd-480b-b82b-3bbfb925e32e · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.278740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.794349Z digest=sha256:17a8ad1babf6f611820f3f1a559e1a3caf018e2997c45ea115f38025e4d2ac63

Observation 04f3e484-de05-46b8-91b7-3b8f7a01426c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.270413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.797434Z digest=sha256:d17ce4500b82a05d92dcc10358a5f56d7f6c9c40f5852458000496fdc32b9bcb

Observation 2fb2dafd-00bc-47e2-977d-ad1b43a34929 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Distilling the Knowledge in a Neural Network

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.800197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.800197Z digest=sha256:c689c76d74292c6b681ac4b7c400ffacfb60a47cdba29f8db2a40a1225e66cba

Observation 15bbbb5d-93ec-47fb-964e-e2320581e371 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.261860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.803531Z digest=sha256:2b69a402e7355ea582d54a33394af8db4459a5d4d025cea2f12e51b3ac4fe455

Observation c58284bd-b098-4379-af3f-343171e0e46b · outbound

This paper cites The Kinetics Human Action Video Dataset.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living The Kinetics Human Action Video Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.806472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.806472Z digest=sha256:56ed229f70d2808ee7f5798b08806667b623324123d0a7ad624394b1be21eb37

Observation 5f68b827-35f3-43fd-920a-3bd5c78e0da8 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.252627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.809749Z digest=sha256:78e33c162f638f8abeef1ded3f36e42640e8f90eae4244c9315a9dfe023bf687

Observation f0185f63-076d-4d4e-bb3c-4941aa59afee · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.812526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.812526Z digest=sha256:38fa1a74ede8255d0e5ab59521d3cb372bbfe97da10a7e543666a12e2587e9fb

Observation 03564d78-d7cf-46c8-9c14-5f20059224c2 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.243761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.816041Z digest=sha256:20da88e5d1289537c56a3014bd0e77461904b3ffcf1a6e923f81fc2b07d1c1dc

Observation 18821cef-6e58-4eb4-af69-e4f3a8c919e3 · outbound

This paper cites Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.818749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.818749Z digest=sha256:89dcec32e177dfba8269a81c0fa02959e5c4cb47520b37293845b94d040a8b60

Observation 67c4e852-cfd2-4f61-99a9-f7bc149e8cb3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.234701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.821991Z digest=sha256:8c72526c0979aa2d93580c4e3c93e17fd170ec95f4730d94481c219e0565d5f3

Observation 7c18f423-58a6-4cc6-bb56-b31f4dae4b8e · outbound

This paper cites The Llama 3 Herd of Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living The Llama 3 Herd of Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.824885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.824885Z digest=sha256:81bfc0f73094b0eff764f8f3fd05dc952655856c97c5f1343c96c0e50d41a48e

Observation ff28a171-010d-40ad-8c36-95aef13f6509 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.225734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.827928Z digest=sha256:31988cbdce0c66390c8f90e946d22e3e87f6810aeb0777f6cabad191fc8a990d

Observation 749eccfb-893d-40db-ab5d-dcbd09789a76 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.217434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.830716Z digest=sha256:6e9cabb576edb54390564cac8b3e4c010ddf750f4cc8f633500dec83ae5f9b5b

Observation b78052dd-89a8-4c07-bc01-c7ccf6886ff5 · outbound

This paper cites Multimodal Open-Vocabulary Video Classification via Pre-Trained Vision and Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Multimodal Open-Vocabulary Video Classification via Pre-Trained Vision and Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.833389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.833389Z digest=sha256:702f12c7bb3d42186360d0eef08e40a9f2c4e13ee4e4a51b30f52b6ae85177a4

Observation 49db148f-90d3-4744-b915-d2662f195c39 · outbound

This paper cites W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.836457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.836457Z digest=sha256:e898a7422e355e594d4c1af7500b70c6d888f98fe582179ed5b499e022087d4a

Observation 0525f685-af0a-46c7-b05d-aeb541a94456 · outbound

This paper cites U.; Maaz, M.; Khan, S.; and Khan, F.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living U.; Maaz, M.; Khan, S.; and Khan, F

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.203460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.839261Z digest=sha256:3082eaa57f589d9997e4bde340c6f02f8addfa8ff08c5007a489936736153474

Observation a3eec8cb-faec-43b6-bc86-fae6cf883a5c · outbound

This paper cites LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of Living.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of Living

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T04:44:59.026175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.841940Z digest=sha256:6b505e23a161c0b7e3e41229ed046a43328e0d473f1c3cc719a1331dc975c623

Observation 55c85646-dcd8-473f-b317-6461ed194c75 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.195267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.845095Z digest=sha256:7b7d53c1665e9b51e2e32cbc603f32faee870a3a088bc155dc9b34f8337642c4

Observation dd19305e-74af-4d64-8c16-5ffdbf45c023 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.187474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.847930Z digest=sha256:46365d310919bbf8ed47ec72bd13ef2e053dea5b2e3278afd7df2f63775c874e

Observation 7d778381-38af-498e-91b2-57f95639bc67 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.178768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.850749Z digest=sha256:80bd070a82292f4f1f24297e491a95202de76f42c7b8f894378b878a65b8c0f5

Observation f28fa74d-1357-4fd5-ba01-e31703e02d8b · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.170988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.853607Z digest=sha256:0739d1f5eebe6cc720169838efb1b47cf7cec8f39b3c9722d0c0c610aca38f63

Observation c94a6557-7f97-4894-9c2a-ed49b206b94a · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.162729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.856293Z digest=sha256:09ae738d9e74b321e635e6997e54e8aa97b18106822cd064a5fbfedd050ab27d

Observation d428651c-b0cb-4bde-8702-dc59826bd4d6 · outbound

This paper cites A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.154375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.858982Z digest=sha256:a9e6f1123f2ca30335cc38fc09b03206075a08d3fe3c22fcdf939a0af8560dad

Observation 3c6ceb8f-dc0a-4012-bc1f-c1ed69ce8b63 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.145730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.861661Z digest=sha256:fbffb19e46054cc75b3791466081c37ca9fefb50c44f5b1ce5bb6e059fbca25a

Observation 809ffe0b-5ff6-47f5-86df-a1297608fc55 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.864413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.864413Z digest=sha256:fdc4472dde561471ddad7896df3e06d46852791b76d3bc119300b46c90f16b7e

Observation 9b476259-d7b6-4e9e-afec-275cefff6106 · outbound

This paper cites PandaGPT: One Model To Instruction-Follow Them All.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living PandaGPT: One Model To Instruction-Follow Them All

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.867443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.867443Z digest=sha256:25e22df6863dfcc84db4a934d056625a2393fd5d83e59ba2596969cd3da79707

Observation 7fbbbcde-d7f9-43f9-8396-ae4a38f4a4a1 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.870821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.870821Z digest=sha256:1eaa1ae2e44c3e8cebb225f4534b7636fe84b270b10e6db6546c672c2735f417

Observation 48467c18-75a4-4263-b5f4-cb0eb9e2c9e4 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LLaMA: Open and Efficient Foundation Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.873900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.873900Z digest=sha256:6f6c473b66acf5c588e1a3c3d843253dc97c053d0077f4739aad8e0cf32eabe3

Observation 717a7f45-29f6-4f8d-a1c0-9a0d749f5d34 · outbound

This paper cites ActionCLIP: A New Paradigm for Video Action Recognition.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living ActionCLIP: A New Paradigm for Video Action Recognition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.877082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.877082Z digest=sha256:e42a40ca9246b35d56210a2229e9578447e112be047519e2927a62008e717656

Observation 9f8917dd-0a2a-4a17-889c-e93499148733 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living CogVLM: Visual Expert for Pretrained Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.880265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.880265Z digest=sha256:a16118e67cb652a08b5831b8b9c740e4d536fe074e1a89b89c6b21f4609bd1d2

Observation ea4c7457-1e85-4795-9aef-25e4d221c7f4 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.136172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.883755Z digest=sha256:5465a42b39fe14258d4586866511238f469a4aadfbaa38076bda1075c18515ed

Observation c67b4078-98c3-43da-9dad-e60ba0f55d4f · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.127192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.886845Z digest=sha256:3029f8a848c81f6a1aa7184f52206f2f21bcd15bdc7ab347731e9e5f0c76f176

Observation c24f294b-10e9-4497-82aa-61c797d398aa · outbound

This paper cites CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.889585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.889585Z digest=sha256:f7a9b9b785e282f83c1075ac0c1792e9b9a3cb97704c41eb6cc13ccdec96d136

Observation 5767af60-c35c-4b60-a239-ae281cc10ef7 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.117989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.892781Z digest=sha256:0ac6933e6b31497a88d79a499440fa860b1050b2d237dfe664e1d3527b31df84

Observation dafa6a34-9a96-4a51-b004-cbb2b41f7904 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.895795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.895795Z digest=sha256:d9fe763e5979be6630a89421d9ee5b9521382cc69db25d5a660de88312cfb645

Observation fe6622ee-4080-42c8-9e88-a4f32c5ec3cb · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.898707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.898707Z digest=sha256:593138ddccf426172a472232e00b3b55b6868acb86677d0f334bf2b4854dba0c

Observation 9602cfe7-07a7-48d8-836c-4fcd743cae2c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.108218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.901971Z digest=sha256:a4c6f7e70010ca6b89c1a9fed151ad1caa6943cab6976d6d23e79471345abe10

Observation 91f593e9-4257-4750-9abc-3b818c95ddbc · outbound

This paper cites Hypergraph Transformer for Skeleton-based Action Recognition.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Hypergraph Transformer for Skeleton-based Action Recognition

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.904658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.904658Z digest=sha256:83018cfe5f5959ec2c977e92df5d2e45f6124922c56820cae104294e94a2dfac

Observation 8cae0090-2e98-432c-8121-24a2515bc481 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.099260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.907612Z digest=sha256:c2b5b52ceb581cec5fcc49f58c0182a5c0fce8242455b8a22ca25a6bb2876996

Observation 5d6eaf64-e9e8-4b4f-931f-72eaeb08b46f · outbound

This paper cites LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.910397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.910397Z digest=sha256:1c94f886db51b25eaf05d244ec36a8b02d4da3992ae1d99663bfcd579fc20e8e

Pith citing papers

No inbound Pith citation observations are available.