Pith. sign in

Paper Citation Record · LEDGER

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2502.03459.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03459 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:44:58.910397Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a09c5ffe-d7c5-4a20-a2c9-ce6ba00e2cae · outbound

This paper cites , " * write output.state after.block = add.period write newline.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.739680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.739680Z digest=sha256:3b36e83355550a8222513a5b23ec825b89f67db07f56756633efbc2acc89fbf6

Observation 4f6807ed-6db9-4ec3-98c4-febcecead612 · outbound

This paper cites write newline.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.743392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.743392Z digest=sha256:cff006216683d4136d07e4711b51ca2cdb89e3ac3b2529b0a9a35e1ba815c542

Observation cc067b34-46b2-4595-bee6-e9419a29f7f9 · outbound

This paper cites GPT-4 Technical Report.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.747112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.747112Z digest=sha256:b47b468761a0b287ab79cfdcad347bf2abe3306f132be43b60c203f1e5fbf23b

Observation f80afc5e-07e6-468e-982f-0eec120e20ca · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.384190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.750548Z digest=sha256:95576bf4962f32df105f883621034d42cdefd878d6055675c3ef164eecd292cc

Observation 45365206-6f49-4d6a-b69d-8c516c560bfb · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.375481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.753636Z digest=sha256:ba481eeb16ca995606e199b71d1a3d0db4023d8c3c0665a405737a4eb36b7daa

Observation f7a4e3ea-b264-435d-b50f-aba1b6f51a87 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.367040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.757401Z digest=sha256:c9b19c8d6553b98821e1a2ee5de1411e2b84b6eb3caa4f317eaa8d7c4bff579c

Observation b1657822-4607-476a-9942-8927d45b8dff · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.358290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.760473Z digest=sha256:bf38cd14ec0ce4cce70d0d5921a7ca6cc0896e7bb2850b89dd59ec2448fc9abc

Observation 74124eed-04fb-4b42-aace-8b2990a08336 · outbound

This paper cites E.; Stoica, I.; and Xing, E.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living E.; Stoica, I.; and Xing, E

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.763620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.763620Z digest=sha256:3a01f9fd6829cc3028a9cb7b2a0e1edfd19941a14b9e10bedbf69d460f2a153b

Observation a46e8080-61a8-43a2-a237-a1b6a75d6267 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.344668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.766847Z digest=sha256:20b0f402a245751c2e5ae00f61ed2b00c95f7eeff0da6fa2c8dc1036972def24

Observation 84ed8ecd-1680-455f-8062-2924e9e12524 · outbound

This paper cites Vision Transformers Need Registers.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Vision Transformers Need Registers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.770111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.770111Z digest=sha256:90f58a755dfa1b99bc95a9ec1ab6abafaedde999ee7ea724b042df24cfafd066

Observation 9768034e-e4d3-4f33-86a9-e67abcf99ef3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.335786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.774535Z digest=sha256:b5483aef73217141afb5362f3325cf8d4b274b4e8d80c088d16e686b0bf1f8d2

Observation 6c824462-e257-48e1-9d20-41056c32535c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.327639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.777364Z digest=sha256:cc09f43abb19a0ad6c3ab7f901cc74086f9fa74a9c9aa8ed5c694afe31ba127e

Observation 4862d20d-3eb7-45d6-a138-0dc844aa3e0f · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.319320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.780311Z digest=sha256:bea16f0af28808da49453d3b3425cbf462e5a2dc02573fca241fc211b050669b

Observation bc61fd4b-297a-4c8a-b3d0-f1e4cc34aef3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.310476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.783021Z digest=sha256:b843a8d9d72b84015475656fbf78fc7a9be06b93bbecdfc89fa945686ca61b7f

Observation 1b8533ee-f3aa-4fc6-a219-9df01b1f972b · outbound

This paper cites K.; Sun, Y.; Patel, P.; and Black, M.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living K.; Sun, Y.; Patel, P.; and Black, M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.301511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.785815Z digest=sha256:e9d4801cae07fc971d6e255d89e1a54cb0ef498248a790a4e3b5ff0ff9b4bc1d

Observation 4801c7b1-12b4-448e-a577-15719e8fa9f9 · outbound

This paper cites V.; Joulin, A.; and Misra, I.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living V.; Joulin, A.; and Misra, I

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.788512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.788512Z digest=sha256:0e26f932d826b14a664fcec493ea74cfb0bbe5fd7839a3cf8aaeb5f79c428df6

Observation 25f595f6-74c1-43a4-b478-16ed7a7bf540 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.287126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.791341Z digest=sha256:af3d395bca76fb3f9ba2b44d5a8287b9bdb4d02d50b83eab2c3832073c03c1fc

Observation 8fcb828b-0ccd-480b-b82b-3bbfb925e32e · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.278740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.794349Z digest=sha256:30f24b2a1376d448ab9e764018741a83e401ff580211feba2e5317d467d4125e

Observation 04f3e484-de05-46b8-91b7-3b8f7a01426c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.270413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.797434Z digest=sha256:e4426691378f8d77419dff4d4b90ac877a182bf79e31b3f594b1f4a0dce5631f

Observation 2fb2dafd-00bc-47e2-977d-ad1b43a34929 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Distilling the Knowledge in a Neural Network

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.800197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.800197Z digest=sha256:958c32df3677977d132352e5142d4b7da953d62822cd7a080d85319d04551e3b

Observation 15bbbb5d-93ec-47fb-964e-e2320581e371 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.261860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.803531Z digest=sha256:8799fbff11a854ba5d6e3e28bc6dd8a103eaed08d509a170938d84dcd73354ca

Observation c58284bd-b098-4379-af3f-343171e0e46b · outbound

This paper cites The Kinetics Human Action Video Dataset.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living The Kinetics Human Action Video Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.806472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.806472Z digest=sha256:7a049e243752030bdf83f08c1a20b4203bc236dfafd56d5be2d6b67c555eb33a

Observation 5f68b827-35f3-43fd-920a-3bd5c78e0da8 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.252627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.809749Z digest=sha256:75afb3e007fe3073e02623d328beae423420ac73c265f967d9c2c74b47a59169

Observation f0185f63-076d-4d4e-bb3c-4941aa59afee · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.812526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.812526Z digest=sha256:3110f428d1bca72e1a7b19aa733e20959e28ed7c0dd219bbe233d30576108526

Observation 03564d78-d7cf-46c8-9c14-5f20059224c2 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.243761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.816041Z digest=sha256:60719d78d5c8816908260d304826ef7d5de39223e3e176aa3e0b71c604d5d969

Observation 18821cef-6e58-4eb4-af69-e4f3a8c919e3 · outbound

This paper cites Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.818749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.818749Z digest=sha256:e0a0f4471e5c99d61d9047cda64bdfb1f7d206dd7e2ef935c74db9b810264748

Observation 67c4e852-cfd2-4f61-99a9-f7bc149e8cb3 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.234701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.821991Z digest=sha256:7e49504605c796e28f8a5f799c1ca2071519990d518fa64aa68cc29b4a5897ab

Observation 7c18f423-58a6-4cc6-bb56-b31f4dae4b8e · outbound

This paper cites The Llama 3 Herd of Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living The Llama 3 Herd of Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.824885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.824885Z digest=sha256:8e4d9206d46d3c30f038e7ce4754f3f8a11d6514bc91a3d0c18b779a36cc9bc9

Observation ff28a171-010d-40ad-8c36-95aef13f6509 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.225734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.827928Z digest=sha256:2560868e81dc7f1eb7f4c1ec744d7a48be70514c31f9a3d333c8a4d96cb56048

Observation 749eccfb-893d-40db-ab5d-dcbd09789a76 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.217434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.830716Z digest=sha256:21f08456fd55d0db9f0f3d437d6334d026742cc92b6e6bdbe4a1e7b48e543b6c

Observation b78052dd-89a8-4c07-bc01-c7ccf6886ff5 · outbound

This paper cites Multimodal Open-Vocabulary Video Classification via Pre-Trained Vision and Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Multimodal Open-Vocabulary Video Classification via Pre-Trained Vision and Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.833389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.833389Z digest=sha256:a6ca046743bd6b469c65b6faffe843f78882f7b345c58aa558a4ce17c296986f

Observation 49db148f-90d3-4744-b915-d2662f195c39 · outbound

This paper cites W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.836457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.836457Z digest=sha256:3c8866d6d4af2728b042827966333ec7e8061e312fab28828a745dc5df98fa93

Observation 0525f685-af0a-46c7-b05d-aeb541a94456 · outbound

This paper cites U.; Maaz, M.; Khan, S.; and Khan, F.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living U.; Maaz, M.; Khan, S.; and Khan, F

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.203460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.839261Z digest=sha256:13ab517d5ae7af4eaa70820e8754db866852323dc28c9fd1d3c1513232a85b06

Observation a3eec8cb-faec-43b6-bc86-fae6cf883a5c · outbound

This paper cites LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of Living.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LLAVIDAL: A Large LAnguage VIsion Model for Daily Activities of Living

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T04:44:59.026175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.841940Z digest=sha256:cc6c00430981c9e25d305ec107cb25acb900fd248f239d8dc3572e51cf0f3568

Observation 55c85646-dcd8-473f-b317-6461ed194c75 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.195267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.845095Z digest=sha256:945164510bbc6cdce6e770b4e3765dcaac17a5741d9df124650ebffd9985357d

Observation dd19305e-74af-4d64-8c16-5ffdbf45c023 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.187474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.847930Z digest=sha256:95f014a03e1db8e53402e7d5404b1eb5ee09f43affb4c13fd828535a9b2bc621

Observation 7d778381-38af-498e-91b2-57f95639bc67 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.178768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.850749Z digest=sha256:bbaed09f0c3179ab7efdae3ea1ef1c33da32640b0ecc8b1ebf32710ebd00229c

Observation f28fa74d-1357-4fd5-ba01-e31703e02d8b · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.170988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.853607Z digest=sha256:731e12283a6ac29051bc1e5d46a685e17cd76b1253fb783b746ee87f708d144a

Observation c94a6557-7f97-4894-9c2a-ed49b206b94a · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.162729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.856293Z digest=sha256:29d9d24437f4b2995195ca71258fd33744dd6cf8dce9a6607464f719b3d2899a

Observation d428651c-b0cb-4bde-8702-dc59826bd4d6 · outbound

This paper cites A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:44:59.154375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.858982Z digest=sha256:3d9fde654f295f7c800616fdc332a38774bb9f6d098b0926356aa250b65c918b

Observation 3c6ceb8f-dc0a-4012-bc1f-c1ed69ce8b63 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.145730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.861661Z digest=sha256:b682ecb6cbd3c7077f1a5d25bf6feb7ad67d9fb738630e7b08ba04112b52164c

Observation 809ffe0b-5ff6-47f5-86df-a1297608fc55 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.864413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.864413Z digest=sha256:402c0243a2e7142d94e57b138c20e356815e7c69761ac57e453faca588afd293

Observation 9b476259-d7b6-4e9e-afec-275cefff6106 · outbound

This paper cites PandaGPT: One Model To Instruction-Follow Them All.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living PandaGPT: One Model To Instruction-Follow Them All

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.867443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.867443Z digest=sha256:6798189649e4579ff954ae53b51e12d1769173608bb574b95299f0b3113f8163

Observation 7fbbbcde-d7f9-43f9-8396-ae4a38f4a4a1 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.870821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.870821Z digest=sha256:fdf4c6e1ce127c8fbc371e06786f6b32ff1433f33497d6fd525f6ce1b33b8bb5

Observation 48467c18-75a4-4263-b5f4-cb0eb9e2c9e4 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LLaMA: Open and Efficient Foundation Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.873900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.873900Z digest=sha256:56eecd6d8b69eb83fed6dde372439fc9535f32880d539a5f112dadd418c4f6f5

Observation 717a7f45-29f6-4f8d-a1c0-9a0d749f5d34 · outbound

This paper cites ActionCLIP: A New Paradigm for Video Action Recognition.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living ActionCLIP: A New Paradigm for Video Action Recognition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.877082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.877082Z digest=sha256:e8eed75ed4671ed1e53bc1a1d65b65fa27191c6fdfe48a16f2e54f1b844fa8d3

Observation 9f8917dd-0a2a-4a17-889c-e93499148733 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living CogVLM: Visual Expert for Pretrained Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.880265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.880265Z digest=sha256:6fa8b58543b7d9d9b2e6d86a2ce1cb670378a977ebd793e116d54aa88fbf40a2

Observation ea4c7457-1e85-4795-9aef-25e4d221c7f4 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.136172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.883755Z digest=sha256:491fd0728028a362760ede8b9fdcb1e1be3bbf71a13c32492ea2a40d097e57c4

Observation c67b4078-98c3-43da-9dad-e60ba0f55d4f · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.127192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.886845Z digest=sha256:4970eb56d47500d5c6da99a76c188f40af01237e3d19a0f29311c6ece2ce67e8

Observation c24f294b-10e9-4497-82aa-61c797d398aa · outbound

This paper cites CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.889585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.889585Z digest=sha256:4f5d83c41778ed9f94b937d66f15b2fa2dff9c45b1007af5c02b58cfade5d9bd

Observation 5767af60-c35c-4b60-a239-ae281cc10ef7 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.117989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.892781Z digest=sha256:42b918784d2d7937c1bac83febacdaf02057825a170b64112b800a8571f3f880

Observation dafa6a34-9a96-4a51-b004-cbb2b41f7904 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.895795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.895795Z digest=sha256:33d9d52407ea988706b40c03139beacd3bad4f028fde49e8e7fc472190116353

Observation fe6622ee-4080-42c8-9e88-a4f32c5ec3cb · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.898707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.898707Z digest=sha256:d7e56e8e751bbdf9f43cb374f5f2601954e2dfdba9d71b9d7048bb3559521d5a

Observation 9602cfe7-07a7-48d8-836c-4fcd743cae2c · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.108218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.901971Z digest=sha256:3843458ea83617f7db104ba9e38bd163daa796a22aa31430347d4e7f4d5874e7

Observation 91f593e9-4257-4750-9abc-3b818c95ddbc · outbound

This paper cites Hypergraph Transformer for Skeleton-based Action Recognition.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Hypergraph Transformer for Skeleton-based Action Recognition

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.904658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.904658Z digest=sha256:42e2411b39b7aa701ef214a5f094d3b384cff64078428690875feb9a08c7ffe6

Observation 8cae0090-2e98-432c-8121-24a2515bc481 · outbound

This paper cites an unresolved cited work.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:44:59.099260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T04:44:58.907612Z digest=sha256:1ec32ffee3911e1e8e113b097de4914bd3e624487b4dfd2cef4e9b55d29fc8dc

Observation 5d6eaf64-e9e8-4b4f-931f-72eaeb08b46f · outbound

This paper cites LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment.

SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T04:44:58.910397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:44:58.910397Z digest=sha256:8bfa27e632d0557833c5c10d424ca7cfdcb619851fda1099e487da5b43951d68

Pith citing papers

No inbound Pith citation observations are available.