Pith. sign in

Paper Citation Record · LEDGER

Efficient Egocentric Action Recognition with Multimodal Data

As of 21 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2506.01757.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01757 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:39:47.435799Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3f3e0ee6-6946-4e6a-8ae8-64d936fff3e0 · outbound

This paper cites Symmetric sub-graph spatio-temporal graph convolution and its application in complex activity recognition.

Efficient Egocentric Action Recognition with Multimodal Data Symmetric sub-graph spatio-temporal graph convolution and its application in complex activity recognition

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:40:11.807748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:38:50.051655Z digest=sha256:029fdf802643d438296458b21c14b23da47897e64d5e3f1e58de038616fff2f0

Observation b11ea7eb-3137-4d39-be53-5b9c5ac7ff03 · outbound

This paper cites Avatars grow legs: Generating smooth human motion from sparse tracking in- puts with diffusion model.

Efficient Egocentric Action Recognition with Multimodal Data Avatars grow legs: Generating smooth human motion from sparse tracking in- puts with diffusion model

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:40:11.502806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:38:51.029276Z digest=sha256:a6c826addd53d22b54d6771c3805584e731682152a2ffa644d179c50f4697e46

Observation c645f8be-a0a5-4b0b-8ebc-828cbfe46ec9 · outbound

This paper cites What would you expect? anticipating egocentric actions with rolling- unrolling lstms and modality attention.

Efficient Egocentric Action Recognition with Multimodal Data What would you expect? anticipating egocentric actions with rolling- unrolling lstms and modality attention

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:40:09.939297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:38:53.660511Z digest=sha256:84441a48da9f5a7f71c59d1ebc53d04444f64c22caf25683a359430c4c2c3893

Observation 1ed631a9-c3de-4ec2-b618-7cd3749929bc · outbound

This paper cites Levit: a vision transformer in convnet’s clothing for faster inference.

Efficient Egocentric Action Recognition with Multimodal Data Levit: a vision transformer in convnet’s clothing for faster inference

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:40:09.869444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.246685Z digest=sha256:036a220932a555ad07e04c477cf8113531cf63b5bad00b14a68d49960324ca91

Observation 7ce97a0e-7f1d-4fc4-b3c3-70caf2ad2627 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Efficient Egocentric Action Recognition with Multimodal Data Distilling the Knowledge in a Neural Network

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:39:46.341840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:39:46.341840Z digest=sha256:4a18e0cab38b507ab4865763dd71b1b094ee18890d0d5c924b3d1df1315ea2df

Observation f38d1923-78ed-4f6d-9e7c-390caa507590 · outbound

This paper cites Object detection- based location and activity classification from egocentric videos: A systematic analysis.

Efficient Egocentric Action Recognition with Multimodal Data Object detection- based location and activity classification from egocentric videos: A systematic analysis

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:40:09.643750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.439381Z digest=sha256:ae10d0eee4c24625d6ce3b4801bbe2db007301e3946fde1f96f6e82eda963d2b

Observation 516502a6-3c90-42da-b8d9-ea573d4596ac · outbound

This paper cites H2o: Two hands manipulating objects for first person interaction recognition.

Efficient Egocentric Action Recognition with Multimodal Data H2o: Two hands manipulating objects for first person interaction recognition

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:40:08.044797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.532798Z digest=sha256:7d7f5386cc56ae8504865c28cc84fb29be6b2222a7df1540aa119c5d948d74f6

Observation a7c229e4-8dfb-4aad-9a81-d226cd4a482e · outbound

This paper cites an unresolved cited work.

Efficient Egocentric Action Recognition with Multimodal Data Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:39:53.375353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.600410Z digest=sha256:9960042a05b882b5c580aaa7fc01d3f1ae64f96b8a316abe812a0a339873af13

Observation 2dc08446-bfa0-4498-86d4-e5cc2876e09c · outbound

This paper cites Detecting activi- ties of daily living in first-person camera views.

Efficient Egocentric Action Recognition with Multimodal Data Detecting activi- ties of daily living in first-person camera views

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:53.018600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.703916Z digest=sha256:71be0b18b978f2f5b5b63cacad02025e65f3325684b2cd9599c835d0732ff3af

Observation 4c34fcf6-c1e3-4892-a001-8cb14f8a23dd · outbound

This paper cites Designing network design spaces.

Efficient Egocentric Action Recognition with Multimodal Data Designing network design spaces

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:39:46.785014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:39:46.785014Z digest=sha256:fbbc405df98a465739dd045949d653928f111415aa2f6a772de73b774b6c08a6

Observation b9396302-902f-4b2f-b10f-1993a39383de · outbound

This paper cites On the utility of 3d hand poses for action recognition.

Efficient Egocentric Action Recognition with Multimodal Data On the utility of 3d hand poses for action recognition

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:52.771809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.848276Z digest=sha256:1ef3bf50887668edb395d66deaffcb72e61c9601290a5665b3d7f808bc4dbf76

Observation b0a4d25c-844a-433c-ab87-95ebc459548c · outbound

This paper cites Convolutional lstm network: A machine learning approach for precipitation nowcasting.

Efficient Egocentric Action Recognition with Multimodal Data Convolutional lstm network: A machine learning approach for precipitation nowcasting

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:52.465269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.916905Z digest=sha256:2adf6a66ebd4cd4512981719324b495710a5f3b870e1f11f2a0aced80b1ddee8

Observation 47fdd9d5-f6e6-4d0a-aa23-6fcfd1f2c9bc · outbound

This paper cites Two-stream con- volutional networks for action recognition in videos.

Efficient Egocentric Action Recognition with Multimodal Data Two-stream con- volutional networks for action recognition in videos

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:52.330510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.953874Z digest=sha256:55ac16fcb223f52d6630c2d38784745bdf24badb2375caaf2d5a81fafbb6a417

Observation f23fc727-1b6f-4e00-accb-497998304f5d · outbound

This paper cites Convolutional long short-term memory networks for recognizing first per- son interactions.

Efficient Egocentric Action Recognition with Multimodal Data Convolutional long short-term memory networks for recognizing first per- son interactions

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:52.179259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:46.968419Z digest=sha256:47524f80810e0fe61b26c13eef13ad8256e4b38682ba39c52cac71656b1f6576

Observation 698690a4-ebe7-47b2-80d9-2c727c70ff5e · outbound

This paper cites Attention is All We Need: Nailing Down Object-centric Attention for Egocentric Activity Recognition.

Efficient Egocentric Action Recognition with Multimodal Data Attention is All We Need: Nailing Down Object-centric Attention for Egocentric Activity Recognition

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:39:47.677342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:47.127877Z digest=sha256:a73331c466b6718d4243336527c3732c565942233cd1ac9da6734ac1a1af16dc

Observation df9a95ff-65a0-4fe1-b40c-b0b457a6cc12 · outbound

This paper cites Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world.

Efficient Egocentric Action Recognition with Multimodal Data Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:51.998471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:47.302341Z digest=sha256:5195d4df4e74d2e677b4beee1e16f67f4c0e6fe34ba76d5797855c5c300ad70a

Observation 5d09577e-6fc1-4c7a-8170-40e07c45bd0a · outbound

This paper cites Spatial tempo- ral graph convolutional networks for skeleton-based action recognition.

Efficient Egocentric Action Recognition with Multimodal Data Spatial tempo- ral graph convolutional networks for skeleton-based action recognition

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:39:51.789981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:39:47.435799Z digest=sha256:9dc1ea7e7186d6c5d328fb613973a5513a9ac76cc48199d2d5078cc6fc131a6d

Pith citing papers

No inbound Pith citation observations are available.