Pith. sign in

Paper Citation Record · LEDGER

MotionLLM: Understanding Human Behaviors from Human Motions and Videos

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2405.20340.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20340 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:06:27.907738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:39:29.096987Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8ce634ec-3c2a-48dc-938a-0d2ab47f9bc8 · inbound

Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward cites this paper.

Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:27.907738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:27.907738Z digest=sha256:9a62c2ae872b4efa885eef96fe5a78930907ca2903d5ea04b9fbdb5f4866f547

Observation 54b94232-4936-4c5b-83ae-3360eb5bea48 · inbound

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding cites this paper.

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:48.674994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:48.674994Z digest=sha256:f7f125b5055fe3408143ce03f5cc2a990b05d37c4e14003d4512d1c565e366cd

Observation 41e14b63-2c7c-442a-a5df-bdb37a9e620a · inbound

KptLLM++: Towards Generic Keypoint Comprehension with Large Language Model cites this paper.

KptLLM++: Towards Generic Keypoint Comprehension with Large Language Model MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:22:08.784420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:22:08.784420Z digest=sha256:dd9b46b4b7dde71236a3305b012346ae7dc3a62ed43af94d3833627104a9483d

Observation f7cd73b6-a9aa-41d4-97c9-144cc1f960c5 · inbound

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos cites this paper.

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T15:33:45.839617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:33:45.839617Z digest=sha256:bdeee985d29c4dcc2ccc0fb25ac67c490baff70d4ce684ac29565b0f702143a1

Observation c0d45bf0-6888-4ddc-82c2-125238f37983 · inbound

Being-M0.5: A Real-Time Controllable Vision-Language-Motion Model cites this paper.

Being-M0.5: A Real-Time Controllable Vision-Language-Motion Model MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T21:54:59.349922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:54:59.349922Z digest=sha256:6055d9516d0397a55401b15090f45e64a2dac99a32121377891fa3b12d90a0de

Observation 5a7c04f5-6443-4215-8b94-e45c166ae1ed · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:09.468537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:09.468537Z digest=sha256:a561cd0c19f16b9eecabe2a520f22a2b1cf994f8302bcbbf6b18e999efb668f4

Observation 65a20522-797f-474f-bae5-05e15825bd9c · inbound

Hierarchical Motion Captioning Utilizing External Text Data Source cites this paper.

Hierarchical Motion Captioning Utilizing External Text Data Source MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T12:35:04.892673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:35:04.892673Z digest=sha256:ea51cad5a87289b3b22a2c3dfc0b89c597fbca1c78619da695fa7137ac15ab58

Observation 05ee7cd5-197e-498a-9d01-c2db73514fbb · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:59:08.668703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:c79ff61adf23b27749a1484fd4399469340ed5fab07c1804146d8c15c510d2bd

Observation a3a45041-6704-43b5-98d6-654eeb4c829f · inbound

Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation cites this paper.

Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T05:29:27.699878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:29:27.699878Z digest=sha256:329163062bc0257d8efd0f8e5cf43d1e11f688b607666595b5e9531f0a6ca930

Observation ca3d3740-6eab-443e-9fd9-2dfdf2a49c07 · inbound

UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation cites this paper.

UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T20:17:37.389126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T20:17:37.389126Z digest=sha256:495c8327802df90f9e11f1de015f5a2129ac6b3a705640ef50f82c936b72ee9b

Observation 76407d53-ee9b-4346-81d1-5900af9b42db · inbound

Seeing Without Eyes: 4D Human-Scene Understanding from Wearable IMUs cites this paper.

Seeing Without Eyes: 4D Human-Scene Understanding from Wearable IMUs MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:21:07.101462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T21:59:00.442135Z digest=sha256:75f66af7feb27823ea16a038d0f26a5ce1aafc47724133bd5415cdef99d0055b

Observation 5735b75f-2e32-4b48-8fbb-c2bace7baf84 · inbound

MotionHiFlow: Text-to-motion via hierarchical flow matching cites this paper.

MotionHiFlow: Text-to-motion via hierarchical flow matching MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:10.922429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T08:28:42.524111Z digest=sha256:1484e25f1f012e69681700299d6b4588764dbefbd1c967c9c4849c845561359e

Observation 4d25b8dc-0f5e-4756-b49e-222cac66069f · inbound

WirelessSenseLLM: Zero-Shot Human Activity Understanding by Bridging Wireless Signals and Human Language cites this paper.

WirelessSenseLLM: Zero-Shot Human Activity Understanding by Bridging Wireless Signals and Human Language MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:28:30.749765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:26:59.318141Z digest=sha256:ef416e0843ee60fab4ccd2181e94a161ac405ca186333354acfd7a0f5c260a8f

Observation b8cca54f-72ac-49c7-bfe2-839abf946976 · inbound

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models cites this paper.

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:26:46.201459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T06:55:22.332372Z digest=sha256:b78f27c1629986005659631bf3a298bd4aebee22b8822bd0737f80d31e9aea3b

Observation 871a9618-42eb-4be6-9d14-4a500309b68d · inbound

Fine-grained Human Motion Understanding with Language Models cites this paper.

Fine-grained Human Motion Understanding with Language Models MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:39:29.100697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:55:49.866744Z digest=sha256:6590db597b3c43c60c72e746b5d97a68aafcd68b39d31e8a25a2e5a92f2a8d93