Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:56:23.789593Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2501.00432.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:56:23.789593Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 87e86ae7-c9a5-44ca-a0f7-d84752eff623 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38df0fac-9e36-4d14-92a5-48fe7147a714 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Gemma: Open Models Based on Gemini Research and Technology
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2a1420b-504d-4ce1-8243-97e034835968 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bb9d0b30-a686-48cc-9731-7bcfe6af0515 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Language-grounded dynamic scene graphs for interactive object search with mobile manipulation,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4c4088ad-c387-4a9e-86b6-dd42d59b8338 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 065d66d5-1452-4bf7-a252-c822fe8f44e6 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Code Llama: Open Foundation Models for Code
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc779ece-c635-45cb-bb9a-a23e874f2ed6 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03ec1088-e661-4fce-8483-36849b6b1f6b · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Text me the data: Generating ground pressure sequence from textual descriptions for har,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d096f72a-b7db-47f1-8beb-f690a8bd9a96 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ec3e0b-5e59-4793-b5ea-8f0473dbf639 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Bliva: A simple multimodal llm for better handling of text-rich visual questions,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6cdbf847-6e40-47f2-885b-ce49ad5fcced · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Infogcn: Representation learning for human skeleton-based action recognition,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e2742733-127e-4afb-b79f-0fe78d0b7222 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models MuJo: Multimodal Joint Feature Space Learning for Human Activity Recognition
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c0d010bc-eef7-4ae9-850f-5435a80c912e · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Decoupled spatial- temporal attention network for skeleton-based action-gesture recogni- tion,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bde806a5-987e-46eb-a94e-645309497de0 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models ALS-HAR: Harnessing Wearable Ambient Light Sensors to Enhance IMU-based Human Activity Recogntion
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4617cad2-2272-4829-9dea-f1deff563e5c · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Human- to-human interaction detection,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 64f0c954-6907-4a33-89d6-f0f86dd997b9 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 395dfb42-dcd6-46c8-b341-a0f9637baa1d · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Hargpt: Are llms zero-shot human activity recognizers?,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 657d9849-0e6d-4ebf-8902-ccfec0004185 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Unsupervised Human Activity Recognition through Two-stage Prompting with ChatGPT
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23e7e97a-c047-4ddd-9efe-647631874ae8 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Segment anything,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 33c1a0c1-e98c-4e43-85b1-977c4e80d32a · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5133d8c2-772f-4b07-92b2-f6779863812e · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd55a662-375e-404f-a130-6b7fa18f9898 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models GPT-4 Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3bafba6-9543-4549-bc43-8a2c5ec89cca · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Track Anything: Segment Anything Meets Videos
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b899e3e7-ec58-4afc-83ad-3c78f8d0bb6b · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Actions in context,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f2ed5518-f78d-4e2f-aa19-02a0c6d631ca · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Sportshhi: A dataset for human-human interaction detection in sports videos,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 81615868-aa76-4666-8700-2b1e3b646c9e · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models High five: Recognising human interactions in tv shows.,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 775c389b-6dff-4080-9af3-82b417bfe29c · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Two-person interaction detection using body- pose features and multiple instance learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation edb78eb0-4d28-49be-9d84-56f615613598 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models First-person activity recognition: What are they doing to me?,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 11898bc6-75b9-4896-8fde-d47ebddf5a3d · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Interaction relational net- work for mutual action recognition,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b0c9739a-7184-421c-ac69-12b35725793c · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models The Kinetics Human Action Video Dataset
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa397625-c615-4051-ad4f-7954a0ce4db9 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Air- act2act: Human–human interaction dataset for teaching non-verbal social behaviors to robots,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 579d3cd5-eaba-4fd1-82c2-27bad80a2ba5 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Human behavior under- standing,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ab50a7af-e2f7-4ff8-bc9e-dc65f055ff5a · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Ntu rgb+ d 120: A large-scale benchmark for 3d human activity understanding,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b619f521-e1ed-4712-b6f6-6b22246a16ba · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Caption Anything: Interactive Image Description with Diverse Multimodal Controls
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12207379-a3be-4412-a41d-7ea6de79d4da · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Vitpose: Sim- ple vision transformer baselines for human pose estimation,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b58fb453-5b87-4516-9d1e-db0f902034fe · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models Tokens- to-token vit: Training vision transformers from scratch on imagenet,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a500a090-c4bf-43d6-acef-9ca4d3e8d58d · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91384ef2-fb44-493f-8492-a01fe3816933 · outbound
OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models LoRA: Low-Rank Adaptation of Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.