Pith. sign in

Paper Citation Record · LEDGER

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration

As of 20 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2504.19847.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19847 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:46:49.074505Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 02ea5475-83e0-437c-8ff2-ce56285ddb16 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.290650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T05:46:49.006517Z digest=sha256:0e93cd7b378857ecc8822058ba45505c190cafdf0b8d26f373949bd6504930d3

Observation 7f851866-9bdf-43f3-a7f7-00b6dc322be3 · outbound

This paper cites LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.011495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.011495Z digest=sha256:4136e19916d6f332658a913dda44e86d43278e0ee56c96bcf95f31cda1e51a32

Observation 8ff6e376-9cc5-4927-acdf-669bf89f9d46 · outbound

This paper cites iCAN: Instance-Centric Attention Network for Human-Object Interaction Detection.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration iCAN: Instance-Centric Attention Network for Human-Object Interaction Detection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.021627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.021627Z digest=sha256:d441ff525db76b81c8dba5089d948bfdf55e809bafa7075dde77f1fa88ca4720

Observation b8f8ea44-11ed-4612-872d-cb1f2a7d6cc1 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.278559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T05:46:49.031370Z digest=sha256:a82e68bfcc6a77beae380c472cd0ca6c658990b38b195d29025353d63b948a9f

Observation 53607f26-2daf-445b-b3fa-b9201fa7bfed · outbound

This paper cites InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.045303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.045303Z digest=sha256:2aecf5f980e59410c36205a3ecba609b0d6eed83af5fb22d21bcd4518dcae841

Observation 5b19f973-773e-4c95-af69-7112db5b5483 · outbound

This paper cites HOI4ABOT: Human-Object Interaction Anticipation for Human Intention Reading Collaborative roBOTs.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration HOI4ABOT: Human-Object Interaction Anticipation for Human Intention Reading Collaborative roBOTs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.055521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.055521Z digest=sha256:19e8bb913afb7e187691ec793ec7cecf895633ae9e685dd59c7a205a07613170

Observation 99ae6d95-c307-4ba6-bb05-3a60f445332c · outbound

This paper cites DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.069455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.069455Z digest=sha256:8c252a410dffea7785f3fe614ea1c4963773f5ac3b6fb30c8e7fa472e322be97

Observation b97a4b0f-ea87-46a8-81cd-e7c9be54cc2f · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.074505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.074505Z digest=sha256:ae3a940c08fe69d281e67bc870758d6fc2607394d30c2bd6bf2ca0742b3f53ac

Observation 69fd76ab-d1c2-4bb5-b7d0-5fcd3a33e81c · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 2014

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.266577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T05:46:49.035826Z digest=sha256:f358a2fb6d6ee98b7d2a7f2b233185eaca92905af7d37adcf639bba6b3139a23

Observation 7d4d5ae1-453f-4e44-be16-9dbce1c2165b · outbound

This paper cites Visual Semantic Role Labeling.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Visual Semantic Role Labeling

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.026653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.026653Z digest=sha256:84396480ff39464dd65fd9a0b7b695416f9c21a1536746df64fe97b18bedad8d

Observation 000c0eaf-9edc-41ad-b5fa-5f15b6197962 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 2016

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.239473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T05:46:49.060359Z digest=sha256:2dcbc95fdfccd6d0d46f77142dcf9cdb48038920b8a99e69376ce4025805fd65

Observation c0ccbfd0-cdea-4a78-8d0d-0737ae1c1618 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 2018

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.302854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T05:46:49.001638Z digest=sha256:90c4b32b492824ee68d7caa8934d1ea5e23e897ace592b821cd92625b66595b0

Observation e0027bfc-2a46-4ce4-9c31-95518fe02a17 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.016559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.016559Z digest=sha256:1436f5b11e0db5680a82f9f73653ea3ee46779f21182455a9f6d7ffc63405d76

Observation 9d98046a-a8ce-40f1-ba50-bbf7e160daae · outbound

This paper cites Computational Intelligence and Neuroscience 2021, 9922697.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Computational Intelligence and Neuroscience 2021, 9922697

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:46:49.253759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T05:46:49.050782Z digest=sha256:71757ad50a88d8826e7a6855edeab51428305452d95050cec69c43497ef72b1f

Observation 12c53fd1-4979-40be-8f94-cdf54f4238d7 · outbound

This paper cites DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.040494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.040494Z digest=sha256:7bc703dd66910c0617153ae487d5b571dddf55f2893c20fc1603205bad03b65d

Observation 5430cd55-03d5-49a6-9a97-f832eb95c26a · outbound

This paper cites GPT-4 Technical Report.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:48.995926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:48.995926Z digest=sha256:1747b25e8d83b0304632d2ac24cfdf5b31e1d4c7f828638101cdcfa69ffcfa6b

Observation f6671e11-45bd-4de5-842c-84b6e5b99c4c · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration SAM 2: Segment Anything in Images and Videos

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.064632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.064632Z digest=sha256:bb22828842e89cd241d588a619f2120cf15ff6991de658285864628acfe26d09

Pith citing papers

No inbound Pith citation observations are available.