Pith. sign in

Paper Citation Record · LEDGER

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration

As of 23 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2504.19847.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19847 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:46:49.074505Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 02ea5475-83e0-437c-8ff2-ce56285ddb16 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.290650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T05:46:49.006517Z digest=sha256:93d3ef7b89551e83b5357aa9febdce9f554c2bcb2731626eaaf2be30a2b460ea

Observation 7f851866-9bdf-43f3-a7f7-00b6dc322be3 · outbound

This paper cites LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.011495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.011495Z digest=sha256:4136e19916d6f332658a913dda44e86d43278e0ee56c96bcf95f31cda1e51a32

Observation 8ff6e376-9cc5-4927-acdf-669bf89f9d46 · outbound

This paper cites iCAN: Instance-Centric Attention Network for Human-Object Interaction Detection.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration iCAN: Instance-Centric Attention Network for Human-Object Interaction Detection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.021627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.021627Z digest=sha256:17dcfe25d1ee589fd194389c89529d1ef5d41ae591969a51f0620b9ff830731c

Observation b8f8ea44-11ed-4612-872d-cb1f2a7d6cc1 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.278559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T05:46:49.031370Z digest=sha256:ef889b82306c6a800d5cc208a09c04d554548e4daa2008a7abe12842e260d8f1

Observation 53607f26-2daf-445b-b3fa-b9201fa7bfed · outbound

This paper cites InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.045303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.045303Z digest=sha256:2aecf5f980e59410c36205a3ecba609b0d6eed83af5fb22d21bcd4518dcae841

Observation 5b19f973-773e-4c95-af69-7112db5b5483 · outbound

This paper cites HOI4ABOT: Human-Object Interaction Anticipation for Human Intention Reading Collaborative roBOTs.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration HOI4ABOT: Human-Object Interaction Anticipation for Human Intention Reading Collaborative roBOTs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.055521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.055521Z digest=sha256:d37b2dbab3d6241a1247aa28886774d633e6fc6085564e18dd31e2bd8ea00f45

Observation 99ae6d95-c307-4ba6-bb05-3a60f445332c · outbound

This paper cites DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.069455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.069455Z digest=sha256:8c252a410dffea7785f3fe614ea1c4963773f5ac3b6fb30c8e7fa472e322be97

Observation b97a4b0f-ea87-46a8-81cd-e7c9be54cc2f · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.074505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.074505Z digest=sha256:ae3a940c08fe69d281e67bc870758d6fc2607394d30c2bd6bf2ca0742b3f53ac

Observation 69fd76ab-d1c2-4bb5-b7d0-5fcd3a33e81c · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 2014

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.266577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T05:46:49.035826Z digest=sha256:df13babad8649cd7ee5a030a7473545092b89be80ac7c1ca1a1ee36b5c23c470

Observation 7d4d5ae1-453f-4e44-be16-9dbce1c2165b · outbound

This paper cites Visual Semantic Role Labeling.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Visual Semantic Role Labeling

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.026653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.026653Z digest=sha256:84396480ff39464dd65fd9a0b7b695416f9c21a1536746df64fe97b18bedad8d

Observation 000c0eaf-9edc-41ad-b5fa-5f15b6197962 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 2016

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.239473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T05:46:49.060359Z digest=sha256:fbb575b1428026bc03dc140f33985ce2aec2f030871b8a0f05cdb1aca8b637e5

Observation c0ccbfd0-cdea-4a78-8d0d-0737ae1c1618 · outbound

This paper cites an unresolved cited work.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Unresolved cited work

Reference 2018

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:46:49.302854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T05:46:49.001638Z digest=sha256:2cbd4b0b34a73d455d198ec129838f7948743c3914ff4eb17fe46bf883fbdfdd

Observation e0027bfc-2a46-4ce4-9c31-95518fe02a17 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.016559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.016559Z digest=sha256:1436f5b11e0db5680a82f9f73653ea3ee46779f21182455a9f6d7ffc63405d76

Observation 9d98046a-a8ce-40f1-ba50-bbf7e160daae · outbound

This paper cites Computational Intelligence and Neuroscience 2021, 9922697.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration Computational Intelligence and Neuroscience 2021, 9922697

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:46:49.253759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T05:46:49.050782Z digest=sha256:bb3ea5a7a33aacc8ed899861c3bcf8204dc1cd7e9bb02e7ea62b35e641e35033

Observation 12c53fd1-4979-40be-8f94-cdf54f4238d7 · outbound

This paper cites DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.040494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.040494Z digest=sha256:7bc703dd66910c0617153ae487d5b571dddf55f2893c20fc1603205bad03b65d

Observation 5430cd55-03d5-49a6-9a97-f832eb95c26a · outbound

This paper cites GPT-4 Technical Report.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:48.995926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:48.995926Z digest=sha256:1747b25e8d83b0304632d2ac24cfdf5b31e1d4c7f828638101cdcfa69ffcfa6b

Observation f6671e11-45bd-4de5-842c-84b6e5b99c4c · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration SAM 2: Segment Anything in Images and Videos

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T05:46:49.064632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:46:49.064632Z digest=sha256:bb22828842e89cd241d588a619f2120cf15ff6991de658285864628acfe26d09

Pith citing papers

No inbound Pith citation observations are available.