Pith. sign in

Paper Citation Record · LEDGER

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction

As of 22 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2507.15729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15729 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:31:10.316818Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact3
  • verified fuzzy21
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 361260b7-9a8e-4c6d-ab29-6170a9dc4aa0 · outbound

This paper cites Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:10.234237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:10.234237Z digest=sha256:f2a3e7cb5307d5e8461d500be3bc4c7457e778859752326f46f122354b104b55

Observation a03007f2-5213-4ddd-ab4a-2b1f79a0a540 · outbound

This paper cites LaMI: Large language models for multi-modal human-robot interaction,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction LaMI: Large language models for multi-modal human-robot interaction,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.592974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.237721Z digest=sha256:ecaf4e3e7f75a4e4a63fc478707afb4f80027a1b879696c1c062bc392295474d

Observation 6df25037-7d45-4f1f-a502-7012d5366ca0 · outbound

This paper cites Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:10.240655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:10.240655Z digest=sha256:1c24f4d0811813363da1d1265dc3c3388a0ed85035014ef54535a58ca4d7ef86

Observation bfc743cc-6e2d-46b9-b5f3-737a032ea219 · outbound

This paper cites Evaluating Efficiency and Engagement in Scripted and LLM-Enhanced Human-Robot Interactions.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Evaluating Efficiency and Engagement in Scripted and LLM-Enhanced Human-Robot Interactions

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:31:10.398015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.243547Z digest=sha256:f93b4eb984d323af738f63d3ba0b9e803031c3b6bd441177cb97ba463bcf1c40

Observation a4a1ab47-e5f4-4851-9f85-fe646b0b470a · outbound

This paper cites Anticipatory robot control for efficient human-robot collaboration,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Anticipatory robot control for efficient human-robot collaboration,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.585630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.247150Z digest=sha256:a6537fd2d59bf32993c2f827b08d55e80f262f86d2c607e748205a7dc04d1066

Observation 78a99687-856d-46e5-959a-0743fe938871 · outbound

This paper cites The Effect of Anthropomorphism on Trust in an Industrial Human-Robot Interaction,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction The Effect of Anthropomorphism on Trust in an Industrial Human-Robot Interaction,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.578746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.250318Z digest=sha256:725f6fd5c7b03589e14269ff40e968ea7579960865d3eb90d4f57891a74faefd

Observation a8f9051b-ef02-4de0-aaee-7aea8ea8f73a · outbound

This paper cites Advan- tages of Multimodal versus Verbal-Only Robot-to-Human Communi- cation with an Anthropomorphic Robotic Mock Driver,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Advan- tages of Multimodal versus Verbal-Only Robot-to-Human Communi- cation with an Anthropomorphic Robotic Mock Driver,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.571024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.253487Z digest=sha256:dc6632c78efd4320af3cc703d70821ffb6ca38fe3f323b67449c8ba5d3accb9e

Observation d7b17e0b-7347-4f13-a318-84934e76039d · outbound

This paper cites TH ¨OR-MAGNI: A large-scale indoor motion capture recording of human movement and robot interaction,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction TH ¨OR-MAGNI: A large-scale indoor motion capture recording of human movement and robot interaction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.563389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.256334Z digest=sha256:26ecb76fa46ad4e82202a8084506abe361415ae489ca7434bb7e028f239a27be

Observation 757f796b-1597-41d6-99ab-3b88aeb50531 · outbound

This paper cites Leveraging Large Language Models in Human-Robot Interaction: A Critical Analysis of Potential and Pitfalls.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Leveraging Large Language Models in Human-Robot Interaction: A Critical Analysis of Potential and Pitfalls

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:10.259203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:10.259203Z digest=sha256:4befaf827b4b77db39a7324e02937e96b73d7e65be90331f4d06de09f17c367b

Observation 88156b7f-2e2d-4bf3-af1c-3101df190f09 · outbound

This paper cites Chatgpt for robotics: Design principles and model abilities,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Chatgpt for robotics: Design principles and model abilities,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.556092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.262358Z digest=sha256:cb614dc26ea2430cf85c5a0456ffbc2c5205594d06fbbfd9636568ba51e3fa6a

Observation 0c0d60f9-b9bd-4c1a-88ce-ce2ed8be9e14 · outbound

This paper cites To help or not to help: LLM-based attentive support for human-robot group interactions,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction To help or not to help: LLM-based attentive support for human-robot group interactions,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.548854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.265102Z digest=sha256:e0f5fe66703d5332dfbc9cd85dbbc6d07f7c6f7043787c5bc62627a9a5625efa

Observation 3ecca934-233d-4ada-b0a4-76fbd50b6549 · outbound

This paper cites How to Communicate Robot Motion Intent: A Scoping Review,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction How to Communicate Robot Motion Intent: A Scoping Review,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.541103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.267892Z digest=sha256:74ac46ceda4b2c8f4472ccf913e860e88c7fef77cfd6bb2d50932218e117fe33

Observation 62e00f09-d0a2-45c7-8983-7a70f337b9b2 · outbound

This paper cites Bi-directional navigation intent communication us- ing spatial augmented reality and eye-tracking glasses for improved safety in human–robot interaction,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Bi-directional navigation intent communication us- ing spatial augmented reality and eye-tracking glasses for improved safety in human–robot interaction,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.533575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.270735Z digest=sha256:9950088f39182efdcb6eb7c4ee7073ee6f51de6103b540df3be7e6c6edf0d7cb

Observation 877b384e-fe7b-43e3-93e6-89e15104a115 · outbound

This paper cites When robots get chatty: Grounding multimodal human-robot conversation and collaboration,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction When robots get chatty: Grounding multimodal human-robot conversation and collaboration,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.525722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.274174Z digest=sha256:2835ad8d4a8cf566c002deb64eba1f3ecd3d5a97985e3f8c457abe47e2a2140c

Observation 0669bc59-4725-4cd5-8325-eaac2f200156 · outbound

This paper cites The need for verbal robot explanations and how people would like a robot to explain itself,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction The need for verbal robot explanations and how people would like a robot to explain itself,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.518193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.276891Z digest=sha256:b41d090a7c0cb16f534f2ba4068211575c0586e7c2018208a8cb4e28721e9486

Observation e3adbb96-1c0d-45d1-8a77-cf85f9b0a131 · outbound

This paper cites Spoken language interaction with robots: Recommendations for fu- ture research,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Spoken language interaction with robots: Recommendations for fu- ture research,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.509851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.279659Z digest=sha256:b621e852783a26388cf449280adbe3ab94d707c45d026867ebe66b219d72d277

Observation e9349e9b-bbec-4b96-906f-3d892b2a8c8c · outbound

This paper cites Human Gaze and Head Rotation during Navigation, Exploration and Object Manipulation in Shared Environments with Robots.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Human Gaze and Head Rotation during Navigation, Exploration and Object Manipulation in Shared Environments with Robots

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:31:10.378813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.282538Z digest=sha256:b760f07f565e0de68041d41f52b324fa0a0e3209e1aca0a5e172bb899cd97de6

Observation 9ff00dad-b06c-4320-9e7b-bc2863cef1f8 · outbound

This paper cites Meet me where i’m gazing: how shared attention gaze affects human-robot handover timing,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Meet me where i’m gazing: how shared attention gaze affects human-robot handover timing,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.501567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.285887Z digest=sha256:55b38e52905fb13cc43782bfc81db4233fb02ea7ad85e72287756453edd6ed3e

Observation 4c42778b-51c8-48d0-bc9a-09283cf9ebcb · outbound

This paper cites Perceptive Recommendation Robot: Enhancing Receptivity of Product Suggestions Based on Customers’ Nonverbal Cues,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Perceptive Recommendation Robot: Enhancing Receptivity of Product Suggestions Based on Customers’ Nonverbal Cues,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.492551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.288159Z digest=sha256:f5970709da270cd8f8057eb282a4507e14de728f72abf2803f1fad4898173ac4

Observation e7373919-8a96-4fe6-860f-2dcd9dbcad68 · outbound

This paper cites Using human eye gaze patterns as indicators of need for assistance from a socially assistive robot,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Using human eye gaze patterns as indicators of need for assistance from a socially assistive robot,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.482789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.291011Z digest=sha256:458c879f35535258e957fe8db39861414070542ee22b3bf30ae84e08e0869cb8

Observation 222e4af9-03bd-4c83-9684-0d5539093b55 · outbound

This paper cites Enhancing ai interpretation and decision-making: Integrating cognitive computational models with deep learning for advanced uncertain reasoning systems,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Enhancing ai interpretation and decision-making: Integrating cognitive computational models with deep learning for advanced uncertain reasoning systems,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.473744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.293854Z digest=sha256:fa1334fba5f0279f27b0e9fdf9cf243022678c97491c5f95e2254a2448c1af2b

Observation c5b4b333-b90b-4b2d-9a73-00865fa3cb0e · outbound

This paper cites To- ward human-aware robot task planning.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction To- ward human-aware robot task planning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.465340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.296654Z digest=sha256:5995506797c47a11f5f0072e5cb733c0a279eeb6472de7029376acfc200f1503

Observation fb998fb0-166b-48cd-a1b3-444875bef31d · outbound

This paper cites FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:10.299425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:10.299425Z digest=sha256:024ce08781746f5898d94d06cb89db3d9f8bb856bfa17a5c7542021e3e94fa05

Observation 4ec51b74-d2d7-4d00-81fc-3ca32d69d19b · outbound

This paper cites SemanticScanpath: Combining Gaze and Speech for Situated Human-Robot Interaction Using LLMs.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction SemanticScanpath: Combining Gaze and Speech for Situated Human-Robot Interaction Using LLMs

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:31:10.357036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.302195Z digest=sha256:fe37a80d42e831426401a20d37190092884a78786990aeeebde6e176499c524e

Observation 02e13c6d-e819-4016-abff-e5bfa754156b · outbound

This paper cites Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.456739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.304604Z digest=sha256:9b7911edb06d77e5ccdc4d946765017e3fd02b11a38a531693f202f5e42a78fc

Observation cbf3a175-6364-4668-af10-cde8445a43fe · outbound

This paper cites Yolo- world: Real-time open-vocabulary object detection,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Yolo- world: Real-time open-vocabulary object detection,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.448560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.306961Z digest=sha256:63d668c4bf37f0e7b7933675669e71327129a97afa493ab70156e1c6c94091bc

Observation 79f04e8b-1a5f-4d04-996e-5596372afbfd · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction SAM 2: Segment Anything in Images and Videos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:10.309278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:10.309278Z digest=sha256:612c8a3fd396a6a62c63a283003a8631e8843b1f27b72a8388cf0500a324d5c5

Observation ae5186be-1395-4d53-9b61-7cace7a3f7a8 · outbound

This paper cites Robust speech recognition via large-scale weak super- vision,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Robust speech recognition via large-scale weak super- vision,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:10.311723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:10.311723Z digest=sha256:99ecfbb804eb72769466c647bde1e55ca88c41eecfd2c100a9bb817d5763c569

Observation 16a6540e-7958-4287-b001-7aaa4a80ce5c · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Chain-of-thought prompting elicits reasoning in large language models,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.435377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.314374Z digest=sha256:43f57626df28a4a89ddc8e45f52021a300672acf2bb04296d39ab029be1a1fcd

Observation acc5d091-75a9-4c44-913c-479d2d592d21 · outbound

This paper cites Perceived usability evaluation of edu- cational technology using the post-study system usability questionnaire (pssuq): a systematic review,.

Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction Perceived usability evaluation of edu- cational technology using the post-study system usability questionnaire (pssuq): a systematic review,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:31:10.426032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T15:31:10.316818Z digest=sha256:85b6b433cbb046823f278728c2f36eabf7a6b7509edf5e9a7bab41177804f19c

Pith citing papers

No inbound Pith citation observations are available.