Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T01:10:55.850274Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 4 inbound Pith citation observations for arXiv:2503.16492.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T01:10:55.850274Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T19:39:07.075356Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T04:57:38.604855Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a90f630d-ffd0-4942-a234-0f17297e9db1 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Social robots in therapy and care
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c5c96ee-84a7-45b9-b6a9-82fa5872d50b · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Jubileo: An open- source robot and framework for research in human-robot social interac- tion
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f34bb0eb-1fa4-4202-b70c-34b47f84741b · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Human-aware physical human–robot collaborative transportation and manipulation with multiple aerial robots
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d98a3c25-71e5-4f90-b9c5-728f256fc094 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Communicating human intent to a robotic companion by multi-type gesture sentences
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5b6af122-e4a8-40b3-a9be-a587238033ad · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Nvp-hri: Zero shot natural voice and posture-based human–robot interaction via large language model
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e82bc16b-485c-4eec-99fc-960afc86de6d · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Robot reading human gaze: Why eye tracking is better than head tracking for human-robot col- laboration
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8b016c46-0c14-406b-9e6a-28c479563a7d · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech A gaze-speech system in mixed reality for human-robot interaction
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation de8dca7e-1ef4-48da-8c8b-9c8aebb24d00 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Human–robot interaction through eye tracking for artistic drawing
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5c486e0b-366d-4f44-8c52-95c382ebf73b · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Is it possible to recognize a speaker without listening? unraveling conversation dynamics in multi-party interactions using continuous eye gaze
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8466c8fb-5ca1-4abc-a755-aedd77d4e226 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Eye-gaze control of a wheelchair mounted 6dof assistive robot for activities of daily living
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 935e96a0-07e1-43b4-abba-59deb811e27b · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Free-view, 3d gaze-guided, assistive robotic system for activities of daily living
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2ec08df2-47da-4ebf-b7ba-6311edd74354 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Microsaccade-inspired event camera for robotics
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 84a328ec-92e3-407c-bad0-6116ac31cee4 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Project Aria: A New Tool for Egocentric Multi-Modal AI Research
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9599ee3f-bdf7-433d-b998-3f8e424e18fd · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Robust gaze- based intention prediction for real-world scenarios
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ea247167-bc67-49e9-afed-9df33b676395 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Getting to know your robot customers: Automated analysis of user identity and demographics for robots in the wild
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 910f1d15-3185-4d1a-b07e-28c592d36a60 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech A personalized comfort space with variable shape based on environmental information for robot navigation in homes
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation afe87ad1-6531-44e1-a26d-41153c03a66c · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Improving the collision tolerance of high-speed industrial robots via impact-aware path planning and series clutched actuation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a57f45cf-4fa9-494c-8990-9392a22f9248 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech In situ calibration of six- axis force–torque sensors for industrial robots with tilting base
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e243694a-600a-4025-af69-5088701a14c1 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Gesture-informed robot assistance via foundation models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 940510b7-d6c0-4d94-92e9-14a5af3548a0 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Interactive multimodal robot dialog using pointing gesture recognition
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 496eb716-9ef3-4144-b255-ac71be72844e · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Code as policies: Language model programs for embodied control
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b0e1534-af26-4086-bd52-341690d9cfd4 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Progprompt: Generating situated robot task plans using large language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation de495b96-5a16-4344-bf7c-f97d58b016be · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Semi-autonomous robotic arm reaching with hybrid gaze–brain machine interface
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0c579db8-c1f5-44ad-abe7-36cdd32dcd11 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Investigating the usability of collabo- rative robot control through hands-free operation using eye gaze and augmented reality
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c424de04-57ed-4f11-bde6-5bdfb2588bf8 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Human gaze following for human-robot interaction
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f52a08cc-f2d2-4f1a-9642-e0fb93061f79 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Gaze-based attention recognition for human-robot collaboration
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5751a08f-385d-46d0-b29c-f6e50498ead3 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech A novel human-in-the-loop multimodal intention fusion method for human-robot interaction
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8578005a-437a-419a-96ab-d558537fcaee · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Alchemist: Llm-aided end-user development of robot applications
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ecbd186f-2683-4c15-8ba5-4f0afafe9617 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Lami: Large language models for multi-modal human-robot interaction
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bb177b39-f85f-4513-92e2-f0658bc631a8 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Robust speech recognition via large-scale weak supervi- sion
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a72528e2-30a4-4431-aa57-a0b3192fc074 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Grounding DINO: marrying DINO with grounded pre-training for open-set object detection
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ce1f1fe4-5005-4905-ad46-a85942036ea9 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Sam 2: Segment anything in images and videos
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9fda874c-469f-4bb3-a6f7-507ac3e52dcc · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech ORB-SLAM3: An accurate open-source library for visual, visual-inertial and multi-map SLAM
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2c5ad54f-e294-4151-b0f0-2c4d1dcff7a8 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Super- glue: Learning feature matching with graph neural networks
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1f99c207-0fb2-4a07-bf49-3715be11343f · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Design and implementation of a haptic measurement glove to create realistic human-telerobot interactions
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7b540175-6910-4a4e-a4f2-d200020897b2 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Don’t yell at your robot: Physical correction as the collaborative interface for language model powered robots
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 178de768-70af-43cc-90b0-4aaf91676fdc · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Lora: Low-rank adaptation of large language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation acb66482-e5a4-4b3c-9f17-d229533c0146 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Model compression
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 814818ae-8916-4fba-aca1-5f51523fa296 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Retrieval- augmented generation for knowledge-intensive nlp tasks
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b1c3cf0a-1399-45fd-85ed-783303144749 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Human gaze improves vision transformers by token masking
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c56d6b22-0910-4d2f-940e-506fbbffe23a · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Egolife: Towards egocentric life assistant
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a4db9305-d286-4410-9fa0-ff7150e5c81d · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Interactive multimodal robot dialog using pointing gesture recognition
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 539233a1-3390-4a59-a9d1-24b70220ad98 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech The speech recognition error in our system primarily due to misinterpretation of similar-sounding words
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1bf8564e-aeb3-4089-85b9-e3865399e044 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eff27f64-7fbf-4972-a976-62c64a22edbc · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Fea- ture matching using superglue becomes unreliable in the presence of weak object textures, repetitive patterns, or partial occlusions, leading to incorrect correspondences
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation afcba079-64b8-4e3e-ae7d-5b6cbcf93483 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Since FAM-HRI requires the LLM’s response to strictly follow a prede- fined prompt format, any deviation renders the output unusable by subsequent system modules
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ee2e4734-419e-463b-9803-3a8b807a24d2 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cc3d4649-3672-4fd2-80a2-53745af9c675 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 11f0fb02-9e1d-4821-9e3e-14599f2dd333 · outbound
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f9ad8420-686a-4da6-9528-a95aa196d81e · inbound
Mind Meets Space: Rethinking Agentic Spatial Intelligence from a Neuroscience-inspired Perspective FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9f2c4a1-f946-4ca1-afb4-b51790544ce6 · inbound
SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0601019d-8154-4643-a96d-f50bc647a580 · inbound
SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c624a78-df58-47cc-82e1-a232ae2d3f6a · inbound
Hierarchical Policies from Verbal and Egocentric Human Signals for Natural Human-Robot Interaction FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.