Pith. sign in

Paper Citation Record · LEDGER

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model

As of 16 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2501.00785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.00785 v3

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:46:48.174002Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9c102ae2-cb8f-43fc-804b-ba2aa2970ba2 · outbound

This paper cites AIR-Embodied: An Efficient Active 3DGS-based Interaction and Reconstruction Framework with Embodied Large Language Model.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model AIR-Embodied: An Efficient Active 3DGS-based Interaction and Reconstruction Framework with Embodied Large Language Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.016883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.016883Z digest=sha256:3790b12e4270b8a56ea8c10a20c8f1461b96526edf6f82b0e15985dd2575cab0

Observation dd974f31-2ed6-479d-9729-ff2b7ae379b0 · outbound

This paper cites Handle object navi- gation as weighted traveling repairman problem,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Handle object navi- gation as weighted traveling repairman problem,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.022958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.022958Z digest=sha256:4080cb490ad302067c00ab63dce2250e9acf17d66d675459c309040e508fa1c7

Observation c519d170-b4f1-49b0-9190-a4c6a113043d · outbound

This paper cites Medical robots for infectious diseases: Lessons and challenges from the covid-19 pandemic,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Medical robots for infectious diseases: Lessons and challenges from the covid-19 pandemic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.820813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.027125Z digest=sha256:bfb4ac87fa282b4abbea757f50807cf8e3a2230dd2f8cb52ccd49a84b5cffecf

Observation d73ef616-d2ec-4549-9168-43690b31b8bb · outbound

This paper cites Jacquard v2: Refining datasets using the human in the loop data correction method,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Jacquard v2: Refining datasets using the human in the loop data correction method,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.808905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.031267Z digest=sha256:22a89690b61f6aa87b75ecee003fa00e0e03bbdd268fb75a003af2398048e62a

Observation ac94c641-2695-483e-8e0d-4d14e947f2ad · outbound

This paper cites Distilling location proposals of unknown objects through gaze information for human- robot interaction,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Distilling location proposals of unknown objects through gaze information for human- robot interaction,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.796571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.034551Z digest=sha256:15ace37eac4137755c52872efbc692b4c60b8a99458bbf3b3948779e9a1df7db

Observation e7c5a1c3-1664-47e4-be48-d2f9356a1b4e · outbound

This paper cites Communicating human intent to a robotic companion by multi-type gesture sentences,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Communicating human intent to a robotic companion by multi-type gesture sentences,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.784543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.038095Z digest=sha256:ec1a50d9d236c7c489a370905e290163dc01b4cdc99aeb5c35b4942c4ee5bc13

Observation e9e624ed-6994-4805-bf99-645bf470d91f · outbound

This paper cites Analysis of the accuracy and robustness of the leap motion controller,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Analysis of the accuracy and robustness of the leap motion controller,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.773388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.041649Z digest=sha256:16c5a34368c5c2e526dc1776b28990aed43d125384ef2ab013325b14764677f9

Observation ac100f03-6a1d-48a4-907e-991b05cc6964 · outbound

This paper cites Intuitive multi-modal human-robot interaction via posture and voice,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Intuitive multi-modal human-robot interaction via posture and voice,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.762140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.044851Z digest=sha256:0f372a2e61f0b28a067b3f8334fa8411db90aff6107a8f00d7ae2ca0da753d7f

Observation 566c6986-1d54-4722-8777-94c24c9f55fc · outbound

This paper cites Large language models for human–robot interaction: A review,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Large language models for human–robot interaction: A review,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.048283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.048283Z digest=sha256:a2002ae30ac4afe93a601d44a537656f13f2e14f1cc7f07639af18cc16e67053

Observation 60197be4-e353-4d46-b877-6fea7d2438aa · outbound

This paper cites Chat with the environment: Interactive multimodal perception using large language models,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Chat with the environment: Interactive multimodal perception using large language models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.742470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.051507Z digest=sha256:8627529483931fd9b7668a6c91b201cb09b173f7b6d0cebc7c70bf8c07184b03

Observation 40693d29-ab2b-4d89-9fc4-239260494d1f · outbound

This paper cites Generating executable action plans with environmentally-aware language models,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Generating executable action plans with environmentally-aware language models,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.730618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.054861Z digest=sha256:98a00a89bcf4a45aa9ebcdedf72316a666a640932997399f1c91446d5351e84b

Observation f182ba0e-2643-4502-8c01-f65cf561e03e · outbound

This paper cites A fast and light-weight noniterative visual odometry with rgb-d cameras,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model A fast and light-weight noniterative visual odometry with rgb-d cameras,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.717891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.058222Z digest=sha256:568779fd39f35edfabb14754a770eac03333d143c07b9edd079984de436006a6

Observation 7dafc02a-81dd-44bf-b5d0-c74ce750e39f · outbound

This paper cites Sgba: Semantic gaussian mixture model-based lidar bundle adjustment,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Sgba: Semantic gaussian mixture model-based lidar bundle adjustment,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.706622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.061147Z digest=sha256:b71e609d4f63666a89bf5e16613eb938cb37abc28f0f91d73db3d9aa23f6ca0b

Observation 9a3ff5e5-7dec-4726-abce-c25c6f7e6819 · outbound

This paper cites Survey on localization systems and algorithms for unmanned systems,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Survey on localization systems and algorithms for unmanned systems,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.063994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.063994Z digest=sha256:644c1f42c1c20d87c574ac33e7ad6eb5310d05adf2855363b1e4ae6a7a79bdaa

Observation beaf1644-907d-443d-af39-ddaa9b72bc4e · outbound

This paper cites From macro to micro: Autonomous multiscale image fusion for robotic surgery,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model From macro to micro: Autonomous multiscale image fusion for robotic surgery,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.688861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.067463Z digest=sha256:54a5dcd44653bb39c45d675427a8f3a520a4604f0588d653ff2e93c3cc686177

Observation a740985b-9ca9-4d01-ba96-e86057c2f406 · outbound

This paper cites A socially assistive robotic platform for upper-limb rehabilitation: A longitudinal study with pediatric patients,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model A socially assistive robotic platform for upper-limb rehabilitation: A longitudinal study with pediatric patients,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.677123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.071162Z digest=sha256:7b93111bb25dccc1dac25c47a3a73e1bfddbb860d266b845522b78420ef180b2

Observation 1ed8c4bc-2a1a-4322-820e-3b2084da239d · outbound

This paper cites Language-conditioned imitation learning for robot ma- nipulation tasks,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Language-conditioned imitation learning for robot ma- nipulation tasks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.075381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.075381Z digest=sha256:fcdde49e6e7ec97ea8ff060ee26bbf2e6b8c3c1fac773345d46c37ddc6bf9a7f

Observation bab4ed14-f9aa-4577-bc37-0d6fb5af09dc · outbound

This paper cites Towards real-time physical human-robot interaction using skeleton information and hand gestures,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Towards real-time physical human-robot interaction using skeleton information and hand gestures,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.657370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.078959Z digest=sha256:8b55f2df5f6512d52afaf3eb04aeb4e8dc140803de9ddc4e2f52d322b31e2e3c

Observation ebe3fc61-68c9-4ead-b8ad-90530eda6a65 · outbound

This paper cites Control system shell of mobile robot with voice recognition module,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Control system shell of mobile robot with voice recognition module,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.646365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.082519Z digest=sha256:cca217f15d8734743bebfd6815ca154623204d0cf4c9f03d2e20508fe94a4b7f

Observation b8d038e5-ee87-4e6f-b8f4-1261bc9cd8fd · outbound

This paper cites Armar-6: A high- performance humanoid for human-robot collaboration in real-world scenarios,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Armar-6: A high- performance humanoid for human-robot collaboration in real-world scenarios,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.633964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.085860Z digest=sha256:88ad94aa8bab01cbe1f0e4512950d53a886e5974174c4337e2738602a4dfa7ac

Observation 830ec27d-e4ff-4711-8565-8e5f0262f477 · outbound

This paper cites Unsupervised scene categorization, path segmentation and landmark extraction while traveling path,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Unsupervised scene categorization, path segmentation and landmark extraction while traveling path,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.623021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.089379Z digest=sha256:e5baa2a7270b27db8e3c1b3d96c5453cce921513725fc774e0b31be2407e00c3

Observation 7bbf63b1-daa9-42a8-883c-b3998b2198e3 · outbound

This paper cites Working with walt: How a cobot was developed and inserted on an auto assembly line,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Working with walt: How a cobot was developed and inserted on an auto assembly line,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.611520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.093012Z digest=sha256:53394e4489be97a5e5d9d0f834db59577ccb5af47cb2daa46a3b1db91c06b1f3

Observation ef210d54-5d3a-4b27-ae81-fd551cd3d258 · outbound

This paper cites A tool for organizing key characteristics of virtual, augmented, and mixed reality for human–robot interaction systems: Synthesizing vam- hri trends and takeaways,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model A tool for organizing key characteristics of virtual, augmented, and mixed reality for human–robot interaction systems: Synthesizing vam- hri trends and takeaways,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.600536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.096663Z digest=sha256:142840e804c7ee180bd568f5a12d678f2dac7d134d066c05abd72655811cb789

Observation 7fd0aa70-2306-4bf0-99fa-d338b8bfbdf6 · outbound

This paper cites An extensible architecture for robust multimodal human-robot communication,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model An extensible architecture for robust multimodal human-robot communication,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.589005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.100788Z digest=sha256:620472f6053c5cb6722b15dc347a351d695fe71db680752bc344179c113a8579

Observation dc39ffb9-52cc-49e9-b333-9f1fcc6cc54f · outbound

This paper cites Intelligent robotic wheelchair with emg-, gesture-, and voice-based interfaces,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Intelligent robotic wheelchair with emg-, gesture-, and voice-based interfaces,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.578089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.104660Z digest=sha256:effcab89f4ac95eefebad422f1cf483da320ee1b8d93fc0904361322c7ba9f68

Observation 506d7710-ac07-4fed-b45f-c6d290d5878b · outbound

This paper cites Interactive multimodal robot dialog using pointing gesture recogni- tion,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Interactive multimodal robot dialog using pointing gesture recogni- tion,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.566054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.108869Z digest=sha256:ff39d08b60338e04e08a119ad0262d1f70893da0e09a7bcbdb01a669f67d5211

Observation 8e4f1339-1480-4818-b233-d86617383a01 · outbound

This paper cites Adaptive-lio: Enhancing robustness and precision through environmental adaptation in lidar inertial odometry,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Adaptive-lio: Enhancing robustness and precision through environmental adaptation in lidar inertial odometry,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.554230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.112446Z digest=sha256:40ab585c198530efccb05a99dbac04f5a0c3285464f4a34883f54e5b98c9c850

Observation b9519a43-7540-454c-9f52-40a63baba78f · outbound

This paper cites Robust loop closure by textual cues in challenging environments,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Robust loop closure by textual cues in challenging environments,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.542463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.116160Z digest=sha256:f8f67a6e11034b25a1885328f794704fbf006ef14045cc65cab392a3c95975f1

Observation 59309e38-4b27-496e-ab92-fed2f2c928d0 · outbound

This paper cites Chatgpt for robotics: Design principles and model abilities,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Chatgpt for robotics: Design principles and model abilities,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.531417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.120390Z digest=sha256:1983541102f4dc49514d2490f440bf75f4c1fe17064542ae32449f7cd3e95f83

Observation 12d6069f-0c8f-4006-acfc-02326656b876 · outbound

This paper cites Ua-mpc: Uncertainty-aware model predictive control for motorized lidar odom- etry,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Ua-mpc: Uncertainty-aware model predictive control for motorized lidar odom- etry,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.520688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.124279Z digest=sha256:5dfd1f0b94dffaa6d1d1b2b6007111c207a7d578d2e6b1d0c838457d54b9b56f

Observation 999f7006-3bb3-49ae-8cf3-fca1336b6236 · outbound

This paper cites Unsupervised uav 3d trajectories estimation with sparse point clouds,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Unsupervised uav 3d trajectories estimation with sparse point clouds,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.127536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.127536Z digest=sha256:eba1f4a335431199de43906c256aed501023070806931cb71954d1a8d6af2c8e

Observation a5a48dfe-f9fb-4165-8d95-c5f967616be6 · outbound

This paper cites Language mod- els are few-shot learners,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Language mod- els are few-shot learners,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.502949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.130980Z digest=sha256:131429e881bfa027d8051c5fb159743a456ffeeb5d412bbec9597e53d0350716

Observation 27dfbcdd-c85e-432c-8bf1-f87a340a9552 · outbound

This paper cites Evaluation of the efficiency of state-of-the-art speech recognition engines,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Evaluation of the efficiency of state-of-the-art speech recognition engines,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.490155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.133861Z digest=sha256:01777fca2dcfea5444ed6210f06e2137c7dfc7bd9024771036f96b4c041409e7

Observation 62fc647c-412b-433d-ab61-4ba302fde950 · outbound

This paper cites Hand and arm gesture- based human-robot interaction: A review,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Hand and arm gesture- based human-robot interaction: A review,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.478557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.136524Z digest=sha256:a2f2d29b90233d055c4094ce24b3b57204f5f7862f389bc1d0b1678c567b6944

Observation 89c588d9-2eeb-4f4f-9021-56dddfcec6ab · outbound

This paper cites You only look once: Unified, real-time object detection,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model You only look once: Unified, real-time object detection,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.139938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.139938Z digest=sha256:41e184df19f05dccfdd8173792b8b21d6d8845ccc6f87e48b84e2874b6d9cf9f

Observation 9af4ce04-1d1c-4686-94ef-3446478ef5e8 · outbound

This paper cites Yolo-world: Real-time open-vocabulary object detection,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Yolo-world: Real-time open-vocabulary object detection,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.143612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.143612Z digest=sha256:762f668da41e7f0ffaf599e7f1aa51cf94d95f64b920d45879afa70555fa34c4

Observation 0f1214f0-2c30-442f-9ecb-8c06d34f8d10 · outbound

This paper cites Realtime multi-person 2d pose estimation using part affinity fields,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Realtime multi-person 2d pose estimation using part affinity fields,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.147126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.147126Z digest=sha256:2905a198f3b80936e0c4d73f4b914f2cacb08a678b0bed18caf0537d3059697f

Observation 1bc1d1a8-ae67-495f-8e36-60dfda88a907 · outbound

This paper cites Autonomous object level segmentation,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Autonomous object level segmentation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.446082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.150681Z digest=sha256:d40ea62d309ff7a98b726c95258a5ece020be181e71fb3ab9b4f25fe58cd8c37

Observation 7f1ea820-4f0b-4681-a6fd-d64853af6beb · outbound

This paper cites Airslam: An efficient and illumination-robust point-line visual slam system,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Airslam: An efficient and illumination-robust point-line visual slam system,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.435205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.154360Z digest=sha256:236a148f88773fbb565715671bff13d091993ee3ef9b68bd3745d8169e061717

Observation 5843b6df-3874-4cbc-9347-939a582f0609 · outbound

This paper cites Relative localiz- ability and localization for multi-robot systems,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Relative localiz- ability and localization for multi-robot systems,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.424521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.157726Z digest=sha256:b2978d94086c3e34792f28878a6f543ae20dafa93ad5ea384790983a9723efa1

Observation 78aa660b-3d34-4589-ac4c-b457ab40a28a · outbound

This paper cites Large-scale uwb anchor calibration and one- shot localization using gaussian process,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Large-scale uwb anchor calibration and one- shot localization using gaussian process,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.411778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.160837Z digest=sha256:ce2182f49da9c6835c8d416b3a94a563edbafef18d4971cd2db564ec1a3b1073

Observation 751fcb3f-b8d1-4bc2-889a-297889d6f896 · outbound

This paper cites Helmetposer: A helmet-mounted imu dataset for data-driven estimation of human head motion in diverse conditions,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Helmetposer: A helmet-mounted imu dataset for data-driven estimation of human head motion in diverse conditions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.399328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.164396Z digest=sha256:63300172b2b56f4622821eed121d5f7ab9455e50327ecfd67c8f7093baa21df8

Observation 2c7b8bef-9bc3-4607-a15a-2f7d666c6e99 · outbound

This paper cites Heterogeneous stereo: A human vision inspired method for general robotics sensing,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Heterogeneous stereo: A human vision inspired method for general robotics sensing,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.386740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-10T22:46:48.167797Z digest=sha256:1c94c5b0034ffc3f20ea2516e47c9e5adc1a515f7b8a7fc8cc8280ecc9375ad8

Observation 444ebb34-c38f-4a7d-9203-86e37f408d9a · outbound

This paper cites Nvp-hri: Zero shot natural voice and posture-based human–robot interaction via large language model,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Nvp-hri: Zero shot natural voice and posture-based human–robot interaction via large language model,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.170804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.170804Z digest=sha256:8091fbfc206bc922bc7cd4f62df874496ea0411725204a191940af87ddcc8c4d

Observation 829ca9da-7935-4272-bb01-b912871b234b · outbound

This paper cites FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.174002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.174002Z digest=sha256:b3326c8dea9dd07dcf84a6d807bb74bf28688188fb52e59176e5a9fe9450cb5c

Pith citing papers

No inbound Pith citation observations are available.