Pith. sign in

Paper Citation Record · LEDGER

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model

As of 16 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2501.00785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.00785 v3

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:46:48.174002Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9c102ae2-cb8f-43fc-804b-ba2aa2970ba2 · outbound

This paper cites AIR-Embodied: An Efficient Active 3DGS-based Interaction and Reconstruction Framework with Embodied Large Language Model.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model AIR-Embodied: An Efficient Active 3DGS-based Interaction and Reconstruction Framework with Embodied Large Language Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.016883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.016883Z digest=sha256:4e2f5cbdba3a4b4a9679865422968d15f484c0d655dfd0a72e0a5177d4439cf2

Observation dd974f31-2ed6-479d-9729-ff2b7ae379b0 · outbound

This paper cites Handle object navi- gation as weighted traveling repairman problem,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Handle object navi- gation as weighted traveling repairman problem,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.022958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.022958Z digest=sha256:4080cb490ad302067c00ab63dce2250e9acf17d66d675459c309040e508fa1c7

Observation c519d170-b4f1-49b0-9190-a4c6a113043d · outbound

This paper cites Medical robots for infectious diseases: Lessons and challenges from the covid-19 pandemic,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Medical robots for infectious diseases: Lessons and challenges from the covid-19 pandemic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.820813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.027125Z digest=sha256:5b0a7d6051d1bbf281353379c75a97629f1773fcc7a140a598f4f3c5a85c7e82

Observation d73ef616-d2ec-4549-9168-43690b31b8bb · outbound

This paper cites Jacquard v2: Refining datasets using the human in the loop data correction method,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Jacquard v2: Refining datasets using the human in the loop data correction method,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.808905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.031267Z digest=sha256:8cee90b5a5687f5cc8d3af25c3915f485d94c2470133a7b42970fdd3788b0576

Observation ac94c641-2695-483e-8e0d-4d14e947f2ad · outbound

This paper cites Distilling location proposals of unknown objects through gaze information for human- robot interaction,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Distilling location proposals of unknown objects through gaze information for human- robot interaction,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.796571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.034551Z digest=sha256:6af6b51a93eb0723f47aa131598799faa04f09703a7aece29da7b46a2b3a2636

Observation e7c5a1c3-1664-47e4-be48-d2f9356a1b4e · outbound

This paper cites Communicating human intent to a robotic companion by multi-type gesture sentences,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Communicating human intent to a robotic companion by multi-type gesture sentences,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.784543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.038095Z digest=sha256:f1530ce72a2ec472d97ef3cd65f7c8984509ff97fd3902456034f708cc586126

Observation e9e624ed-6994-4805-bf99-645bf470d91f · outbound

This paper cites Analysis of the accuracy and robustness of the leap motion controller,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Analysis of the accuracy and robustness of the leap motion controller,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.773388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.041649Z digest=sha256:5b561ec21c38645671d2acecba05e5ef879ff7bfd567f52a1912cf7d143f3c0c

Observation ac100f03-6a1d-48a4-907e-991b05cc6964 · outbound

This paper cites Intuitive multi-modal human-robot interaction via posture and voice,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Intuitive multi-modal human-robot interaction via posture and voice,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.762140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.044851Z digest=sha256:ca75a6fa72ec020dfce5cda0c159d13a3fc34000d9a00558a7a39a77d98e519d

Observation 566c6986-1d54-4722-8777-94c24c9f55fc · outbound

This paper cites Large language models for human–robot interaction: A review,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Large language models for human–robot interaction: A review,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.048283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.048283Z digest=sha256:a2002ae30ac4afe93a601d44a537656f13f2e14f1cc7f07639af18cc16e67053

Observation 60197be4-e353-4d46-b877-6fea7d2438aa · outbound

This paper cites Chat with the environment: Interactive multimodal perception using large language models,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Chat with the environment: Interactive multimodal perception using large language models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.742470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.051507Z digest=sha256:28ff7fba234169cf479627dd2d7c65fe1431add55e290439b1f8ad0a97fe3f3f

Observation 40693d29-ab2b-4d89-9fc4-239260494d1f · outbound

This paper cites Generating executable action plans with environmentally-aware language models,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Generating executable action plans with environmentally-aware language models,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.730618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.054861Z digest=sha256:22e4846b09714c011ce2b64364f792d6c2a6a37bd2b48917c724517fa48d7ab1

Observation f182ba0e-2643-4502-8c01-f65cf561e03e · outbound

This paper cites A fast and light-weight noniterative visual odometry with rgb-d cameras,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model A fast and light-weight noniterative visual odometry with rgb-d cameras,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.717891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.058222Z digest=sha256:fcf1fa9cd1bfcb76f9d71933eda25e23c63f97dd42d9837d89f4964bd1068aaf

Observation 7dafc02a-81dd-44bf-b5d0-c74ce750e39f · outbound

This paper cites Sgba: Semantic gaussian mixture model-based lidar bundle adjustment,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Sgba: Semantic gaussian mixture model-based lidar bundle adjustment,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.706622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.061147Z digest=sha256:75f87eedf1eed48a3c6902ac1dfbee092e983adca8e7620ee76d3b9b51a5a7a7

Observation 9a3ff5e5-7dec-4726-abce-c25c6f7e6819 · outbound

This paper cites Survey on localization systems and algorithms for unmanned systems,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Survey on localization systems and algorithms for unmanned systems,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.063994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.063994Z digest=sha256:644c1f42c1c20d87c574ac33e7ad6eb5310d05adf2855363b1e4ae6a7a79bdaa

Observation beaf1644-907d-443d-af39-ddaa9b72bc4e · outbound

This paper cites From macro to micro: Autonomous multiscale image fusion for robotic surgery,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model From macro to micro: Autonomous multiscale image fusion for robotic surgery,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.688861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.067463Z digest=sha256:f8e0e6d4b90f73ae87c55eb51f0199f17b37e83c235e908ba756af0dfa3adf48

Observation a740985b-9ca9-4d01-ba96-e86057c2f406 · outbound

This paper cites A socially assistive robotic platform for upper-limb rehabilitation: A longitudinal study with pediatric patients,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model A socially assistive robotic platform for upper-limb rehabilitation: A longitudinal study with pediatric patients,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.677123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.071162Z digest=sha256:0b1c6c2f7c6c6d2c79471279ff445fc2863fac488a397b6aaef30dce59c97bd7

Observation 1ed8c4bc-2a1a-4322-820e-3b2084da239d · outbound

This paper cites Language-conditioned imitation learning for robot ma- nipulation tasks,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Language-conditioned imitation learning for robot ma- nipulation tasks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.075381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.075381Z digest=sha256:fcdde49e6e7ec97ea8ff060ee26bbf2e6b8c3c1fac773345d46c37ddc6bf9a7f

Observation bab4ed14-f9aa-4577-bc37-0d6fb5af09dc · outbound

This paper cites Towards real-time physical human-robot interaction using skeleton information and hand gestures,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Towards real-time physical human-robot interaction using skeleton information and hand gestures,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.657370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.078959Z digest=sha256:b27884640976e858c9f6c19adba1a9356eab76a8b5315245eaaff43c1fafab46

Observation ebe3fc61-68c9-4ead-b8ad-90530eda6a65 · outbound

This paper cites Control system shell of mobile robot with voice recognition module,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Control system shell of mobile robot with voice recognition module,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.646365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.082519Z digest=sha256:6d7cbb08c6eaebea97d829832989a67d3c9d0d607f22ac5af2bea42ad695d7ae

Observation b8d038e5-ee87-4e6f-b8f4-1261bc9cd8fd · outbound

This paper cites Armar-6: A high- performance humanoid for human-robot collaboration in real-world scenarios,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Armar-6: A high- performance humanoid for human-robot collaboration in real-world scenarios,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.633964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.085860Z digest=sha256:b1f5c2a7b23cac23140b985fe2f9dee7c70017217446a26c3d36b4b2b189d06b

Observation 830ec27d-e4ff-4711-8565-8e5f0262f477 · outbound

This paper cites Unsupervised scene categorization, path segmentation and landmark extraction while traveling path,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Unsupervised scene categorization, path segmentation and landmark extraction while traveling path,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.623021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.089379Z digest=sha256:655477dbaa111f5929ecd3c6229b3d66731a22c32ee50866bcc1f55f3f4cce8c

Observation 7bbf63b1-daa9-42a8-883c-b3998b2198e3 · outbound

This paper cites Working with walt: How a cobot was developed and inserted on an auto assembly line,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Working with walt: How a cobot was developed and inserted on an auto assembly line,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.611520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.093012Z digest=sha256:69966cc2044880d3e7490349814e7e39a5afe9faf76d47528402d525ae73c732

Observation ef210d54-5d3a-4b27-ae81-fd551cd3d258 · outbound

This paper cites A tool for organizing key characteristics of virtual, augmented, and mixed reality for human–robot interaction systems: Synthesizing vam- hri trends and takeaways,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model A tool for organizing key characteristics of virtual, augmented, and mixed reality for human–robot interaction systems: Synthesizing vam- hri trends and takeaways,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.600536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.096663Z digest=sha256:97c14b8275b698e9ce5de1723d1e0dc02f056dd5e0264707a4283961717bd942

Observation 7fd0aa70-2306-4bf0-99fa-d338b8bfbdf6 · outbound

This paper cites An extensible architecture for robust multimodal human-robot communication,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model An extensible architecture for robust multimodal human-robot communication,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.589005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.100788Z digest=sha256:93b016f6a93d5a16b0ec253f14397999c27efa2d7e09fec642b70d02f02e18af

Observation dc39ffb9-52cc-49e9-b333-9f1fcc6cc54f · outbound

This paper cites Intelligent robotic wheelchair with emg-, gesture-, and voice-based interfaces,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Intelligent robotic wheelchair with emg-, gesture-, and voice-based interfaces,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.578089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.104660Z digest=sha256:93c40ff4f483145624c26c3765848a5617c2ee5087d52b80472c56c463d4ce86

Observation 506d7710-ac07-4fed-b45f-c6d290d5878b · outbound

This paper cites Interactive multimodal robot dialog using pointing gesture recogni- tion,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Interactive multimodal robot dialog using pointing gesture recogni- tion,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.566054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.108869Z digest=sha256:df0c9b961fa1799f34fd2556b0752ebd0585474e409c18877323dd4509da862d

Observation 8e4f1339-1480-4818-b233-d86617383a01 · outbound

This paper cites Adaptive-lio: Enhancing robustness and precision through environmental adaptation in lidar inertial odometry,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Adaptive-lio: Enhancing robustness and precision through environmental adaptation in lidar inertial odometry,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.554230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.112446Z digest=sha256:dd2f14cd8797aa12ba04bcf28e25cb97c88e1afd4429600ab6c4ac6b3e47c95a

Observation b9519a43-7540-454c-9f52-40a63baba78f · outbound

This paper cites Robust loop closure by textual cues in challenging environments,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Robust loop closure by textual cues in challenging environments,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.542463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.116160Z digest=sha256:0dd205ee4f401bae3718ddfbdaa1428c981bd04a23489a3f16bb93aaeb248dcb

Observation 59309e38-4b27-496e-ab92-fed2f2c928d0 · outbound

This paper cites Chatgpt for robotics: Design principles and model abilities,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Chatgpt for robotics: Design principles and model abilities,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.531417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.120390Z digest=sha256:07dbd9428ae11bda4e7be6445d5b9e78a7b1c1441681ee93b6078ab773686874

Observation 12d6069f-0c8f-4006-acfc-02326656b876 · outbound

This paper cites Ua-mpc: Uncertainty-aware model predictive control for motorized lidar odom- etry,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Ua-mpc: Uncertainty-aware model predictive control for motorized lidar odom- etry,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.520688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.124279Z digest=sha256:95b3b2134addc14e87debf49bcddadbb6247a07403c15d63f72537dff0c52c66

Observation 999f7006-3bb3-49ae-8cf3-fca1336b6236 · outbound

This paper cites Unsupervised uav 3d trajectories estimation with sparse point clouds,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Unsupervised uav 3d trajectories estimation with sparse point clouds,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.127536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.127536Z digest=sha256:eba1f4a335431199de43906c256aed501023070806931cb71954d1a8d6af2c8e

Observation a5a48dfe-f9fb-4165-8d95-c5f967616be6 · outbound

This paper cites Language mod- els are few-shot learners,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Language mod- els are few-shot learners,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.502949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.130980Z digest=sha256:a522ae77e420fab6ed2f8477691be999a50d3e69c6d22cbfa720afa50183e088

Observation 27dfbcdd-c85e-432c-8bf1-f87a340a9552 · outbound

This paper cites Evaluation of the efficiency of state-of-the-art speech recognition engines,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Evaluation of the efficiency of state-of-the-art speech recognition engines,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.490155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.133861Z digest=sha256:00fc055a587b1dbdd055ce3c24e7be508fdd3c6d775ed43c83199d19450a351a

Observation 62fc647c-412b-433d-ab61-4ba302fde950 · outbound

This paper cites Hand and arm gesture- based human-robot interaction: A review,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Hand and arm gesture- based human-robot interaction: A review,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.478557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.136524Z digest=sha256:232b56256a622d93de3458f5946b4aac89f5b7af76114542a1f7ece6402d0943

Observation 89c588d9-2eeb-4f4f-9021-56dddfcec6ab · outbound

This paper cites You only look once: Unified, real-time object detection,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model You only look once: Unified, real-time object detection,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.139938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.139938Z digest=sha256:41e184df19f05dccfdd8173792b8b21d6d8845ccc6f87e48b84e2874b6d9cf9f

Observation 9af4ce04-1d1c-4686-94ef-3446478ef5e8 · outbound

This paper cites Yolo-world: Real-time open-vocabulary object detection,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Yolo-world: Real-time open-vocabulary object detection,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.143612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.143612Z digest=sha256:762f668da41e7f0ffaf599e7f1aa51cf94d95f64b920d45879afa70555fa34c4

Observation 0f1214f0-2c30-442f-9ecb-8c06d34f8d10 · outbound

This paper cites Realtime multi-person 2d pose estimation using part affinity fields,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Realtime multi-person 2d pose estimation using part affinity fields,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.147126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.147126Z digest=sha256:2905a198f3b80936e0c4d73f4b914f2cacb08a678b0bed18caf0537d3059697f

Observation 1bc1d1a8-ae67-495f-8e36-60dfda88a907 · outbound

This paper cites Autonomous object level segmentation,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Autonomous object level segmentation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.446082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.150681Z digest=sha256:38772baef797e0e74eea90cf8941a7fffa4fc37966d6ea2f8fe82b7e7139349c

Observation 7f1ea820-4f0b-4681-a6fd-d64853af6beb · outbound

This paper cites Airslam: An efficient and illumination-robust point-line visual slam system,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Airslam: An efficient and illumination-robust point-line visual slam system,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.435205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.154360Z digest=sha256:5954757aad54fccd5c54591b573a1046d6e5f7f944e5ef0d962d12e80ff4a03b

Observation 5843b6df-3874-4cbc-9347-939a582f0609 · outbound

This paper cites Relative localiz- ability and localization for multi-robot systems,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Relative localiz- ability and localization for multi-robot systems,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.424521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.157726Z digest=sha256:3944373a31a0c9bc68ec28433a60616effaa91d677ab3ff4969ea254a1b668ce

Observation 78aa660b-3d34-4589-ac4c-b457ab40a28a · outbound

This paper cites Large-scale uwb anchor calibration and one- shot localization using gaussian process,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Large-scale uwb anchor calibration and one- shot localization using gaussian process,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.411778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.160837Z digest=sha256:189bd4330bd894206368db42337cd55ef3ab1a9775fc97ce4291359ea91cdf73

Observation 751fcb3f-b8d1-4bc2-889a-297889d6f896 · outbound

This paper cites Helmetposer: A helmet-mounted imu dataset for data-driven estimation of human head motion in diverse conditions,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Helmetposer: A helmet-mounted imu dataset for data-driven estimation of human head motion in diverse conditions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.399328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.164396Z digest=sha256:ad2c9de91dbbf00f09222e902075f5539f6fdac5208be786786d31155b3860e6

Observation 2c7b8bef-9bc3-4607-a15a-2f7d666c6e99 · outbound

This paper cites Heterogeneous stereo: A human vision inspired method for general robotics sensing,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Heterogeneous stereo: A human vision inspired method for general robotics sensing,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:46:48.386740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:46:48.167797Z digest=sha256:b00ef7cad839e360b2e008302748ee3e9a56ad53c864e8eb71bdc189db160ccc

Observation 444ebb34-c38f-4a7d-9203-86e37f408d9a · outbound

This paper cites Nvp-hri: Zero shot natural voice and posture-based human–robot interaction via large language model,.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model Nvp-hri: Zero shot natural voice and posture-based human–robot interaction via large language model,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.170804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.170804Z digest=sha256:8091fbfc206bc922bc7cd4f62df874496ea0411725204a191940af87ddcc8c4d

Observation 829ca9da-7935-4272-bb01-b912871b234b · outbound

This paper cites FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech.

Natural Multimodal Fusion-Based Human-Robot Interaction: Application With Voice and Deictic Posture via Large Language Model FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T22:46:48.174002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:46:48.174002Z digest=sha256:b3326c8dea9dd07dcf84a6d807bb74bf28688188fb52e59176e5a9fe9450cb5c

Pith citing papers

No inbound Pith citation observations are available.