Pith. sign in

Paper Citation Record · LEDGER

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy

As of 5 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2304.11193.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.11193 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-24T09:31:34.871215Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact7
  • verified fuzzy35
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 964f22e7-8ba9-4d64-96dc-440a344e110f · outbound

This paper cites Stochastic Variational Video Prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Stochastic Variational Video Prediction

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:34:16.962494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:75db01f487b44a683f7f739f188c2461e2674e241b0b43f6cfe3b0bb531beeef

Observation 2b7d0b99-ad1f-466d-ba51-473cbd7140c3 · outbound

This paper cites Recognising action as clouds of space-time interest points.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Recognising action as clouds of space-time interest points

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.791172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:7efff2dd8ad3e140f45ffe301cd0443ef6f051da720c7bec6038a89e0b4b41d3

Observation 5ae496b3-c122-4aa8-be74-5eacd08ed1ca · outbound

This paper cites Semantic object classes in video: A high-definition ground truth database.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Semantic object classes in video: A high-definition ground truth database

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.782685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:d6095fc6e94112ce8580edbbf937f59168d40b3b6851a04a39393c418e7c0621

Observation b953916f-0ff2-45f8-b0f6-32336ff07726 · outbound

This paper cites Visual-tactile cross-modal data generation using residue- fusion gan with feature-matching and perceptual losses.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Visual-tactile cross-modal data generation using residue- fusion gan with feature-matching and perceptual losses

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.857830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:c5136d1b2d4afe5f3592f674216ee5880c36a7d7d3c515152e2dc4b016cb49e5

Observation a34cd83c-6674-4c9e-a1b0-8f26d6c487c0 · outbound

This paper cites The ycb object and model set: Towards common benchmarks for manipulation research.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy The ycb object and model set: Towards common benchmarks for manipulation research

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.854696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:c29cec30c13d017300e81afaee5fb6680eb7ea066b45154ed50c5c386f6c2cc0

Observation 99a34c11-a5dc-4acd-bc09-3d1ca1f3b107 · outbound

This paper cites RoboNet: Large-Scale Multi-Robot Learning.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy RoboNet: Large-Scale Multi-Robot Learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:34:16.957694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:6fc58b2f87558089498dd663b2856cd455553d3bc5cf7238b73722b671a79750

Observation 5578e691-3cd1-4769-8b6d-00274cb976a7 · outbound

This paper cites Stochastic video gener- ation with a learned prior.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Stochastic video gener- ation with a learned prior

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.851135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:2b487561ca64a1899bfcf025f7f900bb7a2130557d41fd1b3ee691652d7c3ea8

Observation 9448d66a-71c2-4774-944e-6ea2f062f2d1 · outbound

This paper cites Self-supervised visual planning with temporal skip connections.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Self-supervised visual planning with temporal skip connections

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.847922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:f456d89e0d66e55a627d49cb5e778bf4b1e7a5fa051b0e1b84dc7fce4de8e736

Observation a5c53439-2f54-4d7e-9c6c-626c06cfbad8 · outbound

This paper cites Unsupervised Learning for Physical Interaction through Video Prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Unsupervised Learning for Physical Interaction through Video Prediction

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:34:16.971721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:72eccbd56a92791d1c8cd06a1d3210b93ea1cf0e3cace6645741ee51a622a247

Observation db3adc37-5836-47c9-8e24-021a0d87a711 · outbound

This paper cites Vision meets robotics: The kitti dataset.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Vision meets robotics: The kitti dataset

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.844020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:612f3153ba6ec2d79b069c00a884922463c6bb1ae326f6694c027d35103f8a2b

Observation 5218f110-a678-480a-b86f-48bf91337017 · outbound

This paper cites Possible anatomical pathways for short- latency multisensory integration processes in primary sensory cortices.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Possible anatomical pathways for short- latency multisensory integration processes in primary sensory cortices

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.840537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:13b5ccbc39794f2ded7b779c719418be6d7ceec0576fb64527376790a8699db4

Observation f4ede5a6-6f25-4d08-9a44-1cba3bab29b6 · outbound

This paper cites The apolloscape dataset for autonomous driving.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy The apolloscape dataset for autonomous driving

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.837114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:d6f49a29a48cedb9c3716e427e31184898f383b53231c57b55a7de97f2d3bbd5

Observation 76c725d3-53fc-4db7-bf4f-2a13329619bf · outbound

This paper cites Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.833796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:256dba0eed9c79f3865ab2bc15b0bd3dfcefc7061b4056f28f7d6c4d8fc35225

Observation fd3dbf21-7e12-4ffa-9af8-6f8a2b117aaf · outbound

This paper cites Coding and use of tactile signals from the fingertips in object manip- ulation tasks.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Coding and use of tactile signals from the fingertips in object manip- ulation tasks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.829928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:2e79ee719713049e4f5376acb7d86e60191ae6eabb9a35c18f146908f52a838a

Observation 0f1044c5-2b59-4bb6-9f93-dcdc7a6969f2 · outbound

This paper cites On infor- mation and sufficiency.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy On infor- mation and sufficiency

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.826494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:16eda03b8a169661054884aa6533b8d45eb2faa0637f1b9fa2256171cb9b1b58

Observation f6de1a81-8bb8-4b2d-9038-dcc7a20d6ae0 · outbound

This paper cites Stochastic Adversarial Video Prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Stochastic Adversarial Video Prediction

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:34:16.980675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:535edbfe22a0431e09c58c0264661d3cacde83e5d69150f00dadd91e8508c31d

Observation 6941889e-abd6-45c8-a352-51cd2ad92eb8 · outbound

This paper cites touching to see.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy touching to see

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.823450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:3c2200bc6f1cc016f98b98b557dd1e4eee634db5e7c248a91f844cd4f4adb340

Observation 90d91ffd-243f-47be-87cd-10aced3df15d · outbound

This paper cites Making sense of vision and touch: Learning multimodal representations for contact-rich tasks.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Making sense of vision and touch: Learning multimodal representations for contact-rich tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.820435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:6fd0261f157de3968b2b29ccda1a125821361e4f20c287fda3e0964eb152c261

Observation c767bc88-ad98-42e5-94a5-519d54342266 · outbound

This paper cites Connecting touch and vision via cross-modal prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Connecting touch and vision via cross-modal prediction

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.817587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:f1d3ad1b652200f46dd758d156bcb95b3415e836d69ce1c9e5a857e6e4d8b1c4

Observation 7bc49ef1-a5e0-4728-b9f9-f878fe6836e9 · outbound

This paper cites Action conditioned tactile prediction: a case study on slip prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Action conditioned tactile prediction: a case study on slip prediction

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.814665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:ec00638c7dead07d8589cd613eaa0f8a262af364d416e8a46f82818c5254aabb

Observation 60c07f7e-a043-46da-8235-3b6e63793502 · outbound

This paper cites Proactive slip control by learned slip model and trajectory adaptation.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Proactive slip control by learned slip model and trajectory adaptation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-24T09:34:16.966918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:596b8273cc747cdd88f7ee643c6f9d1b4c57154e938b27a4a3978d3d4586cf0e

Observation 30d7784d-b8aa-4647-8885-0631bf54234c · outbound

This paper cites From Active Touch to Tactile Communica- tion: What’s Tactile Cognition Got to Do with It? Danish Resource Centre on Congenital Deafblindness.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy From Active Touch to Tactile Communica- tion: What’s Tactile Cognition Got to Do with It? Danish Resource Centre on Congenital Deafblindness

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.811913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:35399bb1147b52bc0db6211d396f24a42a108b5fe49af499d4766271c27533df

Observation db962440-11c2-4f55-a224-d537bdab7ce3 · outbound

This paper cites Action-conditional video prediction using deep networks in atari games.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Action-conditional video prediction using deep networks in atari games

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.808928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:d2d4c0fc91ade004702b57865a3c30222b68f849eccff8bbea2ed70214ef079a

Observation eacc901a-6bc7-40f6-9893-2278fe62fe19 · outbound

This paper cites Sensing characteristics of an optical three-axis tac- tile sensor mounted on a multi-fingered robotic hand.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Sensing characteristics of an optical three-axis tac- tile sensor mounted on a multi-fingered robotic hand

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.806190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:124c3fe69c6a1660eeb3a23a95e2fc47afe8ae2df5fdcf880d86656dc6ad59b5

Observation 0d821e83-0fa5-49ca-ad90-9e6223414eac · outbound

This paper cites A review on deep learning techniques for video pre- diction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy A review on deep learning techniques for video pre- diction

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.803408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:dff221c3788e9b4e15f9e640ee7c70f337a278e34963ea3b826fc2d668c1dd7a

Observation b844a269-2d30-4c5b-beb8-3dcefc51d41b · outbound

This paper cites The curious robot: Learn- ing visual representations via physical interactions.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy The curious robot: Learn- ing visual representations via physical interactions

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.797081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:46dc4a30edf6d5d3f0d81bad5953284ea78979221d2326ed3f72653d5aba9496

Observation c312b4e0-93cc-4882-bed2-268138f149b5 · outbound

This paper cites Video (language) modeling: a baseline for generative models of natural videos.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Video (language) modeling: a baseline for generative models of natural videos

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:34:16.952782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:25f4a95d0defccf8d81d2138ab81c41733b7b5873f96ff0beed44b6aab63f56d

Observation df196a3f-05f7-43cc-9314-63c753b37cff · outbound

This paper cites Xela robotics uskin magnetic tactile sensor.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Xela robotics uskin magnetic tactile sensor

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.793952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:cfb06200da042344ba5fdf6d59f929363666ec03519d615075600ab555d537cb

Observation af2eb7c4-8731-4cf8-9bcb-8004e940718c · outbound

This paper cites Recognizing human actions: a local svm approach.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Recognizing human actions: a local svm approach

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.788423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:1c9a2a5745c45d08cf8c785b033a8b06198125be6dee6fc5e3e64fe5b2305c30

Observation c8c5d036-a710-4354-b4f5-0c0efe207ed8 · outbound

This paper cites On the design and development of vision-based tactile sensors.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy On the design and development of vision-based tactile sensors

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.785494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:4f7c17dc1f1e6e1da975e26c8860df145d0ac8d3d47844b210b720da888ca34a

Observation 6a6584aa-c2df-481c-bb68-278f34aef582 · outbound

This paper cites Unsupervised learning of video repre- sentations using lstms.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Unsupervised learning of video repre- sentations using lstms

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.779801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:c895ed6b88ab824d18b7315f65e6eddb8925a0460c85b92d1f6ac505f228cdda

Observation a3fff6a0-01bd-4c32-8e53-f5dc661c8ae9 · outbound

This paper cites Learning of action through adaptive combination of motor primitives.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Learning of action through adaptive combination of motor primitives

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.776956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:480a8e904b4eed4f5259c742f0788c3721f765eca8af0eb422ae428c827df35b

Observation bde58240-4083-4de5-8e8a-8e9313b87f51 · outbound

This paper cites A review of tactile sensing technologies with applications in biomedical engineering.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy A review of tactile sensing technologies with applications in biomedical engineering

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.774323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:48bce7ca706a93933d8f971a77d18375a902932eaf43a7936bf4af8e84a9b5d0

Observation e672871f-6bdf-41b0-9c03-f394cab9bed7 · outbound

This paper cites Sensory prediction errors drive cerebellum-dependent adaptation of reach- ing.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Sensory prediction errors drive cerebellum-dependent adaptation of reach- ing

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.771500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:70683739217b10cb278075b0a07ed9b2800e0ac2b6bd5a2e4ec75f9e7592eb53

Observation e8687343-b78f-4de5-8ce9-afbfc55d8f9f · outbound

This paper cites High fidelity video prediction with large stochastic recurrent neural networks.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy High fidelity video prediction with large stochastic recurrent neural networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.768577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:f372973e6713db99ac15a35d26e0309b7a610e4005a1cb53f3e4953e6639da99

Observation af9875d5-bba9-43aa-a7c8-18b2bb1aa1b4 · outbound

This paper cites Decomposing Motion and Content for Natural Video Sequence Prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Decomposing Motion and Content for Natural Video Sequence Prediction

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:34:16.975899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:7256ea213f9e6afb530a2688f529cf9a7697934d09a47fc4490692af2832c52f

Observation 4480048e-1f0c-4e0c-881e-0fbf42cb7970 · outbound

This paper cites Learning to generate long-term future via hierarchical prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Learning to generate long-term future via hierarchical prediction

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.765546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:d354c4f7a8286748bb03dfe54a7fbe4b646930fb71e7801bfcdb420268372cbd

Observation 3382f763-a802-4bae-8053-111b3121717c · outbound

This paper cites Tactip—tactile finger- tip device, challenges in reduction of size to ready for robot hand integration.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Tactip—tactile finger- tip device, challenges in reduction of size to ready for robot hand integration

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.762015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:03e3a612d20f41a7c8c13d37ba293ec1cc5dfb3e2f4aae5d9e3b141d780c5ff8

Observation e5ab922e-d5d3-43a3-aa76-6a0d0a7a7322 · outbound

This paper cites Motor prediction.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Motor prediction

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.758564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:483d095a4b89afb9a99212975eb9d493960bf801cac053b04915c3eeabb453a9

Observation f8f8a923-dace-4328-8b7e-2dfccac9d420 · outbound

This paper cites Gelsight: High-resolution robot tactile sensors for esti- mating geometry and force.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Gelsight: High-resolution robot tactile sensors for esti- mating geometry and force

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.755442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:12edc2dc8e55a1986f4bddb73c64df6165df83a6b2356e867509b1a8a9ee32d9

Observation cf11080c-3170-44b1-90f5-5ec65dd6c202 · outbound

This paper cites Sen- sorization of robotic hand using optical three-axis tactile sensor: Evaluation with grasping and twisting motions.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Sen- sorization of robotic hand using optical three-axis tactile sensor: Evaluation with grasping and twisting motions

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.752281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:310e91d5f1d671e58b647215103824b713d2f93ab6987ca3bb212ae78bb0dec9

Observation 49c03405-2ce5-4730-a925-a3dd71e0b259 · outbound

This paper cites Learning to predict friction and classify contact states by tactile sensor.

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy Learning to predict friction and classify contact states by tactile sensor

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-24T09:34:17.800362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-24T09:31:34.871215Z digest=sha256:5a737f523739b357f008d410d9fad34bea7c0329c47206b358ff3ea80fb26d98

Pith citing papers

No inbound Pith citation observations are available.