Pith. sign in

Paper Citation Record · LEDGER

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2506.20757.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20757 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:47:45.226599Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T12:52:15.138790Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T12:53:17.415884Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy16
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a2dbe33b-d4a9-4342-a152-74f7379a3cd3 · outbound

This paper cites The objectfolder benchmark: Multisensory learning with neural and real objects,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations The objectfolder benchmark: Multisensory learning with neural and real objects,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:48.530388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:41.983648Z digest=sha256:795381e80043e55ad2796a8935db6f5d61ce126e9ed69d6f7862f6273f992e5c

Observation 4da439d4-45c3-426e-9b5f-4cbab7ba9f81 · outbound

This paper cites Robotic tactile perception of object properties: A review,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Robotic tactile perception of object properties: A review,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:42.042535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:42.042535Z digest=sha256:4fd4cab32d342f609c051439907201eb2388336e884e90d0b466fe9c175ee4cb

Observation 4154fa40-a8a9-4c4e-acb7-d99e1b294349 · outbound

This paper cites Deep domain adaptation regression for force calibration of optical tactile sensors,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Deep domain adaptation regression for force calibration of optical tactile sensors,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:48.371959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.108192Z digest=sha256:cc88a1267e97faa92b532c77141180958b455481ec00f9f50940192549128a10

Observation f520930e-25ab-4f94-8a02-e95d851df094 · outbound

This paper cites The feeling of success: Does touch sensing help predict grasp outcomes?.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations The feeling of success: Does touch sensing help predict grasp outcomes?

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:48.247438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.160106Z digest=sha256:0f595bd7478bceab976d91eae6dc3b51120a42e6e83f5014d252f292c01c2a4f

Observation edc4bb20-8b94-496a-a9f9-77f94d135ff3 · outbound

This paper cites Spatio-temporal attention model for tactile texture recognition,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Spatio-temporal attention model for tactile texture recognition,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:48.114054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.288000Z digest=sha256:1302bdf1c11ce0fd79ca652235338e3b18cc3135f57e5d3de49b1d5e84e2874f

Observation 277a0eba-7f2b-4917-8425-ffd971fa4dea · outbound

This paper cites Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:42.351404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:42.351404Z digest=sha256:e4a2cb613981347b08f22f4d2acea25c6797f4e037ff98c8e705a60b7e8d2792

Observation 6afc803a-f56e-4201-82ed-1c0a607cfcaa · outbound

This paper cites Vitac: Feature sharing between vision and tactile sensing for cloth texture recognition,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Vitac: Feature sharing between vision and tactile sensing for cloth texture recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.957776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.441201Z digest=sha256:357ec186306c0b9e8c4691b0fcc5e228992f9ce3f720ff971b8d4aefd529e18e

Observation 1a79521f-a2b5-4693-ae2a-35fa4ef736d8 · outbound

This paper cites Gelsight: High-resolution robot tactile sensors for estimating geometry and force,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Gelsight: High-resolution robot tactile sensors for estimating geometry and force,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:42.536949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:42.536949Z digest=sha256:d003f9b1135fb1e36b6d12efea828144c2d55e8f12b7ffdb1acef33c87a1d87d

Observation 1e81ef4f-8dd5-4088-912e-a1d52e5782ee · outbound

This paper cites Geltip: A finger-shaped optical tac- tile sensor for robotic manipulation,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Geltip: A finger-shaped optical tac- tile sensor for robotic manipulation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.793020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.616582Z digest=sha256:540294daf9f949970dc7ef1ff31f00e7392918821c563edccfad172dea5fe70c

Observation c5e5deb8-56ed-4fb3-8588-be017b34e46a · outbound

This paper cites TouchRoller: A Rolling Optical Tactile Sensor for Rapid Assessment of Large Surfaces.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations TouchRoller: A Rolling Optical Tactile Sensor for Rapid Assessment of Large Surfaces

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:47:45.654026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.696229Z digest=sha256:b6305f633cfa19edf2ca67d1e17b24fbbbb84b855516adcf8f2f0aec7f1532b3

Observation 7a537dac-b045-47d4-9ab7-91796330a325 · outbound

This paper cites Self-attention based visual-tactile fusion learning for predicting grasp outcomes,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Self-attention based visual-tactile fusion learning for predicting grasp outcomes,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:42.822767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:42.822767Z digest=sha256:628af10c0b3a7ba9f8aafc114e9f5d3834ff4879cb589faef1ce44ec156e8260

Observation bfea05a0-45bb-440d-b8d0-ebf764fb2b5f · outbound

This paper cites Touch and go: Learning from human-collected vision and touch,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Touch and go: Learning from human-collected vision and touch,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.666861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:42.915679Z digest=sha256:ce093b6cfb721159eb85651c4447abd1bc791414463161146d8a57714d9ef727

Observation 5ed7388f-1502-4e34-a916-254c4ab1ec05 · outbound

This paper cites Self-Supervised Visuo-Tactile Pretraining to Locate and Follow Garment Features.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Self-Supervised Visuo-Tactile Pretraining to Locate and Follow Garment Features

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:43.020932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:43.020932Z digest=sha256:5085dd3e639985724e7c5d93f0ae7251c08fd267393385c58d4f5299b767d88a

Observation 4c423e40-77c8-47bd-b4e6-86222056ff5e · outbound

This paper cites Multisensory integration: How visual experience shapes spatial perception,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Multisensory integration: How visual experience shapes spatial perception,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.538582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:43.126146Z digest=sha256:56003b592a83a3e0404537116b4d4fb92b87b25a025e57e2739c51ccbf5e2eda

Observation fecdef0c-f844-4c2a-93b1-70f248893e7c · outbound

This paper cites Learning transferable visual models from natural language supervision,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Learning transferable visual models from natural language supervision,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.420916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:43.234130Z digest=sha256:e36d80dad0fec5f36e9a3584d8eac62fd5b725cf6ae958df4e1c2ed29a13503f

Observation 9fe73797-e9b0-42da-85fb-d504923d5cb2 · outbound

This paper cites A simple frame- work for contrastive learning of visual representations,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations A simple frame- work for contrastive learning of visual representations,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:43.342993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:43.342993Z digest=sha256:d55a10406ad66c71a7b55cf3f39d562c9c8f6b4802ee76688d5997aa770e7a6e

Observation 77264659-08b8-4501-a9ac-e35cdfdabde9 · outbound

This paper cites TransForce: Transferable Force Prediction for Vision-based Tactile Sensors with Sequential Image Translation.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations TransForce: Transferable Force Prediction for Vision-based Tactile Sensors with Sequential Image Translation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:43.488139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:43.488139Z digest=sha256:1e278991568cc24c2d54c593d8081c29d41df1fa40e0c0953b60e1073afe985d

Observation 209146a6-735f-48e4-a417-f406356d5e28 · outbound

This paper cites Learning continuous grasp stability for a humanoid robot hand based on tactile sensing,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Learning continuous grasp stability for a humanoid robot hand based on tactile sensing,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.295965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:43.601123Z digest=sha256:835501dc95c354335de3d0dd3040f62b322208707352da24e41cad18882b8320

Observation acb6eae7-dcae-46a7-b277-dbda8000389a · outbound

This paper cites Visuo-tactile transformers for manipulation,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Visuo-tactile transformers for manipulation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:43.736262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:43.736262Z digest=sha256:7050e6b93ef145102e62c666b4d6ecbfcf739ab0b7e6374e106e957b1f010e48

Observation 34c90ca0-09da-47ac-b1b9-3a46c6a3f1a6 · outbound

This paper cites Contrastive multimodal fusion with tupleinfonce,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Contrastive multimodal fusion with tupleinfonce,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:47.138633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:43.846608Z digest=sha256:225e05df3086b6273b69c1b9b4b4621e4785209a4b9589baa7e362fadedc3e4a

Observation 4e0d04e6-a7da-4f75-bf9a-66dbd6ece3e3 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Representation Learning with Contrastive Predictive Coding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:43.932874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:43.932874Z digest=sha256:94a5c2a47d3c0813157a34d0afa3194595bdc73c2bb053b5e05a824d49433dec

Observation 96e95bd5-fecf-4862-aa4f-412e83c8d28a · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Momentum contrast for unsupervised visual representation learning,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.016226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.016226Z digest=sha256:135351f942186948ebbcfb65f14e5f4a122af0621a4bacaf6c0b222e2116bcf3

Observation 4399acf3-4313-4f3e-8a82-550143f4b944 · outbound

This paper cites Binding touch to everything: Learning unified multimodal tactile representations,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Binding touch to everything: Learning unified multimodal tactile representations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:46.979552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:44.126289Z digest=sha256:18e4eb2802463cf6b551af84790373c8b29f5656e80e508d54b6415fca9af870

Observation 604c3d6b-cc03-4bfe-86e7-1876adb6feb8 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.205254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.205254Z digest=sha256:75a8d89233396abf8a34f0619e3dd27b45110d2d38babca0de94fc49c3fee382

Observation bcf0d628-84e7-4567-8b88-0153adc2d775 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.257788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.257788Z digest=sha256:803a7577ddc66d999c87bc8dfbf259e6b5a05dcb66fe013216aed8f85ae1d912

Observation 370b9bfd-4228-467e-b109-830a7d19d7c1 · outbound

This paper cites S 3m- net: Joint learning of semantic segmentation and stereo matching for autonomous driving,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations S 3m- net: Joint learning of semantic segmentation and stereo matching for autonomous driving,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:46.785858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:44.368680Z digest=sha256:60cdc40b4103792b93cab798e5f7adf4a856113cb9d003a1c703608c77d299be

Observation f4bebb64-4940-4939-a984-b952d7bff1b3 · outbound

This paper cites Sg- roadseg: End-to-end collision-free space detection sharing encoder representations jointly learned via unsupervised deep stereo,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Sg- roadseg: End-to-end collision-free space detection sharing encoder representations jointly learned via unsupervised deep stereo,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:46.460813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:44.475352Z digest=sha256:0397b6c4da68971b82a1a0c33756b4a3830514b41fcd17bdf1555768dfc2d996

Observation 9746fc29-52e7-43bb-ad21-6aecfcf96e4b · outbound

This paper cites High-resolution image synthesis with latent diffusion models,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations High-resolution image synthesis with latent diffusion models,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.592287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.592287Z digest=sha256:d4182fb12c5d4a56327bc64a966802567a88c5fe2de55cffc419e095319cff81

Observation 6050e495-9d76-49cc-8835-4ef550030496 · outbound

This paper cites CDI3D: Cross-guided Dense-view Interpolation for 3D Reconstruction.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations CDI3D: Cross-guided Dense-view Interpolation for 3D Reconstruction

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:47:45.452456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:44.668987Z digest=sha256:a699cab1f00c3bbb09462d9d445b3c020a249a0e42b2ceec865b301f83224bca

Observation 2f76bac4-72bc-4363-ab31-105ee86d963f · outbound

This paper cites DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.736918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.736918Z digest=sha256:5425db21495d78867f82574d4e1291c9857bdc7a319760900db12de95c6799b3

Observation 2a1bb092-e85e-4dc5-bc38-0a6b68b2f7e2 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Adam: A Method for Stochastic Optimization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.818022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.818022Z digest=sha256:552bd1a8ee12fe2cd32c230e35f6bbdb245fe9cf90e6d0245a5cfffd38677640

Observation a484fe48-fcb4-4b23-a25e-a1833cb1e7d8 · outbound

This paper cites Deep residual learning for image recognition,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Deep residual learning for image recognition,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:44.938411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:44.938411Z digest=sha256:2027d59838c40a2e34cef04f5639879af54d2843e7c1306fe68ab80d9cc1be32

Observation 6fc08bea-2af6-445f-b57e-f9346f54b3b7 · outbound

This paper cites Two-dimensional pca: a new approach to appearance-based face representation and recognition,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Two-dimensional pca: a new approach to appearance-based face representation and recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:46.207111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:45.024772Z digest=sha256:e096b5a0e806a2c121f1a8028e9822db02482126c470e3e72e5a34cc5da2cbdd

Observation 39b5e1df-51b7-47c5-92d7-ed1f48250906 · outbound

This paper cites Learning deep multimodal feature representation with asymmetric multi-layer fusion,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Learning deep multimodal feature representation with asymmetric multi-layer fusion,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:47:45.949918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:47:45.135455Z digest=sha256:00b2bf78f1d4061543b136f6255f61c8a41ad3fbe70320a8378e5feaac8a15fd

Observation bc42c198-51b5-42c7-9bcf-e870d74da517 · outbound

This paper cites Multimodal zero- shot learning for tactile texture recognition,.

ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations Multimodal zero- shot learning for tactile texture recognition,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T22:47:45.226599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:47:45.226599Z digest=sha256:34a4e877d0c28fd266bf83d0563af41d8f2eb1846b67b2af1461abdbecf08bc2

Pith citing papers

Observation c06afac8-7083-4411-baa8-72f1f7235935 · inbound

Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms cites this paper.

Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:53:17.418098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T12:52:15.138790Z digest=sha256:f9b053dbab0968268b0774c557620d2111d6733edb010011adb656b5a4bd1710