Pith. sign in

Paper Citation Record · LEDGER

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining

As of 6 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 1 inbound Pith citation observation for arXiv:2602.00937.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.00937 v3

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T08:24:44.943709Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T03:16:33.702802Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T18:08:46.857611Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact19
  • verified fuzzy50
  • unresolved1
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7bd41ea1-fded-4706-a238-697e91a2981e · outbound

This paper cites Big vision.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Big vision

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.316088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:6d41ca3f82745fc631943f15dd5d86f98e248182778c838c9a342e7accd789f6

Observation 7ee11045-bf6a-4a87-b6a7-108aa68eb244 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.903981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:aa2ffecb6f041f6f90b9e7906e11ea1b21c71b5c42b446299330f99aac307213

Observation 58d325f0-6ec9-4a7f-80fc-b8ec4c002598 · outbound

This paper cites JAX: composable transformations of Python+NumPy programs.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining JAX: composable transformations of Python+NumPy programs

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.249407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:b9aba523aa5e54553c386866064907a3bbe1bf2a3c8c50dbdb26081f50d89a49

Observation a337a850-aeb6-4fdc-b08b-b6d5812b12c0 · outbound

This paper cites Internvl: Scaling up vision foun- dation models and aligning for generic visual-linguistic tasks.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Internvl: Scaling up vision foun- dation models and aligning for generic visual-linguistic tasks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.288198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:eb1aea3083dc579a0815d2b969759044b88fa9aea7ffba362eb60b080d618463

Observation bb933ee7-ea04-4cc2-9e95-743713968bb8 · outbound

This paper cites Reproducible scaling laws for contrastive language- image learning.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Reproducible scaling laws for contrastive language- image learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.331038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:9fdc5079f1dda299da06c4e9c0424f98531a591457d76fda8a29663ed7a74f0e

Observation f8efb56c-a83c-46eb-b1e3-ad0c2d3a5113 · outbound

This paper cites Dif- fusion policy: Visuomotor policy learning via action diffusion.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Dif- fusion policy: Visuomotor policy learning via action diffusion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.314279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:f34ebdf8257262c0c981d0755c187f45e15736713d81dc099b706e2cb39eaed3

Observation 5167d138-83dd-4a06-858c-1588e01c6c2d · outbound

This paper cites author Dong, W.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining author Dong, W

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:27:36.448205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:74e19060ad99fca80d5d7644b6039ad24c50f29dec727f9f69d02a2e3402737d

Observation 5fa29b93-1576-40c1-86c3-ca6ba4a7d7ad · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining An image is worth 16x16 words: Transformers for image recognition at scale

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.357613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:7c5675fb3f23e4ae6f574711d074a0803d4ebcb1cb6c90bb1e38cd3382ec8559

Observation b898bb92-15c7-4289-97f9-6a08e5b0ce7b · outbound

This paper cites Eva: Exploring the limits of masked visual representation learning at scale.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Eva: Exploring the limits of masked visual representation learning at scale

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.343081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:a3349c5c97f713ffee25c6cd6afb88f04943e6a6aedc1e0c695865e4aa3922ac

Observation 67131aa8-f0a7-49ea-838e-4e193ed5fcd8 · outbound

This paper cites Eva-02: A visual represen- tation for neon genesis.Image and Vision Computing, 149:105171.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Eva-02: A visual represen- tation for neon genesis.Image and Vision Computing, 149:105171

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.307229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:6325ae793aaf56ee34600d0d4974c81f7dff759a8053146148a229ad49d2e754

Observation 9e1fdcae-4494-4c42-a320-d0e9b8d50d4f · outbound

This paper cites Act3d: 3d feature field transform- ers for multi-task robotic manipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Act3d: 3d feature field transform- ers for multi-task robotic manipulation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.313442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:22edc238f11af89d06eccf628837d984f83332cee98c03c87f3554d0d3ccc309

Observation 88193174-f803-4904-88ea-86de96acbe9e · outbound

This paper cites Revisiting point cloud shape classification with a simple and effective baseline.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Revisiting point cloud shape classification with a simple and effective baseline

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.355423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:02db4b874b2793745971165f49ccc50d5afeebc94f9ad71da034a76deea2051c

Observation 32c09b2c-dc2d-4826-9468-7af6653f0cb2 · outbound

This paper cites Rvt: Robotic view transformer for 3d object manipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Rvt: Robotic view transformer for 3d object manipulation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.339217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:7af8aa435d75315150a52084b7a641d950545ad14a5356692e410508037af9a1

Observation 4f3fda71-5160-4826-99aa-162edb970956 · outbound

This paper cites Learning dense visual descriptors using image augmentations for robot manipulation tasks.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Learning dense visual descriptors using image augmentations for robot manipulation tasks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.304933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:58da72458206538a9ab91c972e26748f7174923f9259e03f5c4679a3f8f50752

Observation d8850130-7c24-4f2d-8cf7-b11361967db8 · outbound

This paper cites Pct: Point cloud transformer.Computational visual media, 7(2): 187–199.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pct: Point cloud transformer.Computational visual media, 7(2): 187–199

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.325048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:bf3a0f6306ad1e0be9f9596861027eff2c9212834d685160ab1ec833a9eb1291

Observation e5cc3937-0a6b-476e-9f3b-8a1b2722eb1e · outbound

This paper cites Mvtn: Multi-view transformation network for 3d shape recognition.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Mvtn: Multi-view transformation network for 3d shape recognition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.351117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:606b5b30b281745be3935aaabcd755a28846d24056b41b281c77c6527d1341a8

Observation faf9ce29-ce5e-4ab1-b4f1-2fef47304470 · outbound

This paper cites Voint Cloud: Multi-View Point Cloud Representation for 3D Understanding.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Voint Cloud: Multi-View Point Cloud Representation for 3D Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.877884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:056078f8cda3fc48c29062e4963b1ab3485cccd0cf598e57aaf429b848c5d15e

Observation 6dbecb7a-20aa-46e5-8367-ed134d8c4b97 · outbound

This paper cites Denoising Diffusion Probabilistic Models.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Denoising Diffusion Probabilistic Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.891507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:e5cfacb1c1f851658f0d0cecf815f808afe5f52b7945a5b8b9cf9f35acb8c2e5

Observation e7acc51b-0eb7-4133-b77c-cfd174f56f59 · outbound

This paper cites Pri3d: Can 3d priors help 2d representation learning? InProceedings of the IEEE/CVF International Conference on Computer Vision, pages 5693–5702.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pri3d: Can 3d priors help 2d representation learning? InProceedings of the IEEE/CVF International Conference on Computer Vision, pages 5693–5702

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.309755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:f6183f84b2bae9af9aee8e5a47ac43e01301dcadf65dbb313450fecb5b1154fe

Observation ac154d1c-187d-49da-8043-31e0454d3f21 · outbound

This paper cites A comprehensive survey on contrastive learning.Neurocomputing, 610:128645.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining A comprehensive survey on contrastive learning.Neurocomputing, 610:128645

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.346842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:88a08babf76dbf2b252179c4100e117fef324adb096b43114edc550182b01b61

Observation d9901db9-9d7a-4c63-971e-ae81051efcb5 · outbound

This paper cites An Embodied Generalist Agent in 3D World.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining An Embodied Generalist Agent in 3D World

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-17T14:22:18.773673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:bcfbc6a59d458c8c91caf3e1775b2a1458d6f4e3c256efad090c18c958e02a9a

Observation 0e09e095-f0c7-4a87-b55e-dce7732c0ee7 · outbound

This paper cites Multi-view transformer for 3d visual grounding.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Multi-view transformer for 3d visual grounding

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.323054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:62ea1795c05f29ba054d5ada4726801f6852df3885c3a6f61c18e77f2d08df36

Observation 01779527-60e1-444e-8431-cc55ffcd04bc · outbound

This paper cites Frozen clip transformer is an efficient point cloud en- coder.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Frozen clip transformer is an efficient point cloud en- coder

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.344933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:bb6344679ce0b453fc328e21894fd90324a2e21c5487249a5526aaf5b971c258

Observation ff30f187-bf5b-4e2a-b0ea-69750a3f8f9d · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.858412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:a90a9c680d111946e607b8e4d9ae4bb8b26725a65c838106a3b50a481f9d1621

Observation 9fcef661-6cee-4053-86cb-95c8fc0cf59c · outbound

This paper cites Perceiver IO: A General Architecture for Structured Inputs & Outputs.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Perceiver IO: A General Architecture for Structured Inputs & Outputs

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.861655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:40a621e558dea9addeb6ba72675fe5ee98a15c30a61f141d237d214d6487b0cd

Observation 07cd58b7-0075-445c-a246-60115abcf57a · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Scaling up visual and vision-language representation learning with noisy text supervision

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.298227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:4b8484d42a1ea66bb6f680057472517e733bcca7376ecd85b9f6f8da3f89d424

Observation da2c2b08-b92b-44cb-a2d6-8f604e226792 · outbound

This paper cites Lift3d policy: Lifting 2d foundation models for robust 3d robotic ma- nipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Lift3d policy: Lifting 2d foundation models for robust 3d robotic ma- nipulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.311327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:3f0810136d6469f345f0b201c6142e9d6a84e26d5da2dde9a3f96289de567d9d

Observation 83154ea4-763d-4502-9e44-65bb88ef7b7c · outbound

This paper cites Kingma and Jimmy Ba.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Kingma and Jimmy Ba

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.315093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:f57c1b010fbae8a59ce2fa5ac18eda0cb752a7faa51f3fa99765f242c7d13bd0

Observation 537ba655-8893-47de-9a76-ce547efe1bb8 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Adam: A Method for Stochastic Optimization

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.848675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:16086e3dfe5cd206af4887a400c9a25649aef784bba40e1503c137ab9ab79be1

Observation 2ceeb50a-5650-4e02-bdaf-0c639d236712 · outbound

This paper cites Set transformer: A framework for attention-based permutation-invariant neural networks.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Set transformer: A framework for attention-based permutation-invariant neural networks

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.300542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:e714168e74995b932c3d54e7dba45f2f3936c6e998fdd53566946dff527db892

Observation ed489ff4-d5f6-46a5-acc4-4cbd2898241f · outbound

This paper cites Class: Contrastive learning via action se- quence supervision for robot manipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Class: Contrastive learning via action se- quence supervision for robot manipulation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.321125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:82e82dbf616060372d2bb516507bfa2a959b8ba028ce0c9b6d637482e6b9687f

Observation f829b1b5-d3d0-47a4-9e9e-68a15d8aca5a · outbound

This paper cites Rethinking Network Design and Local Geometry in Point Cloud: A Simple Residual MLP Framework.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Rethinking Network Design and Local Geometry in Point Cloud: A Simple Residual MLP Framework

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.845305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:4f7d9228360222c2355aee39497611fee3ac5aa79820b94a4146eef6c5562173

Observation 51460893-ea02-45a6-9416-e83d77d2620e · outbound

This paper cites A multi-view projection-based object-aware graph network for dense captioning of point clouds.Computers & Graphics, 126:104156.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining A multi-view projection-based object-aware graph network for dense captioning of point clouds.Computers & Graphics, 126:104156

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.252049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:0de2cb8d4008775f235cd59adecbfd5d32b8222847a04792dffa4db0d9a98a83

Observation 047c7a70-f8ac-4ff1-bda9-897b930213fa · outbound

This paper cites TIPS: Text-Image Pretraining with Spatial awareness.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining TIPS: Text-Image Pretraining with Spatial awareness

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.852460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:70682b849ad214b80d9f514ed0e1ff63a4573c070ef552b1fa873b6279ed33c4

Observation 68c7a14d-b60a-468b-8be3-91b8be394917 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.855310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:5ab559e2441a4a3e5947d1789018c2075767f666373ff18fd8abda4c52078478

Observation 430abd26-323b-43d4-a655-5e08b6cc22a8 · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.259955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:ba9b3299646f9f6f413f79a14f41101de2eac39485ba85e68c4e4b169be604f6

Observation 839f2c63-a781-42e8-a31a-6ec7dfa3db20 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.353200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:4521d7c7c74b81c120be7bd08982f04c682a06dd758b2c9e850c2bbc53954b02

Observation afe5c7eb-34a4-47ce-9076-9eab5265b586 · outbound

This paper cites 3d-mvp: 3d multi- view pretraining for manipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining 3d-mvp: 3d multi- view pretraining for manipulation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.320043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:8bc3f55adb5ad5f2f194df21bdacd91c41391794bf7d2eb877db96b4f3d31347

Observation 9caa72a4-47e2-49ab-8433-7ae7b765a9d4 · outbound

This paper cites Learning transferable visual models from natural lan- guage supervision.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Learning transferable visual models from natural lan- guage supervision

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.271836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:8575b34955b09ee6e98d605c747c69df4851de328a0406c831460519d74929b8

Observation bede3e90-fb50-4272-89b2-b80325a74a0e · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.293364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:a9d2c743a240c6df1d18998727b91f8cc3abdbeeeb9b69351969ec46a67c3084

Observation ee44d256-fdd5-4292-9a30-5b5d2139851e · outbound

This paper cites Learning the RoPEs: Better 2D and 3D Position Encodings with STRING.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Learning the RoPEs: Better 2D and 3D Position Encodings with STRING

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.885411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:781eecda1819ef10ac1ef23ff69217b064395e6366935ced072aeedd2f84e651

Observation 1d2d14c6-82c8-4837-880e-83ff708feca9 · outbound

This paper cites Self-supervised visual descriptor learning for dense cor- respondence.IEEE Robotics and Automation Letters, 2 (2):420–427.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Self-supervised visual descriptor learning for dense cor- respondence.IEEE Robotics and Automation Letters, 2 (2):420–427

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.269530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:7d4d4742424b067b279a815013828d8e769a3784e130c915926414b6e2095c86

Observation 6c8f155b-96a9-4714-8ada-96048da4f516 · outbound

This paper cites Perceiver-actor: A multi-task transformer for robotic ma- nipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Perceiver-actor: A multi-task transformer for robotic ma- nipulation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.280195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:9fa53216548e9d1c55a59244d06838048cd7da7ff917ecebf8c79cab42ef12cb

Observation f17a7717-22af-45d5-9ee7-f604b22c5562 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.880704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:5b10ae2d1f9bf13f69aefb5b0f72e943b69d8194cd06cc9b0b8c843fb346feaf

Observation a3d0d134-ca94-4058-a15c-0203fe0f537c · outbound

This paper cites Gemini Robotics 1.5: Pushing the Frontier of Generalist Robots with Advanced Embodied Reasoning, Thinking, and Motion Transfer.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Gemini Robotics 1.5: Pushing the Frontier of Generalist Robots with Advanced Embodied Reasoning, Thinking, and Motion Transfer

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.888636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:90e841666c9a95f603fe44823110020c7f0c545b6d6f2d40d33906ccf0124c64

Observation 91d45c5e-cf4a-40c6-ac8f-ad55ccccae7f · outbound

This paper cites Mujoco: A physics engine for model-based control.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Mujoco: A physics engine for model-based control

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.316971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:b274b3859852b1342a4a5f9f7df2ec65d2be13ceb8f99119f7d6d61a4bfeecb0

Observation bdc581c5-b340-44df-adf1-e1fdf4af5568 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Attention is all you need.Advances in neural information processing systems, 30

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.329087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:5959e059c39beda91159388e94417b67c8590665140b3f8d807695e515db0b73

Observation 2b6369f7-64d1-4cb9-acf8-81daf941fa6e · outbound

This paper cites Dynamic graph cnn for learning on point clouds.ACM Transac- tions on Graphics (tog), 38(5):1–12.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Dynamic graph cnn for learning on point clouds.ACM Transac- tions on Graphics (tog), 38(5):1–12

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.300344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:7a818d1923322771ac4346caec651ac47e607af598f839406a635eb3f971d690

Observation 4efc20a3-2ae3-405a-8504-20c163574875 · outbound

This paper cites ArticuBot: Learning Universal Articulated Object Manipulation Policy via Large Scale Simulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining ArticuBot: Learning Universal Articulated Object Manipulation Policy via Large Scale Simulation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.868328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:0e1857d741d0813c64fc13d34d87f6db2cec5e37643e5e1354d3909c40accbd1

Observation 9564fece-3e91-47a7-9ed4-d50a1a55a16b · outbound

This paper cites Point transformer v2: Grouped vector attention and partition-based pooling.Advances in Neu- ral Information Processing Systems, 35:33330–33342.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Point transformer v2: Grouped vector attention and partition-based pooling.Advances in Neu- ral Information Processing Systems, 35:33330–33342

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.272027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:095454d6bfe6a94c53941f68531bc10bec9654d399ce1463705ede10fb8e32da

Observation dadcd612-037c-46b7-b905-6e183ca0706f · outbound

This paper cites Point transformer v3: Simpler faster stronger.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Point transformer v3: Simpler faster stronger

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.335192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:76f8a4b61c5f39ecc3aea0ae8fd528fd7a1868433de7bd5ec231f27643c86c8f

Observation 5b1a45c1-2f41-4ae9-a02e-5edb7aadeb9d · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pointllm: Empowering large language models to understand point clouds

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.319028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:87ea6bb24e66a54db6bb48d0367076ebc09c2c7f870bd981f421117989bb2a67

Observation 1b797f68-fc98-4ca9-9f56-2c065b47e271 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.341300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:eaa740e161f0dd1f500dc6bedd79d78c28c17e4d100061a3260da2501e902bef

Observation d6223955-d4bb-40a1-a3bd-3df6b9f2d4af · outbound

This paper cites Ulip-2: Towards scalable multimodal pre-training for 3d under- standing.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Ulip-2: Towards scalable multimodal pre-training for 3d under- standing

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.311923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:794702934f1c18b4a67f44a51ecddd195453c51f44a7181fa3eeb1999d0c9110

Observation 0412ddcd-303b-4be0-b8ad-99c62cd14fd1 · outbound

This paper cites mt5: A massively multilingual pre-trained text-to- text transformer.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining mt5: A massively multilingual pre-trained text-to- text transformer

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.288796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:30f51282ee1bcd2c157cecbef1f21618d3f0c69a262455201fcaaf41edc80d2f

Observation 81c4f21a-03e5-4da7-ac7e-0cbcfa3a509a · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:27:36.874628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:b4532bc741ce3c96d6e618b2c90a416a6b55cbfcd43a3432f792143c95b50173

Observation 83e0764d-fcd2-4e59-8198-c9004deeb5c9 · outbound

This paper cites Point-bert: Pre-training 3d point cloud transformers with masked point modeling.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Point-bert: Pre-training 3d point cloud transformers with masked point modeling

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.286470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:4be187e19893aaf4a088e66ef32a4b78019b7e624dc0c0cf37bb82f7a0bd317b

Observation b6b04895-2c6a-4c50-93b5-db3ac162f866 · outbound

This paper cites Florence: A New Foundation Model for Computer Vision.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Florence: A New Foundation Model for Computer Vision

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:38:09.598269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:a0ea1d4e579e1870e38bf18ac207ed1b5b0a1484b0d0b40c2d7fabdf91945eb8

Observation 1055e2f1-16d3-4d74-bf8f-f55a27ad9ae4 · outbound

This paper cites Zakka, Y.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Zakka, Y

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.262459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:c433e940ffd8c5d3ba2f8e5e12a966943f6e3121a3a19e9ca2b20636fa0ed9f7

Observation c265138a-7034-4176-ac32-fdab7a3660c2 · outbound

This paper cites Gnfactor: Multi-task real robot learning with generalizable neural feature fields.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Gnfactor: Multi-task real robot learning with generalizable neural feature fields

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.297944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:852fc0f579e1bc0776287033df87c3b26e208dceed460274c1d4ad82d975f5d6

Observation de0ca59c-5209-4938-98fd-7dbb30b5ee02 · outbound

This paper cites 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.296116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:e19701ba28483650fe9d0eaecf0f1c30411440f4243a04a4794f1e9d0846ee7e

Observation d075399b-e71e-48b7-8bb5-9d9814272f01 · outbound

This paper cites Sigmoid loss for language image pre- training.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Sigmoid loss for language image pre- training

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.283053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:753635ee52d9ec167544f87db1b7df26c186942845decddccfaa1a1cb50936da

Observation aaaab6e4-4841-407a-a800-5e8b18750999 · outbound

This paper cites Point- m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.Advances in neural information processing systems, 35:27061–27074.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Point- m2ae: multi-scale masked autoencoders for hierarchical point cloud pre-training.Advances in neural information processing systems, 35:27061–27074

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.309370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:041d0ab22010ad1400d0131cebac5fa2fc1b266805b2912c555ba5602044c2b6

Observation e7b8d734-f51b-4dad-868b-37813d2d12c3 · outbound

This paper cites Pointclip: Point cloud understanding by clip.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pointclip: Point cloud understanding by clip

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.348693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:d45bc954a368bb69d2bcf9b4cd7260509c12152164b3ede03042bebdcc349546

Observation 37679273-1e9a-4aed-9842-c65850b31bec · outbound

This paper cites Point transformer.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Point transformer

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.265785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:024301fc6fc07d86c0c1215247cdf15bc6d4393bc98e331b120ec7f31f3f52df

Observation 240dbbd1-21ee-443c-a10b-8f49fcc8796f · outbound

This paper cites Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.337356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:16a71ba2ed73a56f9c2dcf887aa84d91670f4c5a25fa08916905e7bfd230badc

Observation 1306913b-d59c-460d-98df-8b5cf4445037 · outbound

This paper cites ALOHA Unleashed: A Simple Recipe for Robot Dexterity.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining ALOHA Unleashed: A Simple Recipe for Robot Dexterity

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.894779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:2461d388f4f5f557d6a804aa8d1fe3ef190bfe657dccb6d147f545d37692dfa8

Observation 2c0ae73b-0ed9-48cf-9a57-8d00a38fbc54 · outbound

This paper cites Uni3D: Exploring Unified 3D Representation at Scale.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Uni3D: Exploring Unified 3D Representation at Scale

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.901002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:7b716b391c0327bb12f7474744cf3ef957d930adf288692b6fc10892b609bb4d

Observation 2236030c-0b3a-4e13-a079-1b2f3693d493 · outbound

This paper cites HACMan: Learning Hybrid Actor-Critic Maps for 6D Non-Prehensile Manipulation.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining HACMan: Learning Hybrid Actor-Critic Maps for 6D Non-Prehensile Manipulation

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:27:36.871516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:db17d0131bde90652dabc1580baee02b8f446ead022e53acd303446237975737

Observation 75a99ab3-edf2-4b00-a637-e63254ad936d · outbound

This paper cites Pointclip v2: Prompting clip and gpt for pow- erful 3d open-world learning.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Pointclip v2: Prompting clip and gpt for pow- erful 3d open-world learning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T08:30:46.326983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:7872d8b2dc7b232754d70820abd18cb9d3b8df84114d49e2623d8cf525a64a05

Observation 9e492f3f-2473-4608-bb71-4fef5b056a1d · outbound

This paper cites If the point cloud contains more than 0.3M points, we randomly subsample to 0.3M; otherwise, we pad with points at the origin (0,0,0) to reach 0.3M points.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining If the point cloud contains more than 0.3M points, we randomly subsample to 0.3M; otherwise, we pad with points at the origin (0,0,0) to reach 0.3M points

Reference 71

Resolution
malformed identifier
raw_fallback, observed 2026-05-16T08:30:46.277798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:e6bbefadc6bc80ca1f5910dad2d66e91256159b5204aadc4e21d66350f99acaf

Observation 63f85e20-4e0d-4abf-8ee1-8b102953c8b9 · outbound

This paper cites an unresolved cited work.

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-05-16T08:30:46.333204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T08:24:44.943709Z digest=sha256:33b17970395d0da19f176a65e130b352d288ef14e1b29a8e0168a83ea37b6511

Pith citing papers

Observation 0c5ffa09-569f-4857-9c50-80a116b0313c · inbound

Contrastive Action-Image Pre-training for Visuomotor Control cites this paper.

Contrastive Action-Image Pre-training for Visuomotor Control CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-07-03T18:08:46.859071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T03:16:33.702802Z digest=sha256:77f823d4e8d067c6b497edd2ae3f29a0be9acaceeddc0d247e6a9cab6919b4ad