Pith. sign in

Paper Citation Record · LEDGER

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

As of 12 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2412.06292.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06292 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:53:22.619176Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:49:39.643676Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T14:49:39.758012Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2de93c0-e140-413a-9279-2a2bee636f38 · outbound

This paper cites Zero-shot 3d shape correspon- dence.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Zero-shot 3d shape correspon- dence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.379194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.379194Z digest=sha256:8b270b8f45dfcf98ebac9bf21cf1711735a025519a14590da11a5b3615d5e270

Observation 4dff6c20-b5f1-463f-8ac6-c950b1c21202 · outbound

This paper cites Satr: Zero-shot semantic segmentation of 3d shapes.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Satr: Zero-shot semantic segmentation of 3d shapes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.521990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.385098Z digest=sha256:c11aa40c0c4b36a9cf0beaecc434ebebbdf0701666170c4ee16796240f3da71d

Observation 582e0e54-aceb-4f25-9f97-0410510bf224 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Flamingo: a visual language model for few-shot learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.389449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.389449Z digest=sha256:717b05db8099585701c2da9827556e0d0590b01fae560dc361703e5111087c7f

Observation 00ab1aeb-bb33-4a83-98a1-05180b557ac2 · outbound

This paper cites Claude 3.5 sonnet model card addendum.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Claude 3.5 sonnet model card addendum

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.489572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.394568Z digest=sha256:1dceedbe7ea9bf5d50982b7afd3250b2351b2e345e3710a15e09909259da6f1f

Observation a087627a-9cb8-40bb-862a-825fbf1e4aed · outbound

This paper cites Ncp: Neural cor- respondence prior for effective unsupervised shape match- ing.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ncp: Neural cor- respondence prior for effective unsupervised shape match- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.468354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.400519Z digest=sha256:3ca644029bb9eaa3799e3f43d9e020264752ff3405795444bde05f77cdacda7a

Observation f7dd7139-ff39-43db-ac0d-2391eb2563cc · outbound

This paper cites Language Models are Few-Shot Learners.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Language Models are Few-Shot Learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.405338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.405338Z digest=sha256:b5c0ee07ca37fba02f21e5ed2ad026aecf3ec898e1fbe4882bb8a3c6deb2cfcb

Observation e4a6b8df-d928-4025-8a99-7c7a08556dc7 · outbound

This paper cites Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.410059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.410059Z digest=sha256:476ccdd8ccb9f713e0ac45f91f03c06c6fc7f9b640141fec0c9197aadcaca91b

Observation df6fb406-9c59-4005-8af0-f8f2fb23d155 · outbound

This paper cites Unsuper- vised learning of intrinsic structural representation points.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsuper- vised learning of intrinsic structural representation points

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.448792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.414870Z digest=sha256:4452877f18e7626dad8754acedd52ef103d60b4041f80031dd676a5fa0ec2232

Observation 53c55fa8-823e-4c88-96d4-b78f4bd50586 · outbound

This paper cites Schelling points on 3d surface meshes.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Schelling points on 3d surface meshes

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.429886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.419429Z digest=sha256:24472d368f965aa8d4b177e7229f3c3f1b3ba796ead056c5b1f95eb67ca9f6b5

Observation 470d239a-0fdf-4916-a21b-88394179106c · outbound

This paper cites 3dmv: Joint 3d-multi- view prediction for 3d semantic scene segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3dmv: Joint 3d-multi- view prediction for 3d semantic scene segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.410607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.423666Z digest=sha256:4ad40205973c24ce3d375a501cb263562595b8bc6560d09ddd38f9a94e361db7

Observation f77e4a32-2e12-4a46-aae2-878e19db2235 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.427754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.427754Z digest=sha256:040bc7dc808726fde3b5f341a5fe970d8e760359c0b9ec6dfa8a71399562996a

Observation b39572e5-31b4-49c3-ab91-3ab8e891485f · outbound

This paper cites Unsupervised learning of category-specific symmetric 3d keypoints from point sets.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsupervised learning of category-specific symmetric 3d keypoints from point sets

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.382461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.432903Z digest=sha256:cc155f62bec01f1d997e15370310acc39ef9455019c4597a51ac59818cfe6a90

Observation a1009d42-3e04-49a0-88bd-73d605595d87 · outbound

This paper cites Mvtn: Multi-view transformation network for 3d shape recognition.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mvtn: Multi-view transformation network for 3d shape recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.365304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.437148Z digest=sha256:35851b3210b213861d7f391f2bee128580865531638ec1799a554fbca830de63

Observation 1871914a-392b-44be-989b-a3599c4d210b · outbound

This paper cites V oint cloud: Multi-view point cloud representation for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models V oint cloud: Multi-view point cloud representation for 3d understanding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.349447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.441286Z digest=sha256:95eb22e2292ae8fb2f07875179438de1b2a0ae0411aaf8a2437fe254b7adbf3a

Observation 84a880fa-5ed3-415a-bb82-e681933076af · outbound

This paper cites Mvtn: Learning multi-view transforma- tions for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mvtn: Learning multi-view transforma- tions for 3d understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.331028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.445167Z digest=sha256:d9ad46813587bb9124e5ac3c14cffd2265696dacc22155e7a479e1a41b1f30b6

Observation bc0da9e4-1860-4f29-9481-693403af7c7b · outbound

This paper cites Unsupervised keypoints from pretrained diffusion models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Unsupervised keypoints from pretrained diffusion models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.316667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.449699Z digest=sha256:c224682a5f384742504abe25281b724c165d7bc55b292edffc108e3718fb9ac7

Observation 596c8d96-1212-4984-93d2-282a39e49667 · outbound

This paper cites 3d-llm: In- jecting the 3d world into large language models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-llm: In- jecting the 3d world into large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.454016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.454016Z digest=sha256:5510664ac21385ada0ebb746a89cbfa3b1a7733e0b9376307c4e36cf415d9cc2

Observation bd089bf1-b03f-462c-8563-3d35bf8b5767 · outbound

This paper cites 3d-sis: 3d se- mantic instance segmentation of rgb-d scans.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-sis: 3d se- mantic instance segmentation of rgb-d scans

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.292681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.458692Z digest=sha256:b523895ab0b169731e829b127ea5dbdb05fc69ed1dcc5482540c7add48a143ea

Observation 454c4372-5c82-4eb0-82ff-789fa2034217 · outbound

This paper cites Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.278453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.462899Z digest=sha256:f403325a16414124d406b6ea7e5e6089e5d7fd632429c1ab62f72afb23a9a40c

Observation 7d10388c-56b0-4ca7-9d4b-a6f7e0b2404c · outbound

This paper cites Keypointdeformer: 9 Unsupervised 3d keypoint discovery for shape control.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Keypointdeformer: 9 Unsupervised 3d keypoint discovery for shape control

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.260658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.467348Z digest=sha256:6efd94c42389c1cb81ba1c98c6cde05a57ec3e9aaf807799a6fb35f77c1b9b89

Observation 9e92acb9-cf9f-4a0e-b476-61628feddb11 · outbound

This paper cites Multi-view pointnet for 3d scene understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Multi-view pointnet for 3d scene understanding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.245437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.472270Z digest=sha256:266924e4aa73026d343e8c0dae2303d7e57046dbcd033c23f78980019dc720bd

Observation d2165763-6c21-4ef5-a648-823767707ba2 · outbound

This paper cites 3d shape segmentation with projective convolutional networks.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d shape segmentation with projective convolutional networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.222504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.476533Z digest=sha256:6a7a0fb36358fbeadd92cb060d5b6facbd196be04f267d16382fd737eb1d78e5

Observation b4bb8479-64b2-47de-849b-dd1fa8933abe · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.480826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.480826Z digest=sha256:a8eea8d6ef6a3fdaa0e25bebafed73f1dacb4a6d3feee3ac5219ffd0ce164a69

Observation 0bdbfc53-5879-42ec-9077-6889205c388f · outbound

This paper cites Virtual multi-view fusion for 3d semantic segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Virtual multi-view fusion for 3d semantic segmentation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.188206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.485866Z digest=sha256:e780d055af9b8cefef96f4fd2151b2e273acd5e11820814116f4e28bb862b598

Observation e2c303d6-0c1c-48af-bc3b-6faf051f9de1 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.490024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.490024Z digest=sha256:3a95cdcc034d507cce7b44741e77a9894e0ba96e942a972ec959e4784e4f6216

Observation eac79e8c-9230-404e-9faa-24db22f008a2 · outbound

This paper cites Visual instruction tuning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Visual instruction tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.494848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.494848Z digest=sha256:8df5bf3b71b29daae2d30afdbe207931527d9cec4cc3245d33fbef01d2f85b00

Observation 72cfa958-02dc-4787-a8b9-329f22a9cae1 · outbound

This paper cites Improved baselines with visual instruction tuning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Improved baselines with visual instruction tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.146801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.499396Z digest=sha256:df3bba0af6ef79ab4c1caf6236c8dc63bfd66a43884622de38f034535d16b311

Observation fcd044aa-136b-48e1-b8d8-b2c514ab1d65 · outbound

This paper cites 3d-to-2d distillation for indoor scene parsing.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3d-to-2d distillation for indoor scene parsing

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.131318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.503575Z digest=sha256:5542bdd8d61b825056b06189fcb7ca540cc26560eeac984f9360a329a84124c1

Observation 71a2aa23-2dfa-40ad-a580-96362a4f41e8 · outbound

This paper cites Learning to segment 3d point clouds in 2d image space.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learning to segment 3d point clouds in 2d image space

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.116382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.508556Z digest=sha256:a4d77891ec2eb02541b123f1619ca5479914f42c3d298cf7305d99cbe586ac7a

Observation 378073f9-a7d3-461a-b7ef-d29ea10de51b · outbound

This paper cites Egoloc: Revisiting 3d object localiza- tion from egocentric videos with visual queries.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Egoloc: Revisiting 3d object localiza- tion from egocentric videos with visual queries

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.093838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.512221Z digest=sha256:12b7a7d03d5c7a54a751122d33ea02746552f745212e0d486b621c7c2aae0bfa

Observation cc9a7285-2371-41ab-9431-02124fdbfb9b · outbound

This paper cites Tracknerf: Bundle adjusting nerf from sparse and noisy views via feature tracks, 2024.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Tracknerf: Bundle adjusting nerf from sparse and noisy views via feature tracks, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.073993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.516273Z digest=sha256:36c625fd06ea08afcf6a14809f6b134730e7eadc897c95368af4d486cf27ebfd

Observation afb41c8c-3f9f-4502-8ed9-0bd41c70655a · outbound

This paper cites Gpt-4 technical report, 2023.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Gpt-4 technical report, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.048343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.520516Z digest=sha256:4a19c565fab0903bc638b4519034e86fb8427d2062a373cef8de88c2c2a4a0cc

Observation 455dfd8b-5d4b-4918-ab37-1a55510ad9ae · outbound

This paper cites Synthesize diagnose and optimize: Towards fine- grained vision-language understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Synthesize diagnose and optimize: Towards fine- grained vision-language understanding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.031427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.525276Z digest=sha256:55dae8e1d2db91b703cea3f9165789ad6d3c3a8ced41bc6f10f15f05dd575b19

Observation 46b9cf11-d76a-4e8f-9c9a-e24ae4926aa2 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Shapellm: Universal 3d object understanding for embodied interaction

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:23.014512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.529670Z digest=sha256:fc25171864e848bbd1076c779e23efc250e1060956b0c95d286491e019cd7946

Observation f1ac552a-836b-475c-b835-1e4d326928ce · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learn- ing transferable visual models from natural language super- vision

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.998627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.534012Z digest=sha256:0a52cee77088b33f226c2fad5caf0753df922357606cc4534bdf88cf7c06af13

Observation 39761a1e-ab19-4615-912c-6fc91ee58bac · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Vision language models are blind: Failing to translate detailed visual features into words

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.538799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.538799Z digest=sha256:67004d6533ac057e7b6dadc39ae8fcd5ab0ee2b26592ca18b64466dfba088cff

Observation 31eb0969-dfc1-488c-8d08-2a5b9fb900fe · outbound

This paper cites The Strategy of Conflict: with a new Preface by the Author.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models The Strategy of Conflict: with a new Preface by the Author

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.979666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.543453Z digest=sha256:c5c6a90de820110929574d99970d4c3b4b6d740ff57380513432d32ab7a70bcf

Observation 925db8d7-fa42-4053-8bc0-8ba92a4a68d8 · outbound

This paper cites Mask3D: Mask Trans- former for 3D Semantic Instance Segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Mask3D: Mask Trans- former for 3D Semantic Instance Segmentation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.548106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.548106Z digest=sha256:e7033d2eb86f42ea4872aa9bf4ec8b2dc210388e8922990e208cc727d0fd56e4

Observation 5b34724c-af20-4f5f-824e-be6c83aea78c · outbound

This paper cites Skele- ton merger: an unsupervised aligned keypoint detector.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Skele- ton merger: an unsupervised aligned keypoint detector

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.949277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.553206Z digest=sha256:0a3c39344f85c8d032b271cc04b9d075a06f32800f72a1ef4313c0c690ea6114

Observation ed0f51ef-4cab-4922-b4fa-84dc6f63c8d0 · outbound

This paper cites What does clip know about a red circle? vi- sual prompt engineering for vlms.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models What does clip know about a red circle? vi- sual prompt engineering for vlms

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.935249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.557535Z digest=sha256:71e64a5fa14dff6ca98bf63895d12b52af217b4771f7b526e54f4c105ba49754

Observation 294a82b1-c7bf-42fa-8116-ce496b9ff8a2 · outbound

This paper cites Discovery of latent 3d key- points via end-to-end geometric reasoning.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Discovery of latent 3d key- points via end-to-end geometric reasoning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.920888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.561859Z digest=sha256:a4f81f99df6ef59755d3cfee9b438eee1008d8948134a1703250fc1825ac2a1b

Observation 353481c6-fc2a-4a3d-854d-4c931b2e2dec · outbound

This paper cites Ldls: 3- d object segmentation through label diffusion from 2-d im- ages.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ldls: 3- d object segmentation through label diffusion from 2-d im- ages

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.905419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.566259Z digest=sha256:141ff9b80d45cc69562d6e312ba930c3c2f6f48722208ea14b3f693cf5d8f9a0

Observation 4dc85b4c-cbf5-42e4-bb6c-dfdffee05be7 · outbound

This paper cites Learning 3d keypoint descriptors for non-rigid shape matching.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Learning 3d keypoint descriptors for non-rigid shape matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.889417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.570113Z digest=sha256:aea97bc86e75df30f3d9f6cab7b005f77a5ee13dd9fc824ccee66e747defdbeb

Observation 6b129e28-9bf1-4a00-ae2f-7e21ddee4aa1 · outbound

This paper cites Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.871745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.574293Z digest=sha256:67014d4916397237dddbb5d59aace9ad4032b2c1d14a1094a97af6e1b0d9a0cd

Observation 88e02f07-2665-424c-a728-65d545fac1b2 · outbound

This paper cites Back to 3d: Few-shot 3d keypoint detection with back-projected 2d features.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Back to 3d: Few-shot 3d keypoint detection with back-projected 2d features

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.856075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.578146Z digest=sha256:4c35805d964b785ae7977c16580d657542039910e8fa7476b485baec8cccdbdc

Observation b8f9746f-4395-4082-abf0-1c4842dc3551 · outbound

This paper cites Clip-dinoiser: Teaching clip a few dino tricks for open- 10 vocabulary semantic segmentation.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Clip-dinoiser: Teaching clip a few dino tricks for open- 10 vocabulary semantic segmentation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.840684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.582597Z digest=sha256:bd3821b964bf16f9d2dbcb78f5e403202b88c97bb63a884778221a07e0585166

Observation d80eac84-56fd-48b9-9b57-fb748e5681a0 · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Pointllm: Empowering large language models to understand point clouds

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.824142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.586781Z digest=sha256:fdc67e5e4164bd4aacbe09b2f6521cb21d65b86315d4c7ba135d068d0039760e

Observation 203fd48a-fece-48b8-b6a6-746e9186fad3 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.809708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.591199Z digest=sha256:56ebedf4a69589966b664b82fae4f8db41b777c9a5229f323f2ef9e83dde5553

Observation aef7633b-abfa-4629-9c13-1e13f93132ff · outbound

This paper cites 3dfeat-net: Weakly su- pervised local 3d features for point cloud registration.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models 3dfeat-net: Weakly su- pervised local 3d features for point cloud registration

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.795136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.595690Z digest=sha256:d0e2125b62fb849f1082ab63fc42960dc987b617a1da9a3b54f4348fee0f3094

Observation 8b26c0c1-3579-445e-97dd-21a712c5e4c7 · outbound

This paper cites KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.600534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.600534Z digest=sha256:cf74cff6543ec03394c719d84735fafde54d3e7cca0b706ecfb6a1e139dde2d9

Observation 276695df-040c-467d-9e47-b5b8d2137c6d · outbound

This paper cites Ukpgan: A general self-supervised keypoint detector.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Ukpgan: A general self-supervised keypoint detector

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.780211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.605173Z digest=sha256:5a19f4cab8a62c4e1ac6b139aca2726b200ee8f874815c49e4442e561f31ffcc

Observation 6f564cf9-386f-447d-8be6-636e7c1d5f14 · outbound

This paper cites Good at captioning bad at counting: Benchmarking gpt-4v on earth observation data.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Good at captioning bad at counting: Benchmarking gpt-4v on earth observation data

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:53:22.765899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:53:22.610243Z digest=sha256:a174155d5183ec55482900ee747b2dec28d103b42bc5bf091c56fe1a3afdc389

Observation 902ee351-1ecc-4d3f-bc1e-01008d542817 · outbound

This paper cites Uni3D: Exploring Unified 3D Representation at Scale.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models Uni3D: Exploring Unified 3D Representation at Scale

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T19:53:22.614137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.614137Z digest=sha256:e2599cfc402f70ee3a59964811934e549601a3bb7191defc28710d70d276cd87

Observation 9b53a1ec-0163-47b0-b616-b3d452b8b755 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 54

Resolution
malformed identifier
no resolver link, observed 2026-08-11T19:53:22.619176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:53:22.619176Z digest=sha256:5a20d15efab7328f2221c0ecc68a40810bc8cf5c1fb3b33763969ebeaab215b1

Pith citing papers

Observation 7f9daa15-1422-4e9f-ba24-35f50659d139 · inbound

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation cites this paper.

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T14:49:39.762699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T14:49:39.643676Z digest=sha256:d5470b84d09a206fe4600fac0c4f81e430356adcb5cc1d47865ee36df8471e94