Pith. sign in

Paper Citation Record · LEDGER

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding

As of 7 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 1 inbound Pith citation observation for arXiv:2507.20110.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20110 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:48:06.435172Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-31T15:16:35.591649Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b835252d-3b07-40c9-a903-615b1f03eb5e · outbound

This paper cites GPT-4 Technical Report.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:00.428975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:00.428975Z digest=sha256:7c1313769a1921eed6c171a61cd42e61c4beb26bc72311fb24af14b7aebb7343

Observation f86a609e-f11a-4be5-a3ef-d675240390ea · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:14.765120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:00.496994Z digest=sha256:f6cee13c037927ac44f85ca021c343bf0d80bf58073c47d3cb9bf8a851b2b293

Observation e36935a6-65fa-4b73-8f8e-fc504753a67f · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Flamingo: a visual language model for few-shot learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:00.578854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:00.578854Z digest=sha256:849d9c1375523ecff867d2d62880c949e18ed963a0e47500087a785595c9db4a

Observation e5e5bd75-cc94-48f5-9124-46710c5f20eb · outbound

This paper cites ShapeNeRF–Text: A Dataset of Neural Object Renderings with Rich Language Descriptions.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding ShapeNeRF–Text: A Dataset of Neural Object Renderings with Rich Language Descriptions

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:14.614753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:00.679459Z digest=sha256:bd1d5ebd21c6a7568ec1c0b93a88326f6f298d2dd490479bbd86697b76b4d888

Observation 859c74ae-622a-44d6-ac40-5754879b1047 · outbound

This paper cites Llana: Large language and nerf assistant.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Llana: Large language and nerf assistant

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:14.356360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:00.771850Z digest=sha256:bbd48e3e6daee75667ef881a1c3d0041c185862860667dbe46297eb738c1baf7

Observation 49a81f53-61a3-430e-ad8c-099f512b945e · outbound

This paper cites Scanqa: 3d question answering for spatial scene understanding.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Scanqa: 3d question answering for spatial scene understanding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:14.077094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:00.863694Z digest=sha256:2a7610547e31af11ceaa36a95640b186b42beaf2df00f680343ffacd852a9e5b

Observation 48de0090-5227-4d2e-bdf0-c66f75c7874f · outbound

This paper cites CompoNeRF: Text-guided Multi-object Compositional NeRF with Editable 3D Scene Layout.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding CompoNeRF: Text-guided Multi-object Compositional NeRF with Editable 3D Scene Layout

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:00.939274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:00.939274Z digest=sha256:47cc2d0a99ca5d22e7e8fa4bd3fbc045eb861087d740e12fda729675b5356b80

Observation a5a0aff7-ee4b-4aac-8dd4-51cc63de8dbc · outbound

This paper cites Connecting nerfs images and text.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Connecting nerfs images and text

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:13.890929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.003325Z digest=sha256:53cbc9f48f63b929fe34708b3cf53b1c7512a82a948459b7661316b38d58005a

Observation 55ebde29-211c-475b-8f50-c6dcd892dc9c · outbound

This paper cites Neural Processing of Tri-Plane Hybrid Neural Fields.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Neural Processing of Tri-Plane Hybrid Neural Fields

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:48:06.723178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.072728Z digest=sha256:0a5bed0bcb15252865a66a20790ec11e42195a23af9ce5283a048559388ad100

Observation e9faa312-f256-4234-ba8a-60dc2e4e8714 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding ShapeNet: An Information-Rich 3D Model Repository

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:01.163710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:01.163710Z digest=sha256:0b4e3b4f94dc44344190f0919c6e147738353dfa5e9aafaa8f19e5a7c4af93d7

Observation cd0eecff-d970-49c6-aabd-b21a89877cfd · outbound

This paper cites Tensorf: Tensorial radiance fields.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Tensorf: Tensorial radiance fields

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:13.652976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.279694Z digest=sha256:7a7aa199eb4e8215a278c0e89effc4d96b675f7f3bce71cee789cbca9991e9e2

Observation 00532071-50f2-4488-8960-5bd6eb9173f0 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:13.481865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.369319Z digest=sha256:0659102dd2d2d126148f9b4e96fc59fde29156886100aa37189082feb7fad530

Observation 73726ca9-4546-48a2-8657-49909d9a5b56 · outbound

This paper cites VideoLLM: Modeling Video Sequence with Large Language Models.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding VideoLLM: Modeling Video Sequence with Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:01.473477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:01.473477Z digest=sha256:7f00a0430f30082746bc75c3b6fc48ecbaa699fc3f4b10204bc96eeb241d48e2

Observation 727cc961-4c35-487d-a8bf-dac8ce28af14 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:01.574668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:01.574668Z digest=sha256:d8028610876de535e6a7bd502233b166d5a02dad2009dc2ada5e86dd55736eaf

Observation dd635596-703e-4a63-891b-c16b2191c493 · outbound

This paper cites Language conditioned spatial relation reasoning for 3d object grounding.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Language conditioned spatial relation reasoning for 3d object grounding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:13.212722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.692980Z digest=sha256:0df7b47a29c7076e86aa19fd2ada008a05b89c21e1cfda9af72863919a25618c

Observation a794db84-808c-4359-958e-c65bedd801ef · outbound

This paper cites MVT: Multi-view Vision Transformer for 3D Object Recognition.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding MVT: Multi-view Vision Transformer for 3D Object Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:01.770602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:01.770602Z digest=sha256:1f989c2d3864c3a33d0e2794aef66901f7407cb357335608d4029e2f1eb114e0

Observation 55d7e0a5-0106-4e12-834e-ce23aca332a4 · outbound

This paper cites End-to-end 3d dense captioning with vote2cap-detr.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding End-to-end 3d dense captioning with vote2cap-detr

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:12.970121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.848099Z digest=sha256:ac1cad7edc32cbf492e1961aa899ba8bc8fc2290da6c617ad1e452d0f1fb45e2

Observation 82fc555f-bfa3-451f-8a38-164ed370c518 · outbound

This paper cites V ote2cap-detr++: Decoupling localization and describing for end-to-end 3d dense captioning.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding V ote2cap-detr++: Decoupling localization and describing for end-to-end 3d dense captioning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:12.694743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.891419Z digest=sha256:d38b19de334a1b07d3de34e9bee303a04f26d322fcd83a6460a71cd8f11a810a

Observation e1446d83-51db-4fcf-a3b8-66c076d2442c · outbound

This paper cites Scan2cap: Context-aware dense captioning in rgb-d scans.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Scan2cap: Context-aware dense captioning in rgb-d scans

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:12.439190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:01.992505Z digest=sha256:d14db74d007df17e9b63e7a8078fd068f74855cbc9f5f3ad23a8e8faa2452248

Observation d6765e62-af05-47a1-9011-d774999b51dd · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.079180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.079180Z digest=sha256:3f98f371b2eed2026e76858f3b0354fb2c6d3b6915506dd4c0404274f4a9607a

Observation af419120-34fd-4192-a135-a411acd978fb · outbound

This paper cites Palm: Scaling language modeling with pathways.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Palm: Scaling language modeling with pathways

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.159298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.159298Z digest=sha256:0743f817f80a4d487f12b34dbc3b15e9e15eaac35b3f9a76898d4c8d74d1e78d

Observation 065b0130-dd3d-409c-a974-e36d59ad6895 · outbound

This paper cites 4d spatio-temporal convnets: Minkowski convolutional neural networks.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding 4d spatio-temporal convnets: Minkowski convolutional neural networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:12.204744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:02.221622Z digest=sha256:88f85feec1cfcaff08eb57efc264d0a45ed413bdc2e7e5f431dcec23fa99e915

Observation cd05eb91-d402-4192-a399-1bf84ed228f3 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.293828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.293828Z digest=sha256:eff52269690f7f250f0491475dfb03ea8c59a20f7f5af84ab6e7ccd6ed99875f

Observation 9926460a-97ca-4bfd-a2bb-22b378356e87 · outbound

This paper cites Deep Learning on Implicit Neural Representations of Shapes.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Deep Learning on Implicit Neural Representations of Shapes

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.358997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.358997Z digest=sha256:9a72564eb59b48a194e416f5feba97d4a58df7ee489297436d8096e44f94bee9

Observation e5fda67a-8459-4d87-b922-00f8072cd5d5 · outbound

This paper cites Plenoxels: Radiance fields without neural networks.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Plenoxels: Radiance fields without neural networks

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:12.001115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:02.442410Z digest=sha256:50362ca701153c300e7593f2cda8edf94e57af6267277b3594fe1a7445b43e19

Observation b9e72428-3e69-451d-aee5-8c228bc6e688 · outbound

This paper cites Imagebind: One embedding space to bind them all.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Imagebind: One embedding space to bind them all

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.507765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.507765Z digest=sha256:ededa1d46736488770ca4e022d7b5a06fe8ebf7d8acd2f9df537302068a44b7b

Observation feff426a-774c-4815-9685-3b213ec08e13 · outbound

This paper cites Submanifold Sparse Convolutional Networks.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Submanifold Sparse Convolutional Networks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.590072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.590072Z digest=sha256:eb1c2eab190baa3c4726de2aa2018d28047ad9f26fd1d3a6c2b0bd8d307a2ed5

Observation cb278c4f-bce7-469a-87bf-a312ab79e322 · outbound

This paper cites Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.658757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.658757Z digest=sha256:6bcf8d522aa8398d27e75397b26562d1d2cf13ba8f0376b31e2cf486c901e07b

Observation d2973317-16b8-4964-b73f-f263517d2860 · outbound

This paper cites ImageBind-LLM: Multi-modality Instruction Tuning.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding ImageBind-LLM: Multi-modality Instruction Tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.757425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.757425Z digest=sha256:97e785c89c3e9ed72911d9705f1462ad6b8c4432f6903e157689363c5d167a53

Observation ca87ad9f-727f-4812-bd8e-5e505887a124 · outbound

This paper cites 3d-llm: Injecting the 3d world into large language models.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding 3d-llm: Injecting the 3d world into large language models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:02.816058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:02.816058Z digest=sha256:38a252f6776e2330eab3b1db55bf2a07b835f3b0f3660c55772dba152220036b

Observation d914a6a7-9378-490b-9cf8-88bbbce0b9ff · outbound

This paper cites Unlocking textual and visual wisdom: Open-vocabulary 3d object detection enhanced by comprehensive guidance from text and image.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Unlocking textual and visual wisdom: Open-vocabulary 3d object detection enhanced by comprehensive guidance from text and image

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:11.700595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:02.907456Z digest=sha256:1d87c671614bb6a198fe46d848758228b6c24a9072c01b2dc90b312b28f0e870

Observation d71c5897-53fa-4b90-9a72-fa17546ffe10 · outbound

This paper cites Cg-nerf: Conditional generative neural radiance fields for 3d-aware image synthesis.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Cg-nerf: Conditional generative neural radiance fields for 3d-aware image synthesis

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:11.411758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:02.982517Z digest=sha256:5affdc2150191ce1521bc689522dab2bc4afdfae7e7d26101e1ee6d95a92e724

Observation 95bd0d92-258c-4530-b5ef-fa2ef0bf7a2a · outbound

This paper cites Lerf: Language embedded radiance fields.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Lerf: Language embedded radiance fields

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:03.089857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:03.089857Z digest=sha256:ad0ff84a7c3524f60cdfa1bf18ede4641e47a2098cb52801aa67073608cc24df

Observation dc9dc57e-ee69-4c7c-b0ed-9d020a56507b · outbound

This paper cites Parameter prediction for unseen deep architectures.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Parameter prediction for unseen deep architectures

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:11.074926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:03.146553Z digest=sha256:e5409c44724e646f3d2be11f75047bd52c5da4907229e3e56deaca685335a62c

Observation dd5eacd8-8f49-4559-ad18-2ec9140bc9e4 · outbound

This paper cites Decomposing nerf for editing via feature field distillation.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Decomposing nerf for editing via feature field distillation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:10.848571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:03.232106Z digest=sha256:f70a54ae7de68b5707175c8efb955262e7bd48ae862414ce5efd6aefe6e8c991

Observation 4fa7979e-e56e-433e-b381-8c28eef6ed37 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:03.295549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:03.295549Z digest=sha256:ab6ba40add5a0d303cc9c30761a65105ce32932a18e93f08d07e6cc8b87f248b

Observation bc8b862d-b2c8-4184-be75-ca34c3a74c6d · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems , 36:34892–34916, 2023.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Visual instruction tuning.Advances in neural information processing systems , 36:34892–34916, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:03.448698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:03.448698Z digest=sha256:e4f63f3e403be4a981aaee76a0e53bb058be7dccddedd6b860ccf1055080cfa4

Observation 6fe358d8-c9ab-4143-9eb9-a6092229c0b9 · outbound

This paper cites SQA3D: Situated Question Answering in 3D Scenes.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding SQA3D: Situated Question Answering in 3D Scenes

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:03.529117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:03.529117Z digest=sha256:a05230d96870548a34f7337031c03036844f29086f4a702f671b0ae54d7ca73a

Observation ed082235-5ada-49a8-b6ce-4448fc336630 · outbound

This paper cites Latent-nerf for shape-guided generation of 3d shapes and textures.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Latent-nerf for shape-guided generation of 3d shapes and textures

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:10.648473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:03.611561Z digest=sha256:f3cc99fbd6b5b23430d407aadcf36276c5b70c4607253fb089fae31b630944c0

Observation 0ab68ad0-8ce8-4a2e-8744-6d0a94f2a2de · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Nerf: Representing scenes as neural radiance fields for view synthesis

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:03.693150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:03.693150Z digest=sha256:541c7c05e63b79cacf2836e3b5172ffb25c4ebb074b0adea43e0eebed6e14b05

Observation 03f8bf14-fc30-4bc0-ae0a-4347db1d2674 · outbound

This paper cites Reference-guided controllable inpainting of neural radiance fields.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Reference-guided controllable inpainting of neural radiance fields

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:10.418059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:03.827743Z digest=sha256:93f3e53eb93f0bbba5c98b7881c10207b068a5897bb2b0066345bd8cdbee1c85

Observation 9c1c88e3-d0f9-4d2b-b417-f781fc681767 · outbound

This paper cites Instant neural graphics primitives with a multiresolution hash encoding.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Instant neural graphics primitives with a multiresolution hash encoding

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:10.270091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:03.944069Z digest=sha256:3787f83169170745e71bce688be4ac0440e6c0456cedd775e0471691902402b0

Observation 9bc5e0dc-e54a-4875-b7d5-06c7b83c3145 · outbound

This paper cites Equivariant architectures for learning in deep weight spaces.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Equivariant architectures for learning in deep weight spaces

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:09.985271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.043245Z digest=sha256:e356878b101a27c619c0c75d27362cb1ebc40504a9b0344b6c0d7748526c906f

Observation 15e8539a-c6ac-4045-ad6e-c6a8c7bd55f3 · outbound

This paper cites Training language models to follow instructions with human feedback.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Training language models to follow instructions with human feedback

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:04.154064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:04.154064Z digest=sha256:20d4b77ce0abb10ea809462565096891b6ef17aeb3f4d4fe0564749055682313

Observation 84b6e44b-455f-48fc-91fa-e73a312161cd · outbound

This paper cites Masked autoencoders for point cloud self-supervised learning.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Masked autoencoders for point cloud self-supervised learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:09.688603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.258152Z digest=sha256:e850019c1936101b65218798e1fd222e3266bab0fe169b6d234e496b386df452

Observation aaaaf3dd-d9ab-4154-806c-bc80019219c5 · outbound

This paper cites Clip-guided vision-language pre-training for question answering in 3d scenes.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Clip-guided vision-language pre-training for question answering in 3d scenes

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:09.476881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.375622Z digest=sha256:ece7ba89cec218befcb38a497789b067227bdc34c28e7cd56b37bc6f57452f6f

Observation c595a8dd-67a0-4e6b-b0e5-2e043fbe2b37 · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:09.286251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.457692Z digest=sha256:e288d965f07a585a5e7dddba41d2671fc28e7864d568da2c8cf8530bb05b1bf7

Observation 4ef90906-4887-4eca-b3ee-c40f8caca8e1 · outbound

This paper cites Pointnet++: Deep hierarchical feature learning on point sets in a metric space.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Pointnet++: Deep hierarchical feature learning on point sets in a metric space

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:04.572299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:04.572299Z digest=sha256:c871a2dec25a205d357c043c42578b78aeebd91137cb888c8d0fdc0f767fa067

Observation 7d5414bd-801c-4c3a-965a-656b6cab5d4f · outbound

This paper cites Gpt4point: A unified framework for point-language understanding and generation.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Gpt4point: A unified framework for point-language understanding and generation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:09.172978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.654947Z digest=sha256:b2f8932ab90237a68d96071e0ff59baa3e998d3c59449c2b6ca76d3df88bd178

Observation f0555abe-8de2-4fe9-bd2f-5259c254993f · outbound

This paper cites Learning transferable visual models from natural language supervision.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Learning transferable visual models from natural language supervision

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:04.733648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:04.733648Z digest=sha256:70f1270288caf9c2b02656416c8523cb19ee9075823435a8112b01af0601ed3e

Observation 7f1f3255-25a4-409f-b8db-ad2aada69b32 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:04.824771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:04.824771Z digest=sha256:7ce7b3ef78c42949f2a0cf6977fa254beddab7d174ee74b4361a38ea7e2a535c

Observation aaafabc3-ada5-4583-a0c4-3fa7d65fcc89 · outbound

This paper cites Deep learning on 3d neural fields.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Deep learning on 3d neural fields

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:09.021935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.878956Z digest=sha256:a6ed3bdcd98c49abd51a747eca23ecd23dfc52abab9a921bfc249caf3d0d2cd0

Observation f5cff35a-9bd2-4f52-b085-9c7540cc7b53 · outbound

This paper cites Self-supervised representation learning on neural network weights for model characteristic prediction.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Self-supervised representation learning on neural network weights for model characteristic prediction

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:08.884051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:04.976553Z digest=sha256:2cffa96904d4fb3ef60299e9fcd460d07d40074bddaa8e3263058488308a26fb

Observation 05ddafe5-0fd2-4c7e-a577-807a5d1a5299 · outbound

This paper cites Hyp-nerf: Learning improved nerf priors using a hypernetwork.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Hyp-nerf: Learning improved nerf priors using a hypernetwork

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:08.684422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.036929Z digest=sha256:9aabaa5ed01d05c10160fa2bcbdf8f9e06863129922ae5aa979f8f9e0ce1149d

Observation c7584855-5479-456e-bce9-44e78b81a67b · outbound

This paper cites DITTO-NeRF: Diffusion-based Iterative Text To Omni-directional 3D Model.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding DITTO-NeRF: Diffusion-based Iterative Text To Omni-directional 3D Model

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.108713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.108713Z digest=sha256:6910710f27932f5e78e3e6ee7351c66330b13630b904bf2b377b7cbce64f0ece

Observation 2da05fd7-dc3e-405a-9a4d-542bcf4e24e0 · outbound

This paper cites Direct voxel grid optimization: Super-fast convergence for radiance fields reconstruction.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Direct voxel grid optimization: Super-fast convergence for radiance fields reconstruction

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:08.521012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.219841Z digest=sha256:fda8c9e33d7f00de011e5f2f6e4990e35f611f2c135393efc05a3a0851a218e2

Observation 8d1602de-0b8a-4ce9-932f-42e1e4f763ba · outbound

This paper cites Nerfeditor: Differentiable style decomposition for 3d scene editing.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Nerfeditor: Differentiable style decomposition for 3d scene editing

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:08.354300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.281748Z digest=sha256:fa68a748fcf92ee8ceac8f7e952740f7a85bf3dea0f85ece5e079a72a8c68a8e

Observation c26e365b-d88e-4c53-a69c-b6714aff5e66 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding LLaMA: Open and Efficient Foundation Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.330695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.330695Z digest=sha256:02f63b82a3b0bc0ad05d328c444b6d10c30ac0f7dd069db1104c88c2038a6699

Observation 7ab09c0b-a41a-48e8-8a2b-cbebe2789cee · outbound

This paper cites Predicting Neural Network Accuracy from Weights.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Predicting Neural Network Accuracy from Weights

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.407676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.407676Z digest=sha256:91c9cc0ead7ca0ca488c66ab64be23cafcd71117192b3f86a3a26bc91343b467

Observation eb4e30a1-0393-473c-82ee-5b520211aa20 · outbound

This paper cites Nerf-art: Text-driven neural radiance fields stylization.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Nerf-art: Text-driven neural radiance fields stylization

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:08.173686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.468093Z digest=sha256:165e2e217c54bd4d5a87aea6efbd2c00a46b12e13f5409bce008f555a64261f0

Observation c31f8e96-e0b5-4c36-90b5-40368345e440 · outbound

This paper cites Sparsenerf: Distilling depth ranking for few-shot novel view synthesis.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Sparsenerf: Distilling depth ranking for few-shot novel view synthesis

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:08.035248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.572795Z digest=sha256:e6211a996c770bcf65ddb8018b1610c916937359aa60e6fed38be9f514189c05

Observation a28ae1b7-f897-400a-b58d-e9bba72b171f · outbound

This paper cites Is Attention All That NeRF Needs?.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Is Attention All That NeRF Needs?

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.618712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.618712Z digest=sha256:c1b08c6bb74c081db6491eb16cdac04790f4e91f5afa74f102d20405b30071c7

Observation e300df26-04ea-4ae6-bf17-4223e3601067 · outbound

This paper cites Dynamic graph cnn for learning on point clouds.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Dynamic graph cnn for learning on point clouds

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:07.858689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.692671Z digest=sha256:cd0daafcf2239572ea1d058593d96075abf464b4b5162c636e5077a4907f07d5

Observation e9f58a95-666a-48b2-a049-35d35c33e215 · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Finetuned Language Models Are Zero-Shot Learners

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.768385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.768385Z digest=sha256:a1097ab6f20470343adfd0bf90d505131df79c6614364a721076a59aa865aeef

Observation b478d53e-5480-4da5-990a-df037edf46ed · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Pointllm: Empowering large language models to understand point clouds

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.820539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.820539Z digest=sha256:bacc04809390960e6af0134f804d9ff80ad3ac2a6598efcf23d334a025eb22cd

Observation 77673273-a29c-48a8-8cc6-4ad89b536166 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:07.719762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:05.907856Z digest=sha256:c67af0aa53e499daace9ae50230b4962078a34b69e6963014961c957ab6bd74a

Observation cc5f76cf-c860-463d-83c8-058c219d105a · outbound

This paper cites Ulip-2: Towards scalable multimodal pre-training for 3d understanding.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Ulip-2: Towards scalable multimodal pre-training for 3d understanding

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:05.963623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:05.963623Z digest=sha256:30efaeb45574c543b37cb59b9fb181971444e96cd6b8faf9cbeb7d7879fffc2a

Observation 3033ed8b-c1f1-48f7-9586-27403ce950e1 · outbound

This paper cites 3d question answering.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding 3d question answering

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:07.607741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:06.031130Z digest=sha256:eea09d60a95481943c21cad16074f925233320c6cf996a781cd721981dc7a3d4

Observation 519ceebc-6da5-49eb-a4ce-2838b03b8532 · outbound

This paper cites Point-bert: Pre-training 3d point cloud transformers with masked point modeling.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Point-bert: Pre-training 3d point cloud transformers with masked point modeling

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:07.431484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:06.098632Z digest=sha256:e97fe0a75f370698e84348dd2ccf5fa7b78b8a804c5a8e36b04c31af7f6aa6c2

Observation 7dd928f8-0435-4871-bd4e-afac0d0200df · outbound

This paper cites X-trans2cap: Cross-modal knowledge transfer using transformer for 3d dense captioning.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding X-trans2cap: Cross-modal knowledge transfer using transformer for 3d dense captioning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:07.306267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:06.200728Z digest=sha256:c22cb42ede55721e53e7312409dfd4bdf2a28bc6e02f2e5d3ddaf40369321e5d

Observation 784c7428-fff6-409c-9438-38c5d308cb57 · outbound

This paper cites LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T13:48:06.283815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:48:06.283815Z digest=sha256:4f0febf0b7c753d65c72708cfe602d25d86f94e00e310f9f813ce83a876088df

Observation 9aba2e05-f1e7-4d15-a6cb-3153795430b5 · outbound

This paper cites Multi3drefer: Grounding text description to multiple 3d objects.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Multi3drefer: Grounding text description to multiple 3d objects

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:07.168973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:06.345597Z digest=sha256:1fff0fded38e1cd9acbc1cb3dbcb9b69dc79f46de34c63338358d8203da04643

Observation b311e841-d93d-4903-8c7a-c1aef28fea23 · outbound

This paper cites Toward explainable 3d grounded visual question answering: A new bench- mark and strong baseline.

NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding Toward explainable 3d grounded visual question answering: A new bench- mark and strong baseline

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:06.944632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:48:06.435172Z digest=sha256:454b358931cd9864f796cd8b57a9ba8cff43597757c5a128f666607fffb16076

Pith citing papers

Observation a33bd187-fc65-4803-86b4-289a43d4f062 · inbound

MAViE: A Multi-scale Adaptive Vision Encoder for Fine-grained Visual Perception and Efficient Multimodal Reasoning cites this paper.

MAViE: A Multi-scale Adaptive Vision Encoder for Fine-grained Visual Perception and Efficient Multimodal Reasoning NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-07-31T15:16:35.591649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T15:16:35.591649Z digest=sha256:25eba8818ad44e6cc29d090082fd955ef32070db65541079073dd7dc8c582a35