Pith. sign in

Paper Citation Record · LEDGER

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE

As of 13 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 1 inbound Pith citation observation for arXiv:2411.16856.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16856 v3

Coverage vector

measured 93 of 93 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:52:10.504301Z

measured 94 of 94 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T14:48:21.787919Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

93 of 93 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved36
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b2a4df05-a876-4940-982b-184584b824ac · outbound

This paper cites GPT-4 Technical Report.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.084365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.084365Z digest=sha256:6d7f378f759b16717d3fe28cd92f33ae9a1bf86d40ef8cb6dc57b633182dd849

Observation 8c4086a9-fbd9-4ef4-9b52-53b7bd4d1d21 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.089742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.089742Z digest=sha256:4d7856eaec3fc6e855b88e54950992908ea1afd7a7cc671e95e25db0298de458

Observation e5f5b9ae-b026-429d-b1cf-e477c957156e · outbound

This paper cites Demystifying mmd gans.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Demystifying mmd gans

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.094793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.094793Z digest=sha256:520e330514371d20a36df1c950d2b3f57896ebf410e3998d6647cdf3ea40e539

Observation df38ee96-2842-4930-a5d3-4feb28a9c736 · outbound

This paper cites Video generation models as world simulators.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Video generation models as world simulators

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.099617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.099617Z digest=sha256:d82744dcf4e4ff7409b15d02aa64f67145c01a60e18c6b623baadfbdf3817a15

Observation f884c0f6-02f5-406a-b35c-001fd00115e9 · outbound

This paper cites Lan- guage models are few-shot learners.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lan- guage models are few-shot learners

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.104973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.104973Z digest=sha256:1107aec1e59163a8eed63f5fe3fc10cf734255da9cadbd0f801ca96c8d85b68d

Observation 3b334e59-2828-4a35-b637-51989331ffd5 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE ShapeNet: An Information-Rich 3D Model Repository

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.109776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.109776Z digest=sha256:a56e62dd5133575e431ed10d9d66c6522df58527f997e4d62b0a5448c9ef2b1f

Observation ddeb8333-e54e-42ea-80c0-a74877a506bc · outbound

This paper cites Maskgit: Masked generative image transformer.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Maskgit: Masked generative image transformer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.114779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.114779Z digest=sha256:7c7f402c41f45881d697c74bce3d26f45f38698242c91433308fc29936542809

Observation 184058bf-2767-4e7a-a36e-aeec197f0482 · outbound

This paper cites Lara: Efficient large-baseline radiance fields.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lara: Efficient large-baseline radiance fields

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.119338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.119338Z digest=sha256:434dc57ce89f7b835feb2744e7e24fbaf443a01e4b93549fa2228c36c9b65bcb

Observation a34a7808-208e-441d-8c59-a5df9574e0ae · outbound

This paper cites Fan- tasia3d: Disentangling geometry and appearance for high- quality text-to-3d content creation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Fan- tasia3d: Disentangling geometry and appearance for high- quality text-to-3d content creation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.124518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.124518Z digest=sha256:a43f154e0d4175ed906c6e7b9bd2386497b582053d1d6df02d647a08ee303710

Observation 8eff5223-12e7-458b-bf06-b92f0982edee · outbound

This paper cites Comboverse: Compositional 3d assets creation using spatially-aware diffusion guidance.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Comboverse: Compositional 3d assets creation using spatially-aware diffusion guidance

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.129061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.129061Z digest=sha256:826d9fbeee65c560ad0b5cddda24c883924facf53d626c065ffcf9ce3d37259c

Observation 717632a5-fac5-472f-aa8f-877600fbbe8b · outbound

This paper cites Meshanything: Artist- created mesh generation with autoregressive transformers.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Meshanything: Artist- created mesh generation with autoregressive transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.134306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.134306Z digest=sha256:64478bcf2694add97adb49eead93682dd23cebe5aa2d5720e2e64fa1fe1e76f8

Observation 4a792050-dd0a-4f1f-865e-0224f030df8c · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.138746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.138746Z digest=sha256:50d1b1eb1f44cecff01efeb973f954eb5ce935463a81c52555dc9ab221a81b3a

Observation 83d55a0b-bab7-4463-8248-10cc1d91c3e9 · outbound

This paper cites Palm: Scaling language modeling with pathways.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Palm: Scaling language modeling with pathways

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.143385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.143385Z digest=sha256:e14515cfbbc63d43c8814e5a4a3a490104c3f875176dffa7c50c57e219070384

Observation cb27c8d0-7c9d-4a11-a828-2c3e028c21eb · outbound

This paper cites Objaverse-xl: A universe of 10m+ 3d objects.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Objaverse-xl: A universe of 10m+ 3d objects

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.147883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.147883Z digest=sha256:404c1d4730c82ff04fbd6d65253a8ac4dfccbbdc053ddf936c399cf7230e88de

Observation 3e8856e4-e1a5-4ecd-9e16-df802fdf20c3 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Objaverse: A universe of annotated 3d objects

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.152354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.152354Z digest=sha256:c21dbc980407a0541f3a0f4f3b1c3a407e83ca0614cdd21d1037ef812198d5fb

Observation 83534297-5288-4a3d-a4d6-5909c0e8c591 · outbound

This paper cites Palm- e: An embodied multimodal language model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Palm- e: An embodied multimodal language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.684263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.156786Z digest=sha256:cb2dde3d59e036bdc82f9f821d585651c5a3a369637c01d032c153ceded4f98e

Observation b154415c-54c7-44a0-a891-e0ca88c199d2 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Taming transformers for high-resolution image synthesis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.161472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.161472Z digest=sha256:22775423931da87838af8a3f0c82bc324547514e740d7105c2421229cca9f978

Observation d718883a-8478-4927-8f57-37d366037e8d · outbound

This paper cites Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.167237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.167237Z digest=sha256:34951f6c15612ca2750ef3911a047f8902b0de3c3a411e1808a5a83687ac1a7a

Observation 93153d77-a56d-4b25-8ba6-f69134e1b7a0 · outbound

This paper cites OpenLRM: Open-source large reconstruction models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE OpenLRM: Open-source large reconstruction models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.660835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.172277Z digest=sha256:e279fc9613d16738388f3d00564132a92ca7fcf95fbb0ffabbe3a11df71bba8d

Observation 3cd612bb-cc44-4dd7-a04e-c3c1fb227622 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.176860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.176860Z digest=sha256:353698d99ff9e6f1119400d9b4758fe98a462f53888eb444c839da03dafb6bc5

Observation 83bdb322-003d-410b-8fea-bd9373a9e823 · outbound

This paper cites Classifier-free diffusion guidance.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Classifier-free diffusion guidance

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.181260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.181260Z digest=sha256:234dfd4b86925d0f553e48d64951cd57799ab3e583453f0b189361373d2e0704

Observation 36b8f98f-24b5-4000-9d08-e51e92ef46ab · outbound

This paper cites Denoising diffu- sion probabilistic models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Denoising diffu- sion probabilistic models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.185923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.185923Z digest=sha256:f6d8565c7e4619d36b8e14a4e521bada0051fa315cca0cb5d83f3a621b84f5b0

Observation 79c6ec6d-3616-425a-bbad-a410c7927808 · outbound

This paper cites 3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.190409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.190409Z digest=sha256:842e4dd1834be05d7712d6936604a872b47fb80327a9773668001aab8702886b

Observation 21ead568-347e-494b-a0ac-b0e44806e213 · outbound

This paper cites 3d-llm: In- jecting the 3d world into large language models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 3d-llm: In- jecting the 3d world into large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.618201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.195390Z digest=sha256:a2ee37216ecaa7a982d48e9a06d4cb9aca5268fe0fe3da55dd3b43ef30d14f5e

Observation 5bff3f0c-8f8f-47fe-ba4d-657e9c54ad42 · outbound

This paper cites Lrm: Large reconstruction model for single image to 3d.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lrm: Large reconstruction model for single image to 3d

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.603468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.199835Z digest=sha256:b9d56980b1727420580fc1c9d8f4922517476b70dc823d662798ed3a8a0be395

Observation a986443c-a237-4e31-aec1-0414030b3afa · outbound

This paper cites 2d gaussian splatting for geometrically ac- curate radiance fields.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 2d gaussian splatting for geometrically ac- curate radiance fields

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.588870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.204127Z digest=sha256:85f4381e9f0bc58d4bb9755d2ddbb67918f983b3096fadf6c430a0e8d60f8742

Observation 1de84dd7-41c9-4a7f-95ef-0f1efe65096c · outbound

This paper cites Shap-E: Generating Conditional 3D Implicit Functions.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Shap-E: Generating Conditional 3D Implicit Functions

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.208683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.208683Z digest=sha256:f5002634e6909fcfa5e24609e901c44c804a1a11c2ff0420ffcbc9300589f821

Observation 6bf5d517-9a33-467a-9189-07787be4c080 · outbound

This paper cites Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.573772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.213565Z digest=sha256:ed5724351142dcb761cc0e0544932921414d2f5b9fbe25bce4270375b90a4a6e

Observation d82dd398-3e3a-4244-bfc4-c2bb78827445 · outbound

This paper cites Musiq: Multi-scale image quality transformer.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Musiq: Multi-scale image quality transformer

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.558286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.218183Z digest=sha256:42b538a7b378df461531b9de05dd300db2c94fe47503f71ff540e5d3a13283b1

Observation 84d6e0a7-162f-4594-893a-f1a0a03f4782 · outbound

This paper cites Kingma and Max Welling.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Kingma and Max Welling

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.542881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.222597Z digest=sha256:419808a3c3cf182dcc9cd3a4956cbc9fc51ced8a49a6a8ca41f086b9af582638

Observation 181b492b-c64a-4e69-90a0-9d0d4ce2a62c · outbound

This paper cites Nerf-vae: A geometry aware 3d scene generative model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Nerf-vae: A geometry aware 3d scene generative model

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.527816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.226855Z digest=sha256:845d344bf32aa0e044eb1075243c081192c18b8071621a8968327d0e3af260cf

Observation 22b3f0fa-a7ca-48ef-a3fc-bccd6703cdf7 · outbound

This paper cites Ln3diff: Scalable latent neural fields diffusion for speedy 3d generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Ln3diff: Scalable latent neural fields diffusion for speedy 3d generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.512855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.231057Z digest=sha256:0353f5f0244988926fa04e8eacc82fb1928af51fedf996b8c7d61d8e6bffecd2

Observation fbffe989-7344-4611-8b0a-7cd8a3bf9d2f · outbound

This paper cites Gaussiananything: Interactive point cloud latent dif- fusion for 3d generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Gaussiananything: Interactive point cloud latent dif- fusion for 3d generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.497786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.235292Z digest=sha256:acd411455aeaf514f992a87a6171b462668ef61d92eb32d40c76e624128b3a4a

Observation 27814b0d-7cac-45d9-9e86-c13f9bb2ff0b · outbound

This paper cites Autoregressive image generation using residual quantization.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Autoregressive image generation using residual quantization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.240211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.240211Z digest=sha256:c3b4552722f183bc515b454a1a78db9f4c43eb18eead8e1c1ee1a19e36440cd8

Observation 04dc7fde-f63d-459e-92e9-b11b3ef14017 · outbound

This paper cites CraftsMan: High-fidelity mesh generation with 3D native generation and interactive geometry refiner.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE CraftsMan: High-fidelity mesh generation with 3D native generation and interactive geometry refiner

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.473350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.245097Z digest=sha256:51794f5277f4eb1f09f88d9c94745fe7ef0adce4a9ee7e9fbad930a490ea119b

Observation 4e718b6e-cb74-4811-b9ef-9792feeede6a · outbound

This paper cites Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.249652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.249652Z digest=sha256:419a7e7fb6baa6e6fc077caa90aa82287082d10088c6afd14d2dda4cbc1a285f

Observation 5a3db5ba-30c3-46c4-9604-9d0458f99b00 · outbound

This paper cites Visual instruction tuning.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Visual instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.458093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.254399Z digest=sha256:4cbd807668b278277f0bbc34e853482550195d83cf475aa0729e305e3c19e0af

Observation e54ce1db-f26d-4554-b3ab-0f3c6fd6a12a · outbound

This paper cites One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimiza- tion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimiza- tion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.443893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.258753Z digest=sha256:766388cc206f0a1916e40a8624d1bb7563d339fa1fe1e27fe8ceac25955de168

Observation cf297b72-e8f7-40f3-8983-f1a1067f77c1 · outbound

This paper cites Zero-1-to- 3: Zero-shot one image to 3d object.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Zero-1-to- 3: Zero-shot one image to 3d object

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.429464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.263015Z digest=sha256:f21fa58f2248119873122caa87af26e8e7e72187f7d1af380efa9425be388848

Observation d50b8a05-abde-4f19-aa22-9f6668fad508 · outbound

This paper cites Wonder3d: Sin- gle image to 3d using cross-domain diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Wonder3d: Sin- gle image to 3d using cross-domain diffusion

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.414849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.267215Z digest=sha256:16335de0b078fb9a498c170d6d6df16a0b7748afd9b8209a835d1dddce5f3668

Observation 08d8f544-0920-4df0-8b45-229b557c645f · outbound

This paper cites Scalable 3d captioning with pretrained models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Scalable 3d captioning with pretrained models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.400539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.271394Z digest=sha256:18196569f31043d73ae6910ddd4513c2421574b3821f2f8f3ba932f2fd046425

Observation abf3f6b4-2824-4c8f-b435-f71e999c5eb3 · outbound

This paper cites View selec- tion for 3d captioning via diffusion ranking.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE View selec- tion for 3d captioning via diffusion ranking

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.386415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.275802Z digest=sha256:0376b036479a02840a18e1dfe9aebfcb0444395bd264dff4f7f43081c8679b68

Observation 227ab008-0fa0-484c-aa81-4b0e5b2a84ea · outbound

This paper cites KOSMOS-2.5: A Multimodal Literate Model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE KOSMOS-2.5: A Multimodal Literate Model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.280261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.280261Z digest=sha256:428e1bb6db46201ee2cbe3538392366d6f49a3ddc3dce585ec548745401975b0

Observation 61a817c7-f2a2-4815-9f90-d634cbd7fff4 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.371963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.284950Z digest=sha256:f54c0a9413f75df4dceba940b5a8bc8733d29fdf28f5b96531b98bf856c4fdc1

Observation 08c8c0b3-89f7-4349-8453-4b8c86293b06 · outbound

This paper cites Autosdf: Shape priors for 3d comple- tion, reconstruction and generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Autosdf: Shape priors for 3d comple- tion, reconstruction and generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.357222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.289490Z digest=sha256:7c4567ac369a5b68b9db8a12832a89529f04552e556cd706d65e9c8367eebea1

Observation 3cc9fa97-dc11-4ef5-af1a-14817b4e84bd · outbound

This paper cites Point-E: A System for Generating 3D Point Clouds from Complex Prompts.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Point-E: A System for Generating 3D Point Clouds from Complex Prompts

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.293892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.293892Z digest=sha256:4e80e839cc45f7ef56016ec57b93c27ce118d8b98f2d846b3303859227e06c36

Observation 3372f6f5-95f2-4b1e-a396-fda875132564 · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Dinov2: Learning robust visual features without supervision

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.342357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.298697Z digest=sha256:017656626f4134fdfd20d9563bf1eb40da50ebbb90e754b6672f1e88a456a1ba

Observation c7ec7313-fb60-4e9f-986c-27154643d9d5 · outbound

This paper cites Training language models to follow instructions with human feedback.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Training language models to follow instructions with human feedback

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.327504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.303249Z digest=sha256:fe49b25f117c2124df78fd5e56639602b4de14c4b18668d8abb170250a49e3da

Observation 94b646af-38b1-4504-adf1-738999ce0ad1 · outbound

This paper cites Scalable diffusion mod- els with transformers.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Scalable diffusion mod- els with transformers

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.312867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.307769Z digest=sha256:981322cfaa2fc8a50b7615877dc0a878bc07c7ee221244bfc39a154e22b05f16

Observation 2f1aacc2-f2d0-41d3-9538-e3d5554644ad · outbound

This paper cites Dreamfusion: Text-to-3d using 2d diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Dreamfusion: Text-to-3d using 2d diffusion

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.297932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.312138Z digest=sha256:04f47c88e01712b798c04c6255d4e9c691afeedaa4313a5eca4a8e116a572e52

Observation 796a8239-893d-4088-a8a6-e2019d5346b1 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Shapellm: Universal 3d object understanding for embodied interaction

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.283016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.316867Z digest=sha256:a94fc37aca9cdc7cbc192462cdff710a4a5b319fb29a689de795b505f2be75ff

Observation 567c8080-6bd6-45e9-bf39-32940f5fc96e · outbound

This paper cites Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to- 3d.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to- 3d

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.268213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.321156Z digest=sha256:ac8d07f234a36e419aa1b3d1e56057451fd9ad62996119ff21de5b5ff2fdeecf

Observation c2cd2614-9ed9-47c4-a5be-c334fbf10bc3 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Learn- ing transferable visual models from natural language super- vision

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.253444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.325324Z digest=sha256:2ce7548fc3c7ae0246c0ec5377978a90739431bce8ba41c9b598fc38ae36847f

Observation 20f051c1-874d-4cf9-82e5-a536a3c78d8c · outbound

This paper cites Zero-shot text-to-image generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Zero-shot text-to-image generation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.238557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.329719Z digest=sha256:fbbae91c6996b19bdd57bbbbf1eca13afb7d39e54989ebeb9ee08e34bbaa0de8

Observation ac0b05aa-cd3b-4429-a435-bfabd014500c · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE High-resolution image synthesis with latent diffusion models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.223192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.334678Z digest=sha256:6fc39543f8cf1e64c927750b57f02432d4a1b17de559617b53da8b49524349db

Observation 3772de48-2a47-4b2b-b4ad-e00100368b5e · outbound

This paper cites Pixelcnn++: Improving the pixelcnn with dis- cretized logistic mixture likelihood and other modifications.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Pixelcnn++: Improving the pixelcnn with dis- cretized logistic mixture likelihood and other modifications

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.208265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.339224Z digest=sha256:e3a592726a6f5a2c5db07b63e4664553e2a091a055d78fa308b3a36fd836cb83

Observation 58f7e6d9-7e1b-44d9-a6ad-1ed8a8b43e61 · outbound

This paper cites Flexible isosurface extraction for gradient-based mesh optimization.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Flexible isosurface extraction for gradient-based mesh optimization

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.192186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.343727Z digest=sha256:243e1c28964e3ee1b830ec693e5bd941b74cecdd14afcbd150520796ed1406bb

Observation 3b37534d-5ddd-42f6-ac33-d82dffeae3b7 · outbound

This paper cites Zero123++: a single image to consistent multi-view dif- fusion base model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Zero123++: a single image to consistent multi-view dif- fusion base model

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.177336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.348218Z digest=sha256:609d45b5150638dfb5d5c0b23e5fd2e60600ff933582db1063c86a163c20b054

Observation be267b2b-3b1f-496d-9b4a-9a4ccc4dc7c4 · outbound

This paper cites Mvdream: Multi-view diffusion for 3d gen- eration.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Mvdream: Multi-view diffusion for 3d gen- eration

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.162340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.352745Z digest=sha256:dc2700d7db89c537fb46b9c3d24ed43b1366b0d0f7bb2cbb47938ba2c0c7b362

Observation 80e3a86d-f7eb-4ab9-8687-43cbec1eac1b · outbound

This paper cites Meshgpt: Generating triangle meshes with decoder-only transformers.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Meshgpt: Generating triangle meshes with decoder-only transformers

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.147220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.357307Z digest=sha256:852542da5ed5d443c082ae78a11fa66618463bef2abd16781fb7bf7f5ebb1c1c

Observation c663576f-1194-4c4d-a4a3-4819f0b16a4b · outbound

This paper cites Light field networks: Neural scene representations with single-evaluation rendering.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Light field networks: Neural scene representations with single-evaluation rendering

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.131501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.361853Z digest=sha256:0f08fff4da1862b8db14b7efc4082e211cc7865c16a7f99e86d9af4f5c8d8fac

Observation 97ba22b4-e441-4f62-80bf-10b60d87018d · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Score-based generative modeling through stochastic differential equa- tions

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.366352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.366352Z digest=sha256:3e385ab9c63eb543ba94022adc94f16cc843cf41688c7cdeaee6bfdf05ece978

Observation 224e466d-c10e-445c-817f-5d8188c05ef5 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.370840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.370840Z digest=sha256:a81522c533ef3a2b78e88a463b3ac3cc74444387c39de779f9a1e333b0204639

Observation 1769e33d-a632-4ad3-a21e-6531e2ef35f7 · outbound

This paper cites Emu: Generative pretraining in multimodality.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Emu: Generative pretraining in multimodality

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.106128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.375560Z digest=sha256:545794cf3dcdbb42c661686252372476df79b5944d168a4a737cff508796c020

Observation d228f497-3fc7-4005-b49f-e901727d11ee · outbound

This paper cites Splatter image: Ultra-fast single-view 3d recon- struction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Splatter image: Ultra-fast single-view 3d recon- struction

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.091254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.379853Z digest=sha256:666c5e932ed68335b1e90248442f6df2b8d2fad6c81704698013960a8b36ade2

Observation a63fc588-d30e-4949-9916-bab4ce75b947 · outbound

This paper cites Make-it-3d: High-fidelity 3d creation from a single image with diffusion prior.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Make-it-3d: High-fidelity 3d creation from a single image with diffusion prior

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.384092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.384092Z digest=sha256:4471f6afadb949a0bdcb53611018a9d63796b8c1ed83ad29bfb67a37cf9fa399

Observation fecfb698-b291-466e-9e08-fd82b0aec056 · outbound

This paper cites Lgm: Large multi-view gaus- sian model for high-resolution 3d content creation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lgm: Large multi-view gaus- sian model for high-resolution 3d content creation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.066493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.388785Z digest=sha256:6857325f3a1f91db405b2c815c9c1c3de38b952cf9267fcfa898102c5c52ce13

Observation ff91c9b5-1d4e-479b-9ec9-c7290ad391b4 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Visual autoregressive modeling: Scalable image generation via next-scale prediction

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.051889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.393225Z digest=sha256:6a8cada035460a06373261fd83b727d795631abfe4912fc6408e90851528aa60

Observation 816442a4-3cdd-4ec8-a585-1c09b6a00ffc · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE LLaMA: Open and Efficient Foundation Language Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.402317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.402317Z digest=sha256:2c9532b84d74ab6af6aa031cd84dd6c628cc9498c2f424e76ca8101786889344

Observation 4b267201-b8f4-4334-b540-8681a8d130c4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.407101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.407101Z digest=sha256:8a6edf06fa998ad8ffb540f48916e04986075faec14ed91585ee2bded32cf481

Observation ba607be5-e79e-4c90-993e-4d829836b422 · outbound

This paper cites Lion: Latent point diffu- sion models for 3d shape generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lion: Latent point diffu- sion models for 3d shape generation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.022066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.411761Z digest=sha256:b62f559d6cca6d2fb3834c1ede80059c57c54a05b46ac63283a55007dfbad70c

Observation 6a826902-1301-4341-8879-43b6d004df67 · outbound

This paper cites Neural discrete representation learning.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Neural discrete representation learning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.006542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.416090Z digest=sha256:59c0f557eb358aa606c6394ac54fea3617b327bc96fa083607309853a9f838b4

Observation 4d3a26fb-8d3d-4ff7-acab-3a21fe65eff1 · outbound

This paper cites Rodin: A generative model for sculpting 3d digital avatars using diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Rodin: A generative model for sculpting 3d digital avatars using diffusion

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.990378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.420602Z digest=sha256:a46aa35d04422d874d2d5af7c1c082124a356d48cef29a507530b2a355cf1a0d

Observation db1fc0c7-1445-4437-b2b6-9d62b50b19a8 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Emu3: Next-Token Prediction is All You Need

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.425149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.425149Z digest=sha256:f579a60af5381aa6c0dcbd094647598ebda807450676db860364bfa41c5b80ef

Observation 05309087-c25b-430e-b3c0-c59489f30767 · outbound

This paper cites Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distilla- tion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distilla- tion

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.975676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.429833Z digest=sha256:a8c63475ce226b0ea349721b21dd25abda52d2b1c6619de454b883dfff387ca6

Observation 701c92f0-bc13-44ac-8a57-c377e6588531 · outbound

This paper cites Crm: Single image to 3d textured mesh with convo- lutional reconstruction model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Crm: Single image to 3d textured mesh with convo- lutional reconstruction model

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.959484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.434286Z digest=sha256:6e75b4c03b66d8efb3e8b9ea0d4603ed61c7afaa9393be073d258d5ff4af3710

Observation 78982d5a-598e-41fb-bfee-b6d6efe7193b · outbound

This paper cites Phidias: A gen- erative model for creating 3d content from text, image, and 3d conditions with reference-augmented diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Phidias: A gen- erative model for creating 3d content from text, image, and 3d conditions with reference-augmented diffusion

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.943707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.438637Z digest=sha256:7406bace4a22fcb4014abcc01cd5b9a257c72f5eb1151a1bbfcdfec25a28f79e

Observation a6583e60-3763-424f-99e0-d0b6b08bac26 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.442970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.442970Z digest=sha256:230c19ba174ef704fd5fe5d0b43628e5957f9db5ac6051bf436d3ac13750f625

Observation 4d699d09-379b-43c4-9bdb-e4399817abd9 · outbound

This paper cites Multiview compres- sive coding for 3d reconstruction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Multiview compres- sive coding for 3d reconstruction

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.928276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.447766Z digest=sha256:cc673ba92126322c577572f600271f7e4c65c379bf37dc32a7e58fa6a58215e1

Observation 9d26c114-c46d-4507-ac0b-ef6fb07808fa · outbound

This paper cites Direct3d: Scal- able image-to-3d generation via 3d latent diffusion trans- former.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Direct3d: Scal- able image-to-3d generation via 3d latent diffusion trans- former

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.912941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.452097Z digest=sha256:aee7a62f08df0dfdc6d942fcba9f0866681fb3a9c8bb39379c712763a37cf0b5

Observation 4aaea32d-9c2e-48a4-ac9b-651ae9c7539c · outbound

This paper cites Latte3d: Large-scale amortized text-to-enhanced3d synthe- sis.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Latte3d: Large-scale amortized text-to-enhanced3d synthe- sis

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.897037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.456524Z digest=sha256:b98249544bbd997dfb718b95af88f23382f0fabcbaaf7391b529dcdaa087795f

Observation d7f248cd-c04c-459b-a3bc-a3b671bb05af · outbound

This paper cites InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.461037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.461037Z digest=sha256:d95da88d613568d2d7441ff4abf1fa7796fec9948cfa0cd02621667b138acff1

Observation 84455d40-91e0-4e4f-98dc-c6d2372321d2 · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Pointllm: Empowering large language models to understand point clouds

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.882030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.465511Z digest=sha256:9720540c46c47dd3c6f4d07a8f8d2c68c1a04615eba355ffa888974621caf298

Observation 1a90409d-3647-47f8-aae7-c39f14720c11 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.469877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.469877Z digest=sha256:16b9fa53503111d3c8584d0ff92de700b12ff3a6f195333ae2856dca7ff81545

Observation 3b67905f-e8e4-4a2f-b743-259b5a04c5ba · outbound

This paper cites Scaling autoregressive models for content-rich text-to-image generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Scaling autoregressive models for content-rich text-to-image generation

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.857578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.474119Z digest=sha256:3a73f19fce78236e2d7c182b8936142606e21d6a19dcb4ccb50d5503c7cf4682

Observation ee76b518-11cd-480a-baf5-f0688ea3ccfd · outbound

This paper cites Language model beats diffusion-tokenizer is key to visual generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Language model beats diffusion-tokenizer is key to visual generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.842554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.478339Z digest=sha256:fa0751203a35b175753cc246521e2c3cbcf13c14e5e8a860071208a5c8e943ac

Observation 57331447-3d86-4891-b60a-9b540cf5a6cb · outbound

This paper cites An image is worth 32 tokens for reconstruction and generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE An image is worth 32 tokens for reconstruction and generation

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.827858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.482482Z digest=sha256:a4792c88a51883443289d5961f4f15b9b50f69c1f95c15b625412b90ebfbb60f

Observation 9de8e88f-e922-4e96-80e6-bd0eb7c5b7c9 · outbound

This paper cites Point-bert: Pre-training 3d point cloud transformers with masked point modeling.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Point-bert: Pre-training 3d point cloud transformers with masked point modeling

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.813443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.486667Z digest=sha256:2aa13d6b1c5be76fd2a7e5e319c7d7cca54809f3db4d6fa483a05522f54af097

Observation c8946e61-1d34-480c-8688-d2a6a489a727 · outbound

This paper cites 3dshape2vecset: A 3d shape representation for neu- ral fields and generative diffusion models.ACM Transactions on Graphics (TOG), 42(4):1–16, 2023.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 3dshape2vecset: A 3d shape representation for neu- ral fields and generative diffusion models.ACM Transactions on Graphics (TOG), 42(4):1–16, 2023

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.490938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.490938Z digest=sha256:bd86ece9453915d9b85bc58705b303e0723db59d6e08ee077a4ef79b1ef66d01

Observation 8e2e7596-eb7c-4c0a-80e4-e1a74289d56c · outbound

This paper cites Clay: A controllable large-scale generative model for creat- ing high-quality 3d assets.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Clay: A controllable large-scale generative model for creat- ing high-quality 3d assets

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.788673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.495714Z digest=sha256:f5de5b80c981e2bdbdbb24b1f6badb8b5a0e32a5ff3ce4fb6707d96448234526

Observation 17f8524b-3f64-43b8-9b2d-4656f83fa133 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE The unreasonable effectiveness of deep features as a perceptual metric

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.773583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.499987Z digest=sha256:46216b46b7e6af0d75efd1172fd45973609801d81b08b967a7f0b16bc7b7f999

Observation a44ad04c-6f0a-4da6-aab8-b1eebe377d5a · outbound

This paper cites Multi-HeadSelf Attention Multi-HeadCross Attention ×N Scale, Shift Layer Norm CLIPT Text Scale Scale ⊕ ⊕ (a) Transformer Block (Text condition) FFN𝛾!,𝛽!Scale, Shift 𝛼! Layer Norm𝛾.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Multi-HeadSelf Attention Multi-HeadCross Attention ×N Scale, Shift Layer Norm CLIPT Text Scale Scale ⊕ ⊕ (a) Transformer Block (Text condition) FFN𝛾!,𝛽!Scale, Shift 𝛼! Layer Norm𝛾

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.757885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.504301Z digest=sha256:2eacdfcdc88142f6094482ad683c575b0a61949d674cc9d4b395a055615ea927

Observation 9cdeec13-9d32-4cab-975a-c39e131d86bd · outbound

This paper cites an unresolved cited work.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T12:52:11.037218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T12:52:10.397825Z digest=sha256:b9e245a0169ee961b2452acab1aff57895cc623298572f7dd3af6f858a1778c0

Pith citing papers

Observation 76acba24-64d5-4712-8ed8-482cebe95928 · inbound

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models cites this paper.

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:50:14.784325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T14:48:21.787919Z digest=sha256:f3a53ad4043fc6cad1dd80526424ac8b99c66ff0bbec934b774923c00ae790bd