Pith. sign in

Paper Citation Record · LEDGER

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE

As of 14 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 1 inbound Pith citation observation for arXiv:2411.16856.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16856 v3

Coverage vector

measured 93 of 93 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:52:10.504301Z

measured 94 of 94 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T14:48:21.787919Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

93 of 93 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved36
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b2a4df05-a876-4940-982b-184584b824ac · outbound

This paper cites GPT-4 Technical Report.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.084365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.084365Z digest=sha256:2d67bf420062e307a64c51eac48f9eb659a282e98e4b6a2ff727f82dce7ba0b6

Observation 8c4086a9-fbd9-4ef4-9b52-53b7bd4d1d21 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.089742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.089742Z digest=sha256:c30d0f1ba510e2a762ca9773baa85f3145ad459741b3d3cd47a8773e517d6891

Observation e5f5b9ae-b026-429d-b1cf-e477c957156e · outbound

This paper cites Demystifying mmd gans.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Demystifying mmd gans

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.094793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.094793Z digest=sha256:924907f3db2b5586b490b758e160deef144c85db20485527e712e2b4eea8aab9

Observation df38ee96-2842-4930-a5d3-4feb28a9c736 · outbound

This paper cites Video generation models as world simulators.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Video generation models as world simulators

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.099617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.099617Z digest=sha256:0caa5dc92a08bd3dca1a4977a26081c49755fb037736586c4cafd7262bbe1cad

Observation f884c0f6-02f5-406a-b35c-001fd00115e9 · outbound

This paper cites Lan- guage models are few-shot learners.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lan- guage models are few-shot learners

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.104973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.104973Z digest=sha256:6fcdf9dff849c15d8a3ed0b0628571ecde2efaaf0fb9c97fde172bd56c203cde

Observation 3b334e59-2828-4a35-b637-51989331ffd5 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE ShapeNet: An Information-Rich 3D Model Repository

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.109776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.109776Z digest=sha256:f996f190ba1dd7a874ab674cedb0b6272dae45231ae575c08aa665bcfdeb3652

Observation ddeb8333-e54e-42ea-80c0-a74877a506bc · outbound

This paper cites Maskgit: Masked generative image transformer.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Maskgit: Masked generative image transformer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.114779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.114779Z digest=sha256:011134aea8028938605dbb7da086cb8e246b2ebd92b411f122a75a03c4bd08dc

Observation 184058bf-2767-4e7a-a36e-aeec197f0482 · outbound

This paper cites Lara: Efficient large-baseline radiance fields.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lara: Efficient large-baseline radiance fields

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.119338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.119338Z digest=sha256:0569a2950685b0a0374b416de6b58e3991a8ed404469b2945b634e16c3d7ef90

Observation a34a7808-208e-441d-8c59-a5df9574e0ae · outbound

This paper cites Fan- tasia3d: Disentangling geometry and appearance for high- quality text-to-3d content creation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Fan- tasia3d: Disentangling geometry and appearance for high- quality text-to-3d content creation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.124518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.124518Z digest=sha256:0feea5c6e5c1764b30e7572cecf916860bdeb90216eef11bab340035e029d089

Observation 8eff5223-12e7-458b-bf06-b92f0982edee · outbound

This paper cites Comboverse: Compositional 3d assets creation using spatially-aware diffusion guidance.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Comboverse: Compositional 3d assets creation using spatially-aware diffusion guidance

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.129061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.129061Z digest=sha256:b2e76bcd30fac924bb5a2e5afd4e9a61b112871358c9b4465ab422bf923d474a

Observation 717632a5-fac5-472f-aa8f-877600fbbe8b · outbound

This paper cites Meshanything: Artist- created mesh generation with autoregressive transformers.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Meshanything: Artist- created mesh generation with autoregressive transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.134306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.134306Z digest=sha256:4b915afc487d73493f1908f779667d4dec97cea2dea672639435d0b6c360ea99

Observation 4a792050-dd0a-4f1f-865e-0224f030df8c · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.138746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.138746Z digest=sha256:deb9a8a76535a4afb8270c49b353f69ffd5a970ec4afa80e2718c86ea5bce262

Observation 83d55a0b-bab7-4463-8248-10cc1d91c3e9 · outbound

This paper cites Palm: Scaling language modeling with pathways.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Palm: Scaling language modeling with pathways

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.143385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.143385Z digest=sha256:e30cfb560a695e9ec7d4ab9eaf9e916ccdbf61e99ea83eb79bfbca837c0d0433

Observation cb27c8d0-7c9d-4a11-a828-2c3e028c21eb · outbound

This paper cites Objaverse-xl: A universe of 10m+ 3d objects.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Objaverse-xl: A universe of 10m+ 3d objects

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.147883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.147883Z digest=sha256:29df844d8a9faf3d97ff8e399bec57d975d0ae7d96c43a4a556ffe323fec739b

Observation 3e8856e4-e1a5-4ecd-9e16-df802fdf20c3 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Objaverse: A universe of annotated 3d objects

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.152354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.152354Z digest=sha256:76e7e921f0c382b6585c7af0b31becf103c72ad12d14ad82233c0896d58034b0

Observation 83534297-5288-4a3d-a4d6-5909c0e8c591 · outbound

This paper cites Palm- e: An embodied multimodal language model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Palm- e: An embodied multimodal language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.684263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.156786Z digest=sha256:20fd578587f9dc8c2c830a51841fb465cf64b460bffce4c7c0b46df01f69506b

Observation b154415c-54c7-44a0-a891-e0ca88c199d2 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Taming transformers for high-resolution image synthesis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.161472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.161472Z digest=sha256:c350c874a4a3287264ab477972e018f92ff74e84c3865c7c6537e9c6d394e59e

Observation d718883a-8478-4927-8f57-37d366037e8d · outbound

This paper cites Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.167237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.167237Z digest=sha256:ec8247f66006172806d8132187d6f4c3b3d912297f7ee92704d4039743e6200c

Observation 93153d77-a56d-4b25-8ba6-f69134e1b7a0 · outbound

This paper cites OpenLRM: Open-source large reconstruction models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE OpenLRM: Open-source large reconstruction models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.660835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.172277Z digest=sha256:e18d03c7b7b0ec0b113490b938cb5e9901c03e2412b1daaf7264262f23dc097c

Observation 3cd612bb-cc44-4dd7-a04e-c3c1fb227622 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.176860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.176860Z digest=sha256:e85d2b110931d9b6cae0affc575e3dd6d152137c73fdc4c9510bd80e32e5adf2

Observation 83bdb322-003d-410b-8fea-bd9373a9e823 · outbound

This paper cites Classifier-free diffusion guidance.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Classifier-free diffusion guidance

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.181260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.181260Z digest=sha256:746ef9fe472725dd7999286209b2aede850479f23ae56a503820dc609ad5c9ae

Observation 36b8f98f-24b5-4000-9d08-e51e92ef46ab · outbound

This paper cites Denoising diffu- sion probabilistic models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Denoising diffu- sion probabilistic models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.185923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.185923Z digest=sha256:80c1dd715d9fe39a01eba476d6a70461f24454d55f5455456624430bc7a48297

Observation 79c6ec6d-3616-425a-bbad-a410c7927808 · outbound

This paper cites 3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.190409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.190409Z digest=sha256:230bf7ec8478a106adfce572c2f510d8c66f796e4bcee0c95787bc5ac112162e

Observation 21ead568-347e-494b-a0ac-b0e44806e213 · outbound

This paper cites 3d-llm: In- jecting the 3d world into large language models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 3d-llm: In- jecting the 3d world into large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.618201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.195390Z digest=sha256:10ef734940146865757566ff55a43479156b7bf9bcf1b33439483341ad28e76e

Observation 5bff3f0c-8f8f-47fe-ba4d-657e9c54ad42 · outbound

This paper cites Lrm: Large reconstruction model for single image to 3d.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lrm: Large reconstruction model for single image to 3d

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.603468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.199835Z digest=sha256:a8efaee29030e71f79cc5107a9698ada68f9d5625aaacd3d90c50c3c6167bba5

Observation a986443c-a237-4e31-aec1-0414030b3afa · outbound

This paper cites 2d gaussian splatting for geometrically ac- curate radiance fields.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 2d gaussian splatting for geometrically ac- curate radiance fields

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.588870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.204127Z digest=sha256:e06d6d867cd8594fbd529155713a7b99701b34b259f8d87f6da30ddced08cc21

Observation 1de84dd7-41c9-4a7f-95ef-0f1efe65096c · outbound

This paper cites Shap-E: Generating Conditional 3D Implicit Functions.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Shap-E: Generating Conditional 3D Implicit Functions

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.208683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.208683Z digest=sha256:8d6f3b6fce74b9ff534622870cf8b2f4c5afae8925e6c7f3717c5fce8ba07461

Observation 6bf5d517-9a33-467a-9189-07787be4c080 · outbound

This paper cites Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.573772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.213565Z digest=sha256:e994049ae7f968f41c981b0383c200f787a451db5982128a03a505232178aed5

Observation d82dd398-3e3a-4244-bfc4-c2bb78827445 · outbound

This paper cites Musiq: Multi-scale image quality transformer.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Musiq: Multi-scale image quality transformer

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.558286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.218183Z digest=sha256:992c8c9e7845513d7555121eff70f5be0ac07c56f804f4fc7f6d505c133abdaa

Observation 84d6e0a7-162f-4594-893a-f1a0a03f4782 · outbound

This paper cites Kingma and Max Welling.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Kingma and Max Welling

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.542881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.222597Z digest=sha256:ea24dd047b9782bcc5ae5928a4eac729b2747765325e7e8f131009294ebca45e

Observation 181b492b-c64a-4e69-90a0-9d0d4ce2a62c · outbound

This paper cites Nerf-vae: A geometry aware 3d scene generative model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Nerf-vae: A geometry aware 3d scene generative model

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.527816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.226855Z digest=sha256:380b8af957b07a2636b6ffbbd7954684ee824bfaf0802b3995c9aab02c061bd0

Observation 22b3f0fa-a7ca-48ef-a3fc-bccd6703cdf7 · outbound

This paper cites Ln3diff: Scalable latent neural fields diffusion for speedy 3d generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Ln3diff: Scalable latent neural fields diffusion for speedy 3d generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.512855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.231057Z digest=sha256:2c8387ecd28578890b5c86ff7b081ec69a789c85d1549b656a31619e9ab5e3a2

Observation fbffe989-7344-4611-8b0a-7cd8a3bf9d2f · outbound

This paper cites Gaussiananything: Interactive point cloud latent dif- fusion for 3d generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Gaussiananything: Interactive point cloud latent dif- fusion for 3d generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.497786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.235292Z digest=sha256:263c1d3979aec32588ee4dfba31501c498956963c8674ed95fec3b7382bc0ba2

Observation 27814b0d-7cac-45d9-9e86-c13f9bb2ff0b · outbound

This paper cites Autoregressive image generation using residual quantization.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Autoregressive image generation using residual quantization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.240211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.240211Z digest=sha256:62709b54b48bc812ea6d307b6d313dd2c9c8d133806e5e64247f27b21f33583e

Observation 04dc7fde-f63d-459e-92e9-b11b3ef14017 · outbound

This paper cites CraftsMan: High-fidelity mesh generation with 3D native generation and interactive geometry refiner.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE CraftsMan: High-fidelity mesh generation with 3D native generation and interactive geometry refiner

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.473350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.245097Z digest=sha256:09f700ea1de4103919597f85431e309df8abef8d47f1c5284d299689d6ba5d36

Observation 4e718b6e-cb74-4811-b9ef-9792feeede6a · outbound

This paper cites Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.249652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.249652Z digest=sha256:6b206e813cc3dc0f9a7753b7458fa7283377fad3366613ee523a810f49d6c051

Observation 5a3db5ba-30c3-46c4-9604-9d0458f99b00 · outbound

This paper cites Visual instruction tuning.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Visual instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.458093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.254399Z digest=sha256:31ed4ec9dd17030a71dc92287815130edd884d7c992a361e4de6fc9300a3d992

Observation e54ce1db-f26d-4554-b3ab-0f3c6fd6a12a · outbound

This paper cites One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimiza- tion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimiza- tion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.443893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.258753Z digest=sha256:215002eebcdf068f2306d70877496dbcca67174410c8c2f9602c060f95f0f12f

Observation cf297b72-e8f7-40f3-8983-f1a1067f77c1 · outbound

This paper cites Zero-1-to- 3: Zero-shot one image to 3d object.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Zero-1-to- 3: Zero-shot one image to 3d object

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.429464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.263015Z digest=sha256:fd94327c532ab2f477427a28e854906eab1faec7a0d631b295aa920b5f73dfa0

Observation d50b8a05-abde-4f19-aa22-9f6668fad508 · outbound

This paper cites Wonder3d: Sin- gle image to 3d using cross-domain diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Wonder3d: Sin- gle image to 3d using cross-domain diffusion

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.414849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.267215Z digest=sha256:5cdfd3b10a7f46ea437e7990b76708437b1f3d6d5fba20d2e15bcc3c91b904ec

Observation 08d8f544-0920-4df0-8b45-229b557c645f · outbound

This paper cites Scalable 3d captioning with pretrained models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Scalable 3d captioning with pretrained models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.400539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.271394Z digest=sha256:9e69e09bcb53e3dc0dac7a13dd9386b66bfe64b25b3e94a3797bb7c0442a526b

Observation abf3f6b4-2824-4c8f-b435-f71e999c5eb3 · outbound

This paper cites View selec- tion for 3d captioning via diffusion ranking.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE View selec- tion for 3d captioning via diffusion ranking

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.386415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.275802Z digest=sha256:7271101b94a9bc04957a13ab12cb597c36805978f0b329284889f577abf4391a

Observation 227ab008-0fa0-484c-aa81-4b0e5b2a84ea · outbound

This paper cites KOSMOS-2.5: A Multimodal Literate Model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE KOSMOS-2.5: A Multimodal Literate Model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.280261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.280261Z digest=sha256:3e66f3053662179189d1051a3982677214d6ec2ac133649cc043650d847f7eb1

Observation 61a817c7-f2a2-4815-9f90-d634cbd7fff4 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.371963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.284950Z digest=sha256:ab106d1fe4cba996bc1706dcabf13433816c66566f0f80e2fddf75efba109c37

Observation 08c8c0b3-89f7-4349-8453-4b8c86293b06 · outbound

This paper cites Autosdf: Shape priors for 3d comple- tion, reconstruction and generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Autosdf: Shape priors for 3d comple- tion, reconstruction and generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.357222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.289490Z digest=sha256:2d64bc32078a14fed9b73041d0dd438d4d70c27ed7f8ca90c433e9d9778cd963

Observation 3cc9fa97-dc11-4ef5-af1a-14817b4e84bd · outbound

This paper cites Point-E: A System for Generating 3D Point Clouds from Complex Prompts.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Point-E: A System for Generating 3D Point Clouds from Complex Prompts

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.293892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.293892Z digest=sha256:2f9f34ba5fcaa2b3cfb572c638d14058cd48248835f7eeefded58b4425647ef0

Observation 3372f6f5-95f2-4b1e-a396-fda875132564 · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Dinov2: Learning robust visual features without supervision

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.342357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.298697Z digest=sha256:2b9e062598067ee92d2291e01ba5f60ac625dc3515ac59723903be0c139ffc71

Observation c7ec7313-fb60-4e9f-986c-27154643d9d5 · outbound

This paper cites Training language models to follow instructions with human feedback.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Training language models to follow instructions with human feedback

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.327504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.303249Z digest=sha256:6b334d372a78f18ac399751c345a66064edd1a4eae28db8600a4278ec742f94f

Observation 94b646af-38b1-4504-adf1-738999ce0ad1 · outbound

This paper cites Scalable diffusion mod- els with transformers.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Scalable diffusion mod- els with transformers

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.312867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.307769Z digest=sha256:d528326d3e5860636e7b13c5751d0901b056d0e9866af524f6293ff93fdf48ec

Observation 2f1aacc2-f2d0-41d3-9538-e3d5554644ad · outbound

This paper cites Dreamfusion: Text-to-3d using 2d diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Dreamfusion: Text-to-3d using 2d diffusion

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.297932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.312138Z digest=sha256:d625770045b81a2a80ed160404ae64f3ee3aaab31076176352227607e7dd95ca

Observation 796a8239-893d-4088-a8a6-e2019d5346b1 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Shapellm: Universal 3d object understanding for embodied interaction

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.283016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.316867Z digest=sha256:e7d5fe7a39435b05380e65ec4b1a7e50101d0bb12588f0c9efc1c0506b3610a8

Observation 567c8080-6bd6-45e9-bf39-32940f5fc96e · outbound

This paper cites Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to- 3d.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to- 3d

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.268213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.321156Z digest=sha256:518573f9f966d381161f5a511988120a483913c901ce11a2cee1f8e457578d1d

Observation c2cd2614-9ed9-47c4-a5be-c334fbf10bc3 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Learn- ing transferable visual models from natural language super- vision

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.253444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.325324Z digest=sha256:aefe9018cbcb0d04ac418218353b6d9ff91aabcab6f5f079c27c48b9eb383264

Observation 20f051c1-874d-4cf9-82e5-a536a3c78d8c · outbound

This paper cites Zero-shot text-to-image generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Zero-shot text-to-image generation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.238557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.329719Z digest=sha256:6d58935b85b12044ff0a35ea97e5150fd83b72f8714735c707330a18a2a0cb48

Observation ac0b05aa-cd3b-4429-a435-bfabd014500c · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE High-resolution image synthesis with latent diffusion models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.223192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.334678Z digest=sha256:4922938cd03e14d3224862eced6c726d242c86a4b2c8eff3ce2131cb5bf1a083

Observation 3772de48-2a47-4b2b-b4ad-e00100368b5e · outbound

This paper cites Pixelcnn++: Improving the pixelcnn with dis- cretized logistic mixture likelihood and other modifications.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Pixelcnn++: Improving the pixelcnn with dis- cretized logistic mixture likelihood and other modifications

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.208265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.339224Z digest=sha256:48cb49006148a24a85c3233f15f10eeeaf3c0ed3385ad7268f5814ff87aa2345

Observation 58f7e6d9-7e1b-44d9-a6ad-1ed8a8b43e61 · outbound

This paper cites Flexible isosurface extraction for gradient-based mesh optimization.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Flexible isosurface extraction for gradient-based mesh optimization

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.192186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.343727Z digest=sha256:873ba2c45a385ca1d8f43cee5167d19b09233f643eebc294156a7d6636e3ce4e

Observation 3b37534d-5ddd-42f6-ac33-d82dffeae3b7 · outbound

This paper cites Zero123++: a single image to consistent multi-view dif- fusion base model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Zero123++: a single image to consistent multi-view dif- fusion base model

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.177336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.348218Z digest=sha256:63f72ba478bba75c20a1871ff0f71edaa217be97ccadf4f5dfb87c52ab9e3d84

Observation be267b2b-3b1f-496d-9b4a-9a4ccc4dc7c4 · outbound

This paper cites Mvdream: Multi-view diffusion for 3d gen- eration.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Mvdream: Multi-view diffusion for 3d gen- eration

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.162340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.352745Z digest=sha256:afd9ff7813f657fcdbdf3b4bbaf8a9c31050504481eeffe3916051061690af27

Observation 80e3a86d-f7eb-4ab9-8687-43cbec1eac1b · outbound

This paper cites Meshgpt: Generating triangle meshes with decoder-only transformers.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Meshgpt: Generating triangle meshes with decoder-only transformers

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.147220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.357307Z digest=sha256:4eda9c93acbbd024396dc845cf11934df07a6d16205908c94cc267f6c7256a78

Observation c663576f-1194-4c4d-a4a3-4819f0b16a4b · outbound

This paper cites Light field networks: Neural scene representations with single-evaluation rendering.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Light field networks: Neural scene representations with single-evaluation rendering

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.131501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.361853Z digest=sha256:3e3ea283b461a5371f4c7055e919aaaf55e38d569cc7d66f958386da6628251e

Observation 97ba22b4-e441-4f62-80bf-10b60d87018d · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Score-based generative modeling through stochastic differential equa- tions

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.366352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.366352Z digest=sha256:93e0e16e9a09fc1d1ce5ba54e38872ce329f8bb7884600df91ad7269f834ac3c

Observation 224e466d-c10e-445c-817f-5d8188c05ef5 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.370840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.370840Z digest=sha256:a8d0c11f999b5f35ea6ae44cf19f9d7dfdfe36a67b4da056627705c93498b373

Observation 1769e33d-a632-4ad3-a21e-6531e2ef35f7 · outbound

This paper cites Emu: Generative pretraining in multimodality.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Emu: Generative pretraining in multimodality

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.106128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.375560Z digest=sha256:14d067801c717a4516c1f6d75304a10179ee7fe7aea47b4bf548fe7c05d33122

Observation d228f497-3fc7-4005-b49f-e901727d11ee · outbound

This paper cites Splatter image: Ultra-fast single-view 3d recon- struction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Splatter image: Ultra-fast single-view 3d recon- struction

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.091254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.379853Z digest=sha256:f77bbe4780a281932a968fac4a326efae82a8026d04824430d271b11867da1d0

Observation a63fc588-d30e-4949-9916-bab4ce75b947 · outbound

This paper cites Make-it-3d: High-fidelity 3d creation from a single image with diffusion prior.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Make-it-3d: High-fidelity 3d creation from a single image with diffusion prior

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.384092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.384092Z digest=sha256:eda5b2f0d255675968fddd78fb8e95a1d3d32d36f3177ad2fa13c27097830f8d

Observation fecfb698-b291-466e-9e08-fd82b0aec056 · outbound

This paper cites Lgm: Large multi-view gaus- sian model for high-resolution 3d content creation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lgm: Large multi-view gaus- sian model for high-resolution 3d content creation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.066493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.388785Z digest=sha256:966d4cd72b703d835c037680c82329e8e1b7305ef31864ca9197a549a42f0f23

Observation ff91c9b5-1d4e-479b-9ec9-c7290ad391b4 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Visual autoregressive modeling: Scalable image generation via next-scale prediction

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.051889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.393225Z digest=sha256:ad28e9a9e41ac0d5aff4cc2970ae279427ae8a7661bea2dd9cd847c881068d11

Observation 816442a4-3cdd-4ec8-a585-1c09b6a00ffc · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE LLaMA: Open and Efficient Foundation Language Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.402317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.402317Z digest=sha256:c9c3776454d526bba3edcc4647e88a21d4f2e98c301b4db9d1bf0c3ec8cb706d

Observation 4b267201-b8f4-4334-b540-8681a8d130c4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.407101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.407101Z digest=sha256:92f3844e81433b7cece77dee77dd2b0057253dbc51618498f2a6706911da97fa

Observation ba607be5-e79e-4c90-993e-4d829836b422 · outbound

This paper cites Lion: Latent point diffu- sion models for 3d shape generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Lion: Latent point diffu- sion models for 3d shape generation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.022066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.411761Z digest=sha256:694562c78bc989993b55a97c549069d41a16237f3159d6b7ce9be8b07859ec69

Observation 6a826902-1301-4341-8879-43b6d004df67 · outbound

This paper cites Neural discrete representation learning.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Neural discrete representation learning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:11.006542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.416090Z digest=sha256:7fe71e6f2136538627e7ca59e3bd107317caff455335cfd1275d5ad883d0f5c1

Observation 4d3a26fb-8d3d-4ff7-acab-3a21fe65eff1 · outbound

This paper cites Rodin: A generative model for sculpting 3d digital avatars using diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Rodin: A generative model for sculpting 3d digital avatars using diffusion

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.990378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.420602Z digest=sha256:7f4975838cb1e2d96f7fe591afdbe4a4188c340c0e99dfdc3997464632254787

Observation db1fc0c7-1445-4437-b2b6-9d62b50b19a8 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Emu3: Next-Token Prediction is All You Need

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.425149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.425149Z digest=sha256:acc94146709adc9e066988870c7ef15f08b86cc667b9d61cfa87821a7ec602b2

Observation 05309087-c25b-430e-b3c0-c59489f30767 · outbound

This paper cites Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distilla- tion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distilla- tion

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.975676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.429833Z digest=sha256:5272951b69571e143aaaf96013c5d770bb53af6547ad2d0b7c4a5502f4703291

Observation 701c92f0-bc13-44ac-8a57-c377e6588531 · outbound

This paper cites Crm: Single image to 3d textured mesh with convo- lutional reconstruction model.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Crm: Single image to 3d textured mesh with convo- lutional reconstruction model

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.959484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.434286Z digest=sha256:30a22f24aa2134cf532c08df8b7a8842b6ebe8bbdc09be61ed03af7139b03bc2

Observation 78982d5a-598e-41fb-bfee-b6d6efe7193b · outbound

This paper cites Phidias: A gen- erative model for creating 3d content from text, image, and 3d conditions with reference-augmented diffusion.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Phidias: A gen- erative model for creating 3d content from text, image, and 3d conditions with reference-augmented diffusion

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.943707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.438637Z digest=sha256:f53dba2952010f19f793da57619fb4a4037ce6a3744b75edc76fff1b8dd6ee71

Observation a6583e60-3763-424f-99e0-d0b6b08bac26 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.442970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.442970Z digest=sha256:337f942074bddb98a69bafb7290d72d107511ff003d8a3edc18d1ad5132ba668

Observation 4d699d09-379b-43c4-9bdb-e4399817abd9 · outbound

This paper cites Multiview compres- sive coding for 3d reconstruction.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Multiview compres- sive coding for 3d reconstruction

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.928276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.447766Z digest=sha256:b8ff3a4738d41e067c9ab442f060d12759d7bdc243b6729f0299e488a49dae81

Observation 9d26c114-c46d-4507-ac0b-ef6fb07808fa · outbound

This paper cites Direct3d: Scal- able image-to-3d generation via 3d latent diffusion trans- former.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Direct3d: Scal- able image-to-3d generation via 3d latent diffusion trans- former

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.912941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.452097Z digest=sha256:21abce3ba6e8c3da4cb7a38f3091640704b6beb61b34b013f799e87f08f1f35d

Observation 4aaea32d-9c2e-48a4-ac9b-651ae9c7539c · outbound

This paper cites Latte3d: Large-scale amortized text-to-enhanced3d synthe- sis.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Latte3d: Large-scale amortized text-to-enhanced3d synthe- sis

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.897037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.456524Z digest=sha256:38c48e4163691ecf23c7a292349bc48b138f35df76743d1de08b1fe566714cd1

Observation d7f248cd-c04c-459b-a3bc-a3b671bb05af · outbound

This paper cites InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.461037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.461037Z digest=sha256:1544f6915d08934830cda6fd0ac14279334d90bc9df45ca0294fd274098e56aa

Observation 84455d40-91e0-4e4f-98dc-c6d2372321d2 · outbound

This paper cites Pointllm: Empowering large language models to understand point clouds.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Pointllm: Empowering large language models to understand point clouds

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.882030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.465511Z digest=sha256:adf579b26cac346262c91afa4aac93cf6b974b5c730422c4bc4ab57af7aaee8f

Observation 1a90409d-3647-47f8-aae7-c39f14720c11 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.469877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.469877Z digest=sha256:102bda8e4bfa1b6979e08b581927e7284a9c0d33842b43e2f9967d58bc8a0839

Observation 3b67905f-e8e4-4a2f-b743-259b5a04c5ba · outbound

This paper cites Scaling autoregressive models for content-rich text-to-image generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Scaling autoregressive models for content-rich text-to-image generation

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.857578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.474119Z digest=sha256:862b3a980c8d8802feac30e7e3c6aace004a592239f97cae8efa4c378fb9ee1f

Observation ee76b518-11cd-480a-baf5-f0688ea3ccfd · outbound

This paper cites Language model beats diffusion-tokenizer is key to visual generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Language model beats diffusion-tokenizer is key to visual generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.842554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.478339Z digest=sha256:21fce46006e3f00bde5acd2e3df883656ff2699c7d7fca8afa388283c45b439a

Observation 57331447-3d86-4891-b60a-9b540cf5a6cb · outbound

This paper cites An image is worth 32 tokens for reconstruction and generation.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE An image is worth 32 tokens for reconstruction and generation

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.827858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.482482Z digest=sha256:6e6f8ad86d423b92fa417a347e2281383f185c9938edf773f1749696ff93ed84

Observation 9de8e88f-e922-4e96-80e6-bd0eb7c5b7c9 · outbound

This paper cites Point-bert: Pre-training 3d point cloud transformers with masked point modeling.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Point-bert: Pre-training 3d point cloud transformers with masked point modeling

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.813443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.486667Z digest=sha256:6be0b47b52751c01f1718cedfef26dae725ec12c2d4ad42c7df8dc481729d112

Observation c8946e61-1d34-480c-8688-d2a6a489a727 · outbound

This paper cites 3dshape2vecset: A 3d shape representation for neu- ral fields and generative diffusion models.ACM Transactions on Graphics (TOG), 42(4):1–16, 2023.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE 3dshape2vecset: A 3d shape representation for neu- ral fields and generative diffusion models.ACM Transactions on Graphics (TOG), 42(4):1–16, 2023

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T12:52:10.490938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:52:10.490938Z digest=sha256:4bf5fb339d6b52899f67354d4a79c7dfab74bbed864d39b0050cea9e31cd7882

Observation 8e2e7596-eb7c-4c0a-80e4-e1a74289d56c · outbound

This paper cites Clay: A controllable large-scale generative model for creat- ing high-quality 3d assets.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Clay: A controllable large-scale generative model for creat- ing high-quality 3d assets

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.788673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.495714Z digest=sha256:eeb4e8ba6f6e680ab43b4117c345b083f8e022e9ae69b944d76ad23e5072130b

Observation 17f8524b-3f64-43b8-9b2d-4656f83fa133 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE The unreasonable effectiveness of deep features as a perceptual metric

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.773583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.499987Z digest=sha256:621afc490e49163a1dee81bda481d9eeb2d93db829cfed9188cadca79f25fbf5

Observation a44ad04c-6f0a-4da6-aab8-b1eebe377d5a · outbound

This paper cites Multi-HeadSelf Attention Multi-HeadCross Attention ×N Scale, Shift Layer Norm CLIPT Text Scale Scale ⊕ ⊕ (a) Transformer Block (Text condition) FFN𝛾!,𝛽!Scale, Shift 𝛼! Layer Norm𝛾.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Multi-HeadSelf Attention Multi-HeadCross Attention ×N Scale, Shift Layer Norm CLIPT Text Scale Scale ⊕ ⊕ (a) Transformer Block (Text condition) FFN𝛾!,𝛽!Scale, Shift 𝛼! Layer Norm𝛾

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:52:10.757885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.504301Z digest=sha256:3092af6a3c4b28938fef6ece8ce0fc2cadefa133cc37cc90ebdf1cb516e55d97

Observation 9cdeec13-9d32-4cab-975a-c39e131d86bd · outbound

This paper cites an unresolved cited work.

SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T12:52:11.037218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T12:52:10.397825Z digest=sha256:e7c4112576e32b29941c89d831d706719214d1ca2b534aabac8c2455ab6f6b3c

Pith citing papers

Observation 76acba24-64d5-4712-8ed8-482cebe95928 · inbound

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models cites this paper.

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:50:14.784325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T14:48:21.787919Z digest=sha256:336abb448e8a467c077f74a3eea19b125ee63aaf18a68a78bba057ff05452329