Pith. sign in

Paper Citation Record · LEDGER

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling

As of 20 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 1 inbound Pith citation observation for arXiv:2411.19492.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19492 v2

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:11:49.782513Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:45:17.233257Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact5
  • verified fuzzy50
  • unresolved44
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c10ed52d-ecbc-492b-a1d2-bede40549c6a · outbound

This paper cites SATR: Zero-shot semantic segmentation of 3D shapes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SATR: Zero-shot semantic segmentation of 3D shapes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.493918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.493918Z digest=sha256:395aaa67d17023686a51e308afb1e6ded4dc821fa2d76d22ea5041f2e706c42f

Observation b0fc3940-b544-4ef7-b7e4-e7d03103bcb2 · outbound

This paper cites SceneCom- plete: Open-world 3D scene completion in complex real world environments for robot manipulation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneCom- plete: Open-world 3D scene completion in complex real world environments for robot manipulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.498994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.498994Z digest=sha256:4e53c4ee49a18ee145334a06d777fea3f49823e192f82a117fc85e1245d7b038

Observation 1f4f9780-9735-48f0-b581-eaa8adef76f4 · outbound

This paper cites Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.502353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.502353Z digest=sha256:51230f25ed2228148835d742900b5865d600360ccba302fb1d556a24abe35bbb

Observation 4569fd82-0eb5-4ccf-8408-1bcf034bcbb4 · outbound

This paper cites Scan2CAD: Learning CAD model alignment in RGB-D scans.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Scan2CAD: Learning CAD model alignment in RGB-D scans

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.505334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.505334Z digest=sha256:e534c320552f5767c80d9b7ff8327f9b2400f7c21ce3224116d72b34d0d93639

Observation e2900653-dc28-4d3e-bff9-dcf7312b8ac8 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Emerg- ing properties in self-supervised vision transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.507805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.507805Z digest=sha256:48f5dbb9251788ae5a390e4fd68673a0af53549675b8e5ecd1a630ac9476e8df

Observation 7da11930-16d1-4549-9077-4548384cca77 · outbound

This paper cites ShapeNet: An Information-Rich 3D Model Repository.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ShapeNet: An Information-Rich 3D Model Repository

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.511213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.511213Z digest=sha256:791bb264b2b9a09f8ab10ad0083f420ae3863d41f83b3e8e0478caa6e12be752

Observation 6fa30b9e-676e-459b-9dad-b82ad1ec18d5 · outbound

This paper cites CLIP2Scene: Towards label-efficient 3D scene un- derstanding by CLIP.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CLIP2Scene: Towards label-efficient 3D scene un- derstanding by CLIP

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.514058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.514058Z digest=sha256:99c2837dcaa1a32ab3c5e562bb34609e481d682ba98d9215d7ed40aff80bc0bf

Observation 3461e84d-5a87-42d0-8f08-242c6964026d · outbound

This paper cites Single-view 3D scene reconstruc- tion with high-fidelity shape and texture.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Single-view 3D scene reconstruc- tion with high-fidelity shape and texture

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.516584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.516584Z digest=sha256:800ee9e59bd52e2952fa49803f36f21a30362dcd803ed0f8a3662f93a7939ef6

Observation d1a26e7b-6693-4eaf-a69e-a43ad1dc52d4 · outbound

This paper cites URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.518832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.518832Z digest=sha256:e550cb82c7ba7143b03939b07781267152b41ac5897fd9a607859dc1503da9ca

Observation c152bf9f-fd1c-4877-9db7-f4ef625b6bb6 · outbound

This paper cites ScanNet: Richly-annotated 3D reconstructions of indoor scenes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ScanNet: Richly-annotated 3D reconstructions of indoor scenes

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.521898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.521898Z digest=sha256:f2d1b7447cb39dde50142ca121461c5fe690143f95ba9148eb33688b1c304691

Observation 6b6e1817-d7fd-48a5-805d-f8c3b396e0e0 · outbound

This paper cites Automated Creation of Digital Cousins for Robust Policy Learning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Automated Creation of Digital Cousins for Robust Policy Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.524207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.524207Z digest=sha256:26a8de864431cf4dfda9ef7086234515b87734bfbee67d7030aeca7193ba007b

Observation 4b986f3b-c3e1-4058-93f5-6e6687f99228 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Objaverse: A universe of annotated 3d objects

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.527276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.527276Z digest=sha256:420c821a39a892a83b51997d5f354e70dc422e6dcd286f9329950805c800faad

Observation 90fe9c9d-96ce-4944-81a5-ffad12978faa · outbound

This paper cites SceneFun3D: Fine-grained functionality and affordance un- derstanding in 3D scenes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneFun3D: Fine-grained functionality and affordance un- derstanding in 3D scenes

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.530240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.530240Z digest=sha256:d1b651d74e50c8765ac54905d85f652712ab333b363e03e75a45819e2c0be672

Observation f034f4e7-ed66-4761-82f8-d719e5bb0389 · outbound

This paper cites PLA: Language-driven open- vocabulary 3D scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PLA: Language-driven open- vocabulary 3D scene understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.532957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.532957Z digest=sha256:36a4968a697421ea3d5c4c3e3df9624422bc0d9d1a6f13da909dc902d843615a

Observation 2ac572d8-4fd1-4b72-b3ac-7d1c6df8ec1c · outbound

This paper cites PanoContext-Former: Panoramic total scene understanding with a transformer.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PanoContext-Former: Panoramic total scene understanding with a transformer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.535892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.535892Z digest=sha256:eaa8ebdcc00ee39df7aad8131f0106c8bada75a5542c9d359aaa7eb24b445fa2

Observation eaaf686a-693b-4f76-8ab8-5e89260c4f29 · outbound

This paper cites CLIP- Away: Harmonizing focused embeddings for removing ob- jects via diffusion models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CLIP- Away: Harmonizing focused embeddings for removing ob- jects via diffusion models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.538456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.538456Z digest=sha256:2ecc99a6cdf63bc6eae596b678476636d74a39e62f21e055726f56ade4b1713e

Observation a4b66764-a5ea-4d09-8c3f-6eb2e3eed339 · outbound

This paper cites Prob- ing the 3D awareness of visual foundation models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Prob- ing the 3D awareness of visual foundation models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.540729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.540729Z digest=sha256:e6f06c10690a54fae8dd8c05f7574d3854ff6e5aace409abd8fb7d7d7c661cf4

Observation 4307cfc7-8e60-4dc0-bb28-44d4669cff56 · outbound

This paper cites A density-based algorithm for discovering clusters in large spatial databases with noise.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling A density-based algorithm for discovering clusters in large spatial databases with noise

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.543346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.543346Z digest=sha256:f860537eba38959a60cfb0af24a26750171222b5229c7ac5aa617ba9d11ef985

Observation 4b65af23-3c8f-4cbc-94a4-b502f79ca123 · outbound

This paper cites Fischler and Robert C.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Fischler and Robert C

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.546454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.546454Z digest=sha256:9af7dd075aaadf2b2448729af5457d66c9b9c78743deddba7eee31d238e06060

Observation 36b5064f-c43d-45c7-b707-0bb01a285626 · outbound

This paper cites Example-based synthesis of 3d object arrangements.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Example-based synthesis of 3d object arrangements

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.549432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.549432Z digest=sha256:aadeefd8e86f6d32b3b8626d3ec746b3e1277a6a9081201cc3c0cf2e1af24b18

Observation 7f63adc0-3367-483c-8376-984dc9be9b26 · outbound

This paper cites 3d-front: 3d furnished rooms with layouts and semantics.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling 3d-front: 3d furnished rooms with layouts and semantics

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.551727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.551727Z digest=sha256:a87d4516e2173a5a9ca7c1052c2989950818911c3bfb78d8410d8c22a1579e41

Observation a4cc9940-54dc-4f87-85a2-32d1c84734fe · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.554635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.554635Z digest=sha256:26127cf78af24158ce74a669b6c9ccd6a646b17f2a785e488f532b4f93e423c1

Observation 4cc51535-0c6e-4a92-87c1-4c6e61dea9bb · outbound

This paper cites Any- home: Open-vocabulary generation of structured and tex- tured 3d homes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Any- home: Open-vocabulary generation of structured and tex- tured 3d homes

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.558068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.558068Z digest=sha256:1f1d5558b30a2f73638a04939c956202314d6be7d35d4e898941a6f8545b2f4d

Observation 11432c41-d44e-4321-9968-19ef40bb787c · outbound

This paper cites DiffCAD: Weakly-supervised probabilistic CAD model retrieval and alignment from an RGB image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DiffCAD: Weakly-supervised probabilistic CAD model retrieval and alignment from an RGB image

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.560446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.560446Z digest=sha256:80a42c67dbb0455c170955cfce461050edb6f3815a1e353be30e5de8dc793447

Observation b99148ee-bb8a-4528-840f-8e326469b597 · outbound

This paper cites GraphDreamer: Compositional 3D scene synthesis from scene graphs.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling GraphDreamer: Compositional 3D scene synthesis from scene graphs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.563920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.563920Z digest=sha256:c7ea5ed6bd86e904ec44417f55dcefa9e76cf6b90157f8c243a6999bd1a02803

Observation aa3bb3ee-95c7-442f-a61e-50183eddfc04 · outbound

This paper cites Zero-shot category-level object pose estimation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Zero-shot category-level object pose estimation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.568274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.568274Z digest=sha256:2d247feb44c9d16369a2e776fd693ebb801fb84cb5c19776f5f6889eb119b830

Observation 44c9488f-8f54-4e50-a6e5-e4f699ea6f0d · outbound

This paper cites ROCA: Robust CAD model retrieval and alignment from a single image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ROCA: Robust CAD model retrieval and alignment from a single image

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.509037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.571479Z digest=sha256:23b1f4fbc0dde597d8e0207b4f60953ee965939ac8c0631107ce555f6c4a5e60

Observation 1389a93c-938b-429c-93de-ea4f091f3eb5 · outbound

This paper cites 3D-LLM: In- jecting the 3D world into large language models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling 3D-LLM: In- jecting the 3D world into large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.501591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.573680Z digest=sha256:be5b4052ed1028a0343cc124f1804d156c3cdd1e9f9e795f173791aa2e44d390

Observation eadedfa4-e170-4b51-94f0-c1486f70a6ed · outbound

This paper cites Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.576177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.576177Z digest=sha256:1eb74ae3fc7f06e11eda4483211020492ec17ba8da0e91ac4cc2003d21e47190

Observation 3914b068-dd94-433a-9877-e4b1564c8537 · outbound

This paper cites Aladdin: Zero-Shot Hallucination of Stylized 3D Assets from Abstract Scene Descriptions.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Aladdin: Zero-Shot Hallucination of Stylized 3D Assets from Abstract Scene Descriptions

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:50.011876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.579347Z digest=sha256:06d2373c2f384554655a618ffd56180917202fdbaf5ab46ca98b682df36dc16e

Observation b1682dc6-5c43-417d-963c-0c57922a5317 · outbound

This paper cites Holistic 3D scene parsing and re- construction from a single RGB image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Holistic 3D scene parsing and re- construction from a single RGB image

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.493482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.582343Z digest=sha256:f820bfa8e36cf73e01f1b6945267254f42cae460400aaa78f4ba1183e670515d

Observation f40c009a-a11c-4b74-b560-43cc7a141fc7 · outbound

This paper cites OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.584612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.584612Z digest=sha256:a8295fc2aa9ae5c06461feb8ec5e0295979a76962cc291572500e9c588684363

Observation e0d13309-f173-4a11-aa4c-20ffb32777b9 · outbound

This paper cites CenterSnap: Single-shot multi-object 3D shape reconstruction and categorical 6D pose and size estimation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CenterSnap: Single-shot multi-object 3D shape reconstruction and categorical 6D pose and size estimation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.485608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.587024Z digest=sha256:e8d37f9deeccf3e1babfb9706c9409572df7d146fb274e114a9713a75ebe5475

Observation 4029bd80-e635-4220-9332-c665f271c800 · outbound

This paper cites an unresolved cited work.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.590378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.590378Z digest=sha256:24f778732418c4041d083f85c60a7ac37c29b76a69c26394527da4051e6abf96

Observation c4df1580-0934-4c3e-ab63-e5cae16d210a · outbound

This paper cites ConceptFusion: Open-set Multimodal 3D Mapping.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ConceptFusion: Open-set Multimodal 3D Mapping

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.592608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.592608Z digest=sha256:0bca642d274ba879c228ea10c4c84e06f0ee1a8f25579651690a3d4f8d44951f

Observation 8d8df2bf-ae13-4266-b896-c28c4e3f334a · outbound

This paper cites SceneVerse: Scaling 3D vision-language learning for grounded scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneVerse: Scaling 3D vision-language learning for grounded scene understanding

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.473781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.595695Z digest=sha256:93556d96637904c456ef95fdf0e045ae51fde4e9bcf8c70eaf049b35215204d3

Observation 9fd86233-c5dd-4801-b59e-9f150ccc453d · outbound

This paper cites LERF: Language embed- ded radiance fields.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling LERF: Language embed- ded radiance fields

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.466992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.598600Z digest=sha256:5c9e54eda1f4dfb3cba40b65582f7eec82e868d0730105c19b492ee04cf946ed

Observation fbf60058-fe52-4163-9c27-40d3938ce0a2 · outbound

This paper cites Habitat synthetic scenes dataset (HSSD-200): An analysis of 3D scene scale and realism tradeoffs for objectgoal naviga- tion.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Habitat synthetic scenes dataset (HSSD-200): An analysis of 3D scene scale and realism tradeoffs for objectgoal naviga- tion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.460093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.600942Z digest=sha256:42f75648a42cecb4007907220da43989e88758883d7de322ef5f65227d739ed3

Observation 78cb8e7f-801f-44dc-aa8b-676e4c882697 · outbound

This paper cites Segment Anything.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Segment Anything

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.604227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.604227Z digest=sha256:e457d975be5d2a979a16a5cf2b34cec2c1346d9b9d03e32113ead42fb4ee47b3

Observation 63ef7ff9-186c-42a7-95a2-83510ec31c31 · outbound

This paper cites Mask2CAD: 3D shape prediction by learning to seg- ment and retrieve.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Mask2CAD: 3D shape prediction by learning to seg- ment and retrieve

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.452832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.607531Z digest=sha256:b8519f9867bb245c83a444911613679fe3ce4cb03b7bc535f2e0c16f80b71b36

Observation b19deefe-7550-41d4-8e8a-3d8bba338e28 · outbound

This paper cites Patch2cad: Patchwise embedding learning for in-the- wild shape retrieval from a single image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Patch2cad: Patchwise embedding learning for in-the- wild shape retrieval from a single image

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.445461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.609993Z digest=sha256:262bb0f40d4039840410f77fbb087e8e61c499e12ce051de1c07c735c2c7fb7a

Observation 49a26a17-6fdb-4f7a-9957-55c070ed88bc · outbound

This paper cites Langer, G.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Langer, G

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.438477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.613122Z digest=sha256:f380a164cea8a845532efb6b22b1eb600cb19fa7f5bde2418f7b061ec93c086f

Observation ad0699e8-db43-43c7-8977-55b4a505aef9 · outbound

This paper cites FastCAD: Real-Time CAD Retrieval and Alignment from Scans and Videos.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling FastCAD: Real-Time CAD Retrieval and Alignment from Scans and Videos

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.617182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.617182Z digest=sha256:c82c9224c61eefebb156c90e659536bdfc541534afeeb03aef279e6267084eb2

Observation 141f1074-31b5-42c3-ad5e-6632d1020d50 · outbound

This paper cites Duoduo CLIP: Efficient 3D Understanding with Multi-View Images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Duoduo CLIP: Efficient 3D Understanding with Multi-View Images

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.619909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.619909Z digest=sha256:8b668414693f6157c51a991b59209347898dfd5cd2da28d85effadeff2a7fe1b

Observation 51b90bbf-2d31-4901-abb9-0b501532ab37 · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.622478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.622478Z digest=sha256:9b3dc0c8c107d191024213f0c627e8754cecc4b6ee30a072ba8697c3f1685e31

Observation 08c3a02c-a967-425e-889e-8e411437f904 · outbound

This paper cites InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.625257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.625257Z digest=sha256:58c631b7ed45d72c01f6d2225a2b750a2edfe04a914ea3c6444071d8a5b948cf

Observation 33e30b14-43ea-407b-af39-8b59d8bfdcd7 · outbound

This paper cites Towards high-fidelity single-view holistic reconstruction of indoor scenes.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Towards high-fidelity single-view holistic reconstruction of indoor scenes

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.430827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.628147Z digest=sha256:ed1e090a7a01b3069222f4f25585be8afa38fce3c633373b0da33235648c27e0

Observation 5afda80b-60e8-4cef-a612-0b847c5398af · outbound

This paper cites LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation Dataset.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation Dataset

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:49.953020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.630597Z digest=sha256:881c318b58d2e1d86926b77131ae9ebe0ff0fbe9300c2f92796f7c007bce4068

Observation 05e95c9d-3b13-421d-b51b-dfc76d29ae6b · outbound

This paper cites PartSLIP: Low-shot part segmentation for 3D point clouds via pretrained image- language models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PartSLIP: Low-shot part segmentation for 3D point clouds via pretrained image- language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.422596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.633616Z digest=sha256:2840c38ca24c9812305e4ce31ce194b9bf9c4126aab1420af420ee586743da6a

Observation a4aa060b-a51c-4e72-ae46-90abf372a5c4 · outbound

This paper cites OpenShape: Scaling up 3D shape representation towards open-world understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenShape: Scaling up 3D shape representation towards open-world understanding

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.415948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.637708Z digest=sha256:034a9fec2ce7b0f98c5b6ebd075be1e99b670fd3a7ba6f3b0f7381d3376554cc

Observation 899abf58-de0f-4be0-92c4-72564e622cab · outbound

This paper cites Open-vocabulary point-cloud object detection without 3D annotation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open-vocabulary point-cloud object detection without 3D annotation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.408637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.640419Z digest=sha256:52b7730163cf4938426c0edf6a88dc20a7b74b452753d40a429122fc5ba90029

Observation 299a2f2d-ca6e-4ec6-a677-6b5b8c355a71 · outbound

This paper cites Vid2CAD: CAD model alignment using multi-view constraints from videos.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Vid2CAD: CAD model alignment using multi-view constraints from videos

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.400234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.642820Z digest=sha256:31145335977cc5be14e29dd984711794a8d88350a841490aac50b71d640ec6d2

Observation 5c630e24-16f8-4a04-8cc6-96b3964b67d9 · outbound

This paper cites Cad-estate: Large-scale cad model annota- tion in rgb videos.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Cad-estate: Large-scale cad model annota- tion in rgb videos

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.393160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.645585Z digest=sha256:8fe36a7ceb76794c30db3aec8294194e87e5474d3cff910d343a9fd59678c2c8

Observation 01ecffa7-6669-4a6e-9d06-1f38217748c9 · outbound

This paper cites Scaling open-vocabulary object detection.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Scaling open-vocabulary object detection

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.386021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.648858Z digest=sha256:1704ff98ae7ecec3109807db245a9015e7ebf08d32ad87a89cb449a5ccbed4d2

Observation aef7d29d-6dd1-4f0a-ad72-4fc34ac6a4bb · outbound

This paper cites Open3DIS: Open-vocabulary 3D instance segmentation with 2D mask guidance.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open3DIS: Open-vocabulary 3D instance segmentation with 2D mask guidance

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.378614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.651350Z digest=sha256:ddf154862fd1c60b48548d5b3aec2a682dd30093a82d54e9bb19197dad28bc21

Observation 2bf060e5-9f48-423b-aa8b-4477f6de6072 · outbound

This paper cites GigaPose: Fast and Robust Novel Ob- ject Pose Estimation via One Correspondence.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling GigaPose: Fast and Robust Novel Ob- ject Pose Estimation via One Correspondence

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.369090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.653661Z digest=sha256:22f2af6ee6291d870fa5323c3e73245b59405ed275b00f9b6512b0ac8a1ac61d

Observation dea198db-3ebb-431c-89e1-4664d7dd54ea · outbound

This paper cites Total3DUnderstanding: Joint layout, object pose and mesh reconstruction for indoor scenes from a single image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Total3DUnderstanding: Joint layout, object pose and mesh reconstruction for indoor scenes from a single image

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.360211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.656294Z digest=sha256:09f47f599733b3ba6edb627ca73855d1afee42a4aee6893065f56a446fc8ca2e

Observation 8e93a9eb-b3e8-4bb6-85ec-f0c01fb757b6 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DINOv2: Learning Robust Visual Features without Supervision

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.658751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.658751Z digest=sha256:141f43b9bce19a7b62b025859321288074b206704d1c52b99146d35f62665e07

Observation 13524e66-3cc2-4a6e-92c3-956d93cbd301 · outbound

This paper cites ATISS: Autore- gressive transformers for indoor scene synthesis.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ATISS: Autore- gressive transformers for indoor scene synthesis

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.351186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.661845Z digest=sha256:f4092b97d1472a1a3b51ad05e223f739ef46e2f2a3d1ddbe0c23b017e46c746d

Observation 264b8f0a-3f84-4350-b2bf-25d1fbaa5973 · outbound

This paper cites OpenScene: 3D scene understanding with open vocabular- ies.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenScene: 3D scene understanding with open vocabular- ies

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.344077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.665002Z digest=sha256:3682501c5e157bd29d6291805c8e270b0807e64373242b7f8be73ef703077eaf

Observation d82d9222-8027-40e7-98db-82eb806064aa · outbound

This paper cites LangSplat: 3D language gaussian splat- ting.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling LangSplat: 3D language gaussian splat- ting

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.336354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.668569Z digest=sha256:b8ceef32df08f400423727d074a8affa03591940323311220d2978fd9b2a535e

Observation 8d9bafd8-7cf4-4c13-8d8a-7564cbd35969 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Learning transferable visual models from natural language supervi- sion

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.671433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.671433Z digest=sha256:17198e2818e931f91e9c41df5989153e6b9260be0f0240ec1ab498b2a0813718

Observation e3a9a5f5-90cc-4ae7-b497-2e2fe0b355e9 · outbound

This paper cites Fast and flex- ible indoor scene synthesis via deep convolutional genera- tive models.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Fast and flex- ible indoor scene synthesis via deep convolutional genera- tive models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.325043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.674137Z digest=sha256:4629e8ecaec843a62ac4eecb80e66bb19b52fc975aa6e7ca174ee8fb0cf0fe30

Observation 787cabb3-1af9-40d0-81f9-3dbcd1c2acb4 · outbound

This paper cites Hypersim: A photorealistic syn- thetic dataset for holistic indoor scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Hypersim: A photorealistic syn- thetic dataset for holistic indoor scene understanding

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.318328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.677118Z digest=sha256:555fd8aa1515fc632fc8e8875eeda4a842929797c5bc353768a3cd30adf25941

Observation 880af085-46f0-4a31-b380-5f96591c318d · outbound

This paper cites Estimating generic 3D room structures from 2D annotations.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Estimating generic 3D room structures from 2D annotations

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.309562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.680491Z digest=sha256:ea5ff19e74f6eb627ac1006fe4ff377c9a48ecde2b93b41dbaf15c0d51274de5

Observation 2676c924-eef1-447c-8041-2becce7e76e2 · outbound

This paper cites Computational geometry.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Computational geometry

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.301546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.683425Z digest=sha256:a9a7e4a566eac9174ba6c5658fd2f1cd04f13483ac715169066870ac8a5ca060

Observation ae21a2f4-160a-4e51-9794-8c91025c928d · outbound

This paper cites PlaneRecTR: Uni- fied query learning for 3D plane recovery from a single view.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PlaneRecTR: Uni- fied query learning for 3D plane recovery from a single view

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.295397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.685824Z digest=sha256:d837f57791cac98cd330f809aa038c40ad1d228d13e9922cf12ae1cac8ef2825

Observation ff8537de-ce15-41b0-82eb-07afa53accbf · outbound

This paper cites General 3D room layout from a single view by render-and-compare.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling General 3D room layout from a single view by render-and-compare

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.287652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.688793Z digest=sha256:2a3830ce142f49e2b2a0bd3cd5c63b0a3715037d47fd929949d94e650760a208

Observation 7d498c06-3dfd-46a6-b560-7a03cd10f95b · outbound

This paper cites Lempitsky.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Lempitsky

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.280447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.691376Z digest=sha256:e01b621240bac24ef0b670327af01618644412c1c5b4c6b0d78c3ac9e8078029

Observation c8d8126c-03d9-4792-a825-8a3a4f38bf44 · outbound

This paper cites Habitat 2.0: Training home assistants to rearrange their habitat.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Habitat 2.0: Training home assistants to rearrange their habitat

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.273410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.693691Z digest=sha256:c9d97a95e5ecb3e53bb79de340874662ff9202e9b7db92b812ec136ff26ee954

Observation 54f70c6a-f0bf-47ed-b740-f7e59808c83c · outbound

This paper cites OpenMask3D: Open-Vocabulary 3D Instance Segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling OpenMask3D: Open-Vocabulary 3D Instance Segmentation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.697493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.697493Z digest=sha256:5d7113252d3d7cace59b91001c0ed240d36ac7e4821f77e374d9ca9ab472dc12

Observation cabec119-1f50-43de-a3f9-688cb1002400 · outbound

This paper cites SceneMotifCoder: Example-driven Visual Program Learning for Generating 3D Object Arrangements.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneMotifCoder: Example-driven Visual Program Learning for Generating 3D Object Arrangements

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.700931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.700931Z digest=sha256:69a121e37fed7a833a95c15af232b4aa522e6f0221a015ec748bfb0de9690f4e

Observation 7d6a11eb-53e0-4c2a-a996-ecd22a6e70e4 · outbound

This paper cites DiffuScene: Denoising diffu- sion models for generative indoor scene synthesis.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DiffuScene: Denoising diffu- sion models for generative indoor scene synthesis

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.266035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.703907Z digest=sha256:b1a4f4d448371a86103e0cfc04997239a5a29ce7dec2046f7f9094a9d4e5d8d6

Observation 41f804fa-f2cc-4271-812a-cb9dc1345bda · outbound

This paper cites Least-squares estimation of transforma- tion parameters between two point patterns.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Least-squares estimation of transforma- tion parameters between two point patterns

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.257927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.707342Z digest=sha256:988cc6c68cb27e698c7a97c470df27d26d4475993210e2efac26a92ba7c88823

Observation aa30b035-e05d-4931-b51e-827bac432aa7 · outbound

This paper cites Deep convolutional priors for indoor scene syn- thesis.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Deep convolutional priors for indoor scene syn- thesis

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.251260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.709763Z digest=sha256:7e016584e3341854e0c57fa2b4324517e5fba6f5ab03d565d16c71a403995ec1

Observation 0815e5db-bc2b-483d-97d8-b42e5d68612d · outbound

This paper cites PlanIT: Planning and in- stantiating indoor scenes with relation graph and spatial prior networks.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling PlanIT: Planning and in- stantiating indoor scenes with relation graph and spatial prior networks

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.244285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.712598Z digest=sha256:3ad22bf2181153a57b1bf47f1cb4ceffcd39b7c8ca6776d6638201d3d53fffc0

Observation 3a365394-dbd9-407f-8cd0-cdf9e83bc041 · outbound

This paper cites Lift3D: Zero-shot lifting of any 2D vi- sion model to 3D.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Lift3D: Zero-shot lifting of any 2D vi- sion model to 3D

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.237503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.715682Z digest=sha256:2d1e6a97b53ef44bf3f4b3887e67fd6a47d018a13a65fa9160e5ae79d861fa43

Observation a4d9314b-7d34-466e-b97b-048b3bb45935 · outbound

This paper cites SceneFormer: Indoor scene generation with transformers.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling SceneFormer: Indoor scene generation with transformers

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.230954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.718055Z digest=sha256:f48945cc35683e2f58e3a3105e26a7fe39cf46acf274686bff08eccd79f8d42b

Observation 5d969a9a-5799-4b47-ad3b-779e1785e78a · outbound

This paper cites Lego-net: Learning regular rearrangements of ob- jects in rooms.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Lego-net: Learning regular rearrangements of ob- jects in rooms

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.224002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.720297Z digest=sha256:bf3d4e14186b2f33b7bf13d805ba36d7536c8419bbde04b317a80b20895169f8

Observation 85a496ba-cd1d-4b62-9b6f-182dd497fc87 · outbound

This paper cites R3ds: Reality-linked 3d scenes for panoramic scene understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling R3ds: Reality-linked 3d scenes for panoramic scene understanding

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.217254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.722620Z digest=sha256:75f56ed147eed0f04e903928370719045e610da2338103c5713ac8a303b20ef9

Observation a7760160-f15f-44c5-b27c-5c3f8fe3d692 · outbound

This paper cites Generalizing single-view 3D shape retrieval to occlu- sions and unseen objects.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Generalizing single-view 3D shape retrieval to occlu- sions and unseen objects

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.209991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.725620Z digest=sha256:98cdaa9b18ddfd2f60b6fb2aad14c26ca9d0de606f979ab47c23aac5334f8fe8

Observation e18b82c3-98e4-4f11-89d0-3f068ae12108 · outbound

This paper cites ULIP-2: Towards scal- able multimodal pre-training for 3D understanding.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ULIP-2: Towards scal- able multimodal pre-training for 3D understanding

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.201800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.728700Z digest=sha256:3718ade3988a263d3fe61162cc1a1cd7ed1e584b9d0237b6f831a443e17bc302

Observation 02d3f13e-c7db-4f0a-82b0-1d96e5be249b · outbound

This paper cites Learning to reconstruct 3d non-cuboid room layout from a single rgb image.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Learning to reconstruct 3d non-cuboid room layout from a single rgb image

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.194470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.731710Z digest=sha256:290bc54fcb6c04313a1a6d2b607c625036b123673b395de158e1f4c37df33d51

Observation d06723d3-9339-40d2-ad98-66f40cf7dde5 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.734860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.734860Z digest=sha256:b7624cff38b1b64cf963da5cb26649ce5b6309411776cf55ee524b678cbc30ad

Observation 11ce0f39-aab4-4e95-a6de-e8952184a690 · outbound

This paper cites Depth Anything V2.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Depth Anything V2

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.737583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.737583Z digest=sha256:55f9da7f83b4521d1d62fce7efebd9a4f47c1c641d7e35a59b98d88bc2c4e7b0

Observation dd72d8e8-af42-4cde-aa27-6e12c9c63ac3 · outbound

This paper cites ImOV3D: Learning Open-Vocabulary Point Clouds 3D Object Detection from Only 2D Images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling ImOV3D: Learning Open-Vocabulary Point Clouds 3D Object Detection from Only 2D Images

Reference 86

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:49.916053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.740308Z digest=sha256:5ca6c6777fb5459dd9aae6ca230554b5d74296a502dd989160b1931776755a26

Observation ae042157-ac9d-4eac-a7dd-24e250697efa · outbound

This paper cites Holodeck: Language guided gen- eration of 3D embodied AI environments.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Holodeck: Language guided gen- eration of 3D embodied AI environments

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.186658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.743249Z digest=sha256:8b8b876684e7a8fa45c64a11b830247d263c0cb47b9ddf42f0f5a19794380bde

Observation cf604c72-7298-4844-9d0f-897e80fac004 · outbound

This paper cites Multi-view Aggregation Network for Dichotomous Image Segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Multi-view Aggregation Network for Dichotomous Image Segmentation

Reference 88

Resolution
verified exact
local_arxiv, observed 2026-08-12T10:11:49.905489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.745506Z digest=sha256:27b25296686258f0cbbbd028e323d50ab86cf207057ecbbdfeab78f140c5aaf5

Observation d35df8f3-c6fb-426b-86db-10310aa4bad0 · outbound

This paper cites Inpaint Anything: Segment Anything Meets Image Inpainting.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Inpaint Anything: Segment Anything Meets Image Inpainting

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.748866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.748866Z digest=sha256:133f8098994835ada2a09ca35c8bb9fb45840bfeb4d15c4514e2dff9be18dc7d

Observation 096e0f4b-243d-4d72-a29f-042671238d5e · outbound

This paper cites Improving 2D feature representations by 3D-aware fine-tuning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Improving 2D feature representations by 3D-aware fine-tuning

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.178196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.751681Z digest=sha256:c5a65c15611980ab0aa2148719f3b207c46b8b6c4963848278f94ef186d7a76f

Observation c9c5a416-f891-4c26-9f76-9ce0382ff57d · outbound

This paper cites DeepPanoCon- text: Panoramic 3D scene understanding with holistic scene context graph and relation-based optimization.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling DeepPanoCon- text: Panoramic 3D scene understanding with holistic scene context graph and relation-based optimization

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.170172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.754700Z digest=sha256:e9e1ba7c5e6cc50e5a52b2f99026bf3cabe25eec103176cf0109540fda3a898c

Observation 8bb5fafc-8cf4-4e2e-939d-5b266b049bb5 · outbound

This paper cites CLIP-FO3D: Learning free open-world 3D scene representations from 2D dense CLIP.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling CLIP-FO3D: Learning free open-world 3D scene representations from 2D dense CLIP

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.162667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.757458Z digest=sha256:d13f6a3b03e7796d15b44e48612d776db33a599529626ed07adf7bab759846b0

Observation 5a632b2e-c84f-44a5-b1b9-8c367523e5e8 · outbound

This paper cites Structured3D: A large photo-realistic dataset for structured 3D modeling.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Structured3D: A large photo-realistic dataset for structured 3D modeling

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.155290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.759672Z digest=sha256:7924e48401e5d7e3c85a946adcd477f4b5c25b87082493caeb2a36240e5eab02

Observation 8dc8a302-6b17-4eb6-a682-4dddcdc1ae71 · outbound

This paper cites Bilateral refer- 12 ence for high-resolution dichotomous image segmentation.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Bilateral refer- 12 ence for high-resolution dichotomous image segmentation

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.147406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.761875Z digest=sha256:1e20be264186d0c34930494d5b2a9971ccbf5e8b757f00eb54f2d3ba8e90648e

Observation ff514cac-978c-4aae-a9e8-ff68e3efb5e8 · outbound

This paper cites Zero-Shot Scene Reconstruction from Single Images with Deep Prior Assembly.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Zero-Shot Scene Reconstruction from Single Images with Deep Prior Assembly

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.764963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.764963Z digest=sha256:001f5761841e2c30f1068c70cd5eaef0b92526a2c4d054256ae31c6dbcdf3579

Observation 601a6c67-5a14-4e0a-b27a-c02bae60cec9 · outbound

This paper cites Open3D: A Modern Library for 3D Data Processing.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Open3D: A Modern Library for 3D Data Processing

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T10:11:49.769541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:11:49.769541Z digest=sha256:91dbe2c20e71d04c0ca78393b77fe1697216fae1ede55f8fb848ee74b09ffc3b

Observation 0326cd5a-3ce0-4e48-8e58-c24586f365e0 · outbound

This paper cites Point- CLIP v2: Prompting CLIP and GPT for powerful 3D open- world learning.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling Point- CLIP v2: Prompting CLIP and GPT for powerful 3D open- world learning

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.138459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.772565Z digest=sha256:77bfa0b78801755d2311dc23e4d8d5363690497c4413483169a58f75e8f6ed53

Observation 843593dc-c620-4229-a2c7-860359cbced8 · outbound

This paper cites GRS: Generating robotic simulation tasks from real-world images.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling GRS: Generating robotic simulation tasks from real-world images

Reference 98

Resolution
verified exact
raw_fallback, observed 2026-08-12T10:11:49.871719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.775705Z digest=sha256:bd68ab4e0de5cb1cb1f522c02d25b46282b4cfd8e6139aa6b0064028ac1c351f

Observation ae5943dc-7d99-4022-ab7e-5ce3bce352b7 · outbound

This paper cites gpt-4o-2024-08-06.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling gpt-4o-2024-08-06

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:11:50.128905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.778819Z digest=sha256:8fcdb8a039857bf244a78c5484aaa85f1b9163d35612d5d016bb4008661bd051

Observation 775dd22d-dda3-4b03-8033-b0e597ed2b6c · outbound

This paper cites a photo of CLASS.

Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling a photo of CLASS

Reference 100

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T10:11:50.120553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-12T10:11:49.782513Z digest=sha256:4cf480330e74e56f2831cf8dd6194e586dcf21c9c1c89a1f36543f117a6bee70

Pith citing papers

Observation 946354f0-d8de-47df-80b8-e47a1d45a800 · inbound

Advances in 4D Representation: Geometry, Motion, and Interaction cites this paper.

Advances in 4D Representation: Geometry, Motion, and Interaction Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling

Reference 299

Resolution
unresolved
no resolver link, observed 2026-08-04T08:45:17.233257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:45:17.233257Z digest=sha256:6ca555acd89d6416ef9f3b959bcfd2be5750bf9fc53c9d32056726d72ef1cb4d