Pith. sign in

Paper Citation Record · LEDGER

AI2-THOR: An Interactive 3D Environment for Visual AI

As of 20 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 100 inbound Pith citation observations for arXiv:1712.05474.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1712.05474 v4

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-12T05:24:28.744397Z

measured 148 of 148 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 100 of 220 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:33:00.779254Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy47
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

327
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation be93fe07-7e0c-4b92-aad3-90c2e6b31025 · outbound

This paper cites Robothor: An open simulation-to-real embodied ai platform.

AI2-THOR: An Interactive 3D Environment for Visual AI Robothor: An open simulation-to-real embodied ai platform

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:28.813151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:22ecfe8fa69565cb36e78b5352f4e980b13d5704c56e473dace889b42014523e

Observation 29c10b23-84ac-45a8-8a9e-a4a67b4d4595 · outbound

This paper cites Procthor: Large-scale embodied ai using procedural generation.

AI2-THOR: An Interactive 3D Environment for Visual AI Procthor: Large-scale embodied ai using procedural generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:28.855357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:e58229dabea969875b763818fd28da6885de74f408e7c4df00718f1115a5f5da

Observation 16efcf7f-82a5-4dc4-b218-b3689b89b6ec · outbound

This paper cites Learning object relation graph and tentative policy for visual navigation.

AI2-THOR: An Interactive 3D Environment for Visual AI Learning object relation graph and tentative policy for visual navigation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:28.927371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:b80ec1d88099defa7deb6a14698a6b162a9d1c6657edccab5adf7065aa26c010

Observation fc325d26-3fb5-4a79-bc37-a96a156f60d5 · outbound

This paper cites What do navigation agents learn about their environment? In CVPR.

AI2-THOR: An Interactive 3D Environment for Visual AI What do navigation agents learn about their environment? In CVPR

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.037718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:9065903b4d682dcd300e1d268bd74efaeb0c7f768378c629bae8f8ee633d3697

Observation 9472ad5d-ef99-48d2-86a8-18ce65deb386 · outbound

This paper cites Manipulathor: A framework for visual object manipulation.

AI2-THOR: An Interactive 3D Environment for Visual AI Manipulathor: A framework for visual object manipulation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.102653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:c7bdf5951902e0c1fb87c5380d918a7cea0f04512c2171089fd8b180a5e57478

Observation e8cd9a68-9d1b-4ae5-9cae-ccbce1c6674f · outbound

This paper cites Segan: Segmenting and generating the invisible.

AI2-THOR: An Interactive 3D Environment for Visual AI Segan: Segmenting and generating the invisible

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.108218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:1d83ba5fb7b16b368dce9e8315da2ef961ca0b77f83ddbc0d4b0637e5c1e650c

Observation 1729c8a9-d1b8-4271-9aed-aa993485f06f · outbound

This paper cites Threedworld: A platform for interactive multi-modal physical simulation.

AI2-THOR: An Interactive 3D Environment for Visual AI Threedworld: A platform for interactive multi-modal physical simulation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.197206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:fb7295d0d92b4d978c99606940c897b8be0767fe9259c669441a826eafc07def

Observation f23cdb4d-92b9-4d2e-a378-54690a22e685 · outbound

This paper cites Look, listen, and act: Towards audio- visual embodied navigation.

AI2-THOR: An Interactive 3D Environment for Visual AI Look, listen, and act: Towards audio- visual embodied navigation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.266722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:c7388b1372455aba26fdbb292834919da36687e311b4f633449ed76a5fa72338

Observation 60849c81-d144-4ab4-b742-158c0ada3cbd · outbound

This paper cites Dialfred: Dialogue- enabled agents for embodied instruction following.

AI2-THOR: An Interactive 3D Environment for Visual AI Dialfred: Dialogue- enabled agents for embodied instruction following

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.337206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:fa3122efa4b802db098bb1790f78ad8e28b1686ee4c9bb86353e2cd463f1f570

Observation 2a6a60e8-ad26-4d56-a02f-6575306b9d94 · outbound

This paper cites Iqa: Visual question answering in interactive environments.

AI2-THOR: An Interactive 3D Environment for Visual AI Iqa: Visual question answering in interactive environments

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.416575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:9b378d76a9fb3319e9cecb2bf2f11d9f164e03245f9d9b9f5f65cef430b615ae

Observation 513149e5-8c53-4dcf-9cda-82479efa9268 · outbound

This paper cites an unresolved cited work.

AI2-THOR: An Interactive 3D Environment for Visual AI Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-12T05:24:29.442582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:de45d0232843fd4f2c670de36b087e9efcff316f93507be4fe9df8439fa979a8

Observation 6a1485bf-7da3-48c6-955b-606d75d2e715 · outbound

This paper cites Schwing, and Aniruddha Kembhavi.

AI2-THOR: An Interactive 3D Environment for Visual AI Schwing, and Aniruddha Kembhavi

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.446751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:75d7dd7c8d5d6d04034615539a320199b4d6f00eb4913e57922e636f4de30b2e

Observation 14129ea1-51f5-45b4-b593-f3f175fc4d3e · outbound

This paper cites Learning adaptive language interfaces through decomposition.

AI2-THOR: An Interactive 3D Environment for Visual AI Learning adaptive language interfaces through decomposition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.486904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:0e3a612b1fbac8d318527933b903ccde639f355591db3ed0b20d4f2fd5d553c0

Observation 4b30b6c7-3d07-404b-8749-6aaff67a3a3d · outbound

This paper cites The design of stretch: A compact, lightweight mobile manipulator for indoor human environments.

AI2-THOR: An Interactive 3D Environment for Visual AI The design of stretch: A compact, lightweight mobile manipulator for indoor human environments

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.526641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:8c369d2a8b342e2219620cf24f16bb687533d62e9622ac0e2a031ad6463e608c

Observation a3158ec2-1665-48f8-acbb-2817333f1497 · outbound

This paper cites Simple but effective: Clip embeddings for embodied ai.

AI2-THOR: An Interactive 3D Environment for Visual AI Simple but effective: Clip embeddings for embodied ai

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.556583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:f00cb51b7a62a8c9730000ba7258a011c88fb92ad2ddcd86fb0a8c8a90bff07f

Observation 102f82cd-c0f6-4050-a31b-c08489b5756b · outbound

This paper cites Contrasting contrastive self- supervised representation learning pipelines.

AI2-THOR: An Interactive 3D Environment for Visual AI Contrasting contrastive self- supervised representation learning pipelines

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.673893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:03369f87559e8b3a7ea8ffc01c1a6035496a25b2baec23540802272603969e2f

Observation 795e5939-e120-4d0f-b1c4-c6800358363b · outbound

This paper cites Interactron: Embodied adaptive object detection.

AI2-THOR: An Interactive 3D Environment for Visual AI Interactron: Embodied adaptive object detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.736720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:d2a04886eaf788d157e302388409aafa84f5c47eced4bb1a71ebd1ee78b0b2f6

Observation 78cdc9c7-5396-4e1f-b006-4979963a9b66 · outbound

This paper cites igibson 2.0: Object-centric simulation for robot learning of everyday household tasks.

AI2-THOR: An Interactive 3D Environment for Visual AI igibson 2.0: Object-centric simulation for robot learning of everyday household tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.854261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:97a865a6110f25b6b09b252a50bfb29df1cca8691c7c386e5703056aeb737293

Observation 436ccb82-c447-401f-8ada-23e73f11df7e · outbound

This paper cites Ifr-explore: Learning inter-object functional relationships in 3d indoor scenes.

AI2-THOR: An Interactive 3D Environment for Visual AI Ifr-explore: Learning inter-object functional relationships in 3d indoor scenes

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:29.986655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:6ab7085be5c04bf31f5d1bbb5fd97028b6ed428c7ac6937bd43d2c4e1162b9bc

Observation 9c92eb2a-a633-4020-93ec-0049fe33281a · outbound

This paper cites Multi-agent embodied visual semantic navigation with scene prior knowledge.

AI2-THOR: An Interactive 3D Environment for Visual AI Multi-agent embodied visual semantic navigation with scene prior knowledge

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.097129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:f3ccad88c1008f0fd4be6987084f7806365451e857a79b41bbdeded0af90ba84

Observation 23cae2b9-6ec8-41d5-8788-b99f93e3a4d3 · outbound

This paper cites Learning about objects by learning to interact with them.

AI2-THOR: An Interactive 3D Environment for Visual AI Learning about objects by learning to interact with them

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.110601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:3d4c4c5d86924fc34c62d60e9cdfa184b0cef3c7ba9b53304b4f660e7f5c22dc

Observation e52dc120-e6c3-4f7d-8421-c96c2b453504 · outbound

This paper cites Mgrl: Graph neural network based inference in a markov network with reinforcement learning for visual navigation.

AI2-THOR: An Interactive 3D Environment for Visual AI Mgrl: Graph neural network based inference in a markov network with reinforcement learning for visual navigation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.176745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:f04b1d2baec8b7da334f4aa1a18c797b4e22319ca342d7fc3a366497d1c63407

Observation 2215ee78-333b-4160-8b68-bf82433a1625 · outbound

This paper cites Film: Following instructions in language with modular methods.

AI2-THOR: An Interactive 3D Environment for Visual AI Film: Following instructions in language with modular methods

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.276690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:20e7d55fa9ce41ff9825c24409713a11acbf0824945d135355fc8e66aa59cc68

Observation fc9fe8f7-6bf9-4625-85cc-d7f1ea29644d · outbound

This paper cites Pyrobot: An open-source robotics framework for research and benchmarking.

AI2-THOR: An Interactive 3D Environment for Visual AI Pyrobot: An open-source robotics framework for research and benchmarking

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.406705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:e36c973b3eeb00715d823a5ac4ef95c0ca294d76f13960e1318a75450fb0fd91

Observation dc00aefc-8524-406b-a58e-626691f56b1b · outbound

This paper cites Learning affordance landscapes for interaction exploration in 3d environments.

AI2-THOR: An Interactive 3D Environment for Visual AI Learning affordance landscapes for interaction exploration in 3d environments

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.527063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:0adf98b2e421cd3d9d184afa0132cbebba238e1308e09efc3e7d1ddde7d1b29f

Observation fae9c8c5-9ef0-4e5a-b551-9345d0f068b1 · outbound

This paper cites Shaping embodied agent behavior with activity-context priors from egocentric video.

AI2-THOR: An Interactive 3D Environment for Visual AI Shaping embodied agent behavior with activity-context priors from egocentric video

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.606632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:21d65a91546b829dd4d5614bf592d1ee2e0c24bd30b01bc3d906c3b211bfc3e9

Observation bda267a4-ccb9-4ec2-b168-de00e6c5b736 · outbound

This paper cites Teach: Task-driven embodied agents that chat.

AI2-THOR: An Interactive 3D Environment for Visual AI Teach: Task-driven embodied agents that chat

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.717375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:03aee0d085da7dde30ddf5c40c29313e9d3dfc7f5c5c02672cc2d26f3e756687

Observation 132bd50d-1d55-4ff1-bfcd-412dfcf1438a · outbound

This paper cites Episodic transformer for vision-and-language navigation.

AI2-THOR: An Interactive 3D Environment for Visual AI Episodic transformer for vision-and-language navigation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.820777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:4c9bb2f9ce4b401ffb39bc9448f2fe1c8ff5536aab48a25e75961d9a631a1504

Observation a68c71d1-16f0-4c10-b0eb-833db1031528 · outbound

This paper cites Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI.

AI2-THOR: An Interactive 3D Environment for Visual AI Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.827895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:88318d60adc9c35713c722f3cd23ee75ce7dc56adeb9af41ae29cb678c18ff7a

Observation 8f381d19-49f9-4664-8946-71e44f9be417 · outbound

This paper cites Habitat: A platform for embodied ai research.

AI2-THOR: An Interactive 3D Environment for Visual AI Habitat: A platform for embodied ai research

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.897086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:64a696ebb863e9f7d3e3ca57100cac5be9c4aa640869950a40b46254f37cc0c6

Observation 8d2972a4-0c46-40e9-a9ce-ed814a55a815 · outbound

This paper cites Alfred: A benchmark for interpreting grounded instructions for everyday tasks.

AI2-THOR: An Interactive 3D Environment for Visual AI Alfred: A benchmark for interpreting grounded instructions for everyday tasks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.901394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:8b496f06b00c57a7b8f4980f13a0b0854aa0c54a78b38dccdf5d62b49bdbc254

Observation dabb1184-7b90-487d-b6ee-62a554a985c9 · outbound

This paper cites Chang, Zsolt Kira, Vladlen Koltun, Jitendra Malik, Manolis Savva, and Dhruv Batra.

AI2-THOR: An Interactive 3D Environment for Visual AI Chang, Zsolt Kira, Vladlen Koltun, Jitendra Malik, Manolis Savva, and Dhruv Batra

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.908522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:5703b02b1ae7842650d9e2f56a6eb3205662d92356a0f0df36bdc5fe20475ee5

Observation 8872e4fb-1201-423f-a0be-523dc5976758 · outbound

This paper cites Multi-agent embodied question answering in interactive environments.

AI2-THOR: An Interactive 3D Environment for Visual AI Multi-agent embodied question answering in interactive environments

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.912288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:b64909a1ebc8d87cc3dc4613261fd9b49e937ff5d34613edefc2021bfc542aa7

Observation b8d82bab-0d17-4822-b733-5737c11154b4 · outbound

This paper cites Visual room rearrangement.

AI2-THOR: An Interactive 3D Environment for Visual AI Visual room rearrangement

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.919459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:51a4f7d94b3f95cb2fc320ca64e6c808b2b6458b152bf8812913630e96b0fcee

Observation 2dc2e787-58cb-4eb0-a899-9d0e93d00e85 · outbound

This paper cites Learning generalizable visual representations via interactive gameplay.

AI2-THOR: An Interactive 3D Environment for Visual AI Learning generalizable visual representations via interactive gameplay

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.947120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:158a4c365e9d1cdedf4aef863c637f24ba2321dcddf3c6d93e2e7a2b00667fb7

Observation 7873169d-2d9c-4c77-bcab-2dd8fb327af7 · outbound

This paper cites Allenact: A framework for embodied AI research.

AI2-THOR: An Interactive 3D Environment for Visual AI Allenact: A framework for embodied AI research

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.960632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:e3bc9fdfaac0960ddeedff16b61e84544d8787c6ef970c2060d89e66acb32354

Observation 7676895c-8847-49b4-abb3-111c6af347c0 · outbound

This paper cites Learning to learn how to learn: Self-adaptive visual navigation using meta-learning.

AI2-THOR: An Interactive 3D Environment for Visual AI Learning to learn how to learn: Self-adaptive visual navigation using meta-learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.970543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:695b70a6c88ffa4b12256be1f382a5896032ccdbdde56dbf54453b7ec62b5c89

Observation b6ac0b0b-0f46-44b6-8a36-91b882ece799 · outbound

This paper cites Communicative learning with natural gestures for embodied navigation agents with human-in-the-scene.

AI2-THOR: An Interactive 3D Environment for Visual AI Communicative learning with natural gestures for embodied navigation agents with human-in-the-scene

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.976336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:7c6a48cba39f7f385061778f6e792e5bd04a23a14a9559e6a26e43d0d0e84129

Observation 8c957aa7-e28b-4a15-b243-2c1e60b86738 · outbound

This paper cites Chang, Leonidas J.

AI2-THOR: An Interactive 3D Environment for Visual AI Chang, Leonidas J

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:30.982338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:c7a4e07168d8d2fcfe50be07937d466b2a7e6e83990a1c4294a720f40a88d043

Observation 7471b744-72ba-41e6-90fe-925c83924597 · outbound

This paper cites Visual semantic navigation using scene priors.

AI2-THOR: An Interactive 3D Environment for Visual AI Visual semantic navigation using scene priors

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.003221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:deb22a6f1d406ed79a8d52e765dd1c5ca0eaf1117082232fddf597cd2c5d4c39

Observation 99b83760-c71b-4208-93ba-d7917bba777e · outbound

This paper cites Peters, Roozbeh Mottaghi, Aniruddha Kembhavi, Ali Farhadi, and Yejin Choi.

AI2-THOR: An Interactive 3D Environment for Visual AI Peters, Roozbeh Mottaghi, Aniruddha Kembhavi, Ali Farhadi, and Yejin Choi

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.007942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:7e6339c5553959240ec513cafde7853b3b0e41f14c9e94a7f1a471ab23854223

Observation b634326b-4c13-4f70-9c9d-1a170d07d0cc · outbound

This paper cites Visual reaction: Learning to play catch with your drone.

AI2-THOR: An Interactive 3D Environment for Visual AI Visual reaction: Learning to play catch with your drone

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.012106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:aa64b71825872535f9b386800b1b8ea84b8836907e959dbd40509074c2befcd4

Observation e655a536-27b4-4ff8-bc86-fa2c33f12b1f · outbound

This paper cites Lumi- nous: Indoor scene generation for embodied ai challenges.

AI2-THOR: An Interactive 3D Environment for Visual AI Lumi- nous: Indoor scene generation for embodied ai challenges

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.023334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:eab561b5522d9789e497f7d6f26b8123845cff12454014891dcbd3cea5426b3a

Observation da27a9b8-8f86-4f00-bb29-eb806ba093f5 · outbound

This paper cites Towards optimal correlational object search.

AI2-THOR: An Interactive 3D Environment for Visual AI Towards optimal correlational object search

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.036352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:34562a357004c3b9f5a6caa1e5ff730a5a22f3523bd3ab40ecb45178e41961f4

Observation 3fcde670-87dd-464c-9696-2b380e82dc43 · outbound

This paper cites Target-driven visual navigation in indoor scenes using deep reinforcement learning.

AI2-THOR: An Interactive 3D Environment for Visual AI Target-driven visual navigation in indoor scenes using deep reinforcement learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.047618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:26c98863d9e07e47214a2150df9ffe1b8b6b1019410751170c54c04dea986c4c

Observation 74fe8c3c-1672-45a1-b684-dd0efff26d30 · outbound

This paper cites AI2-THOR supports many different types of agents, including the Ma- nipulaTHOR, Abstract, and LoCoBot agents.

AI2-THOR: An Interactive 3D Environment for Visual AI AI2-THOR supports many different types of agents, including the Ma- nipulaTHOR, Abstract, and LoCoBot agents

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.054345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:116dd99b9c2bd68969eee9fd9be1dac16bb51240262382fa7942d6b82ba521ad

Observation 1fdc1b5a-c3d7-4294-ac55-7d03af181ac2 · outbound

This paper cites AI2-THOR.

AI2-THOR: An Interactive 3D Environment for Visual AI AI2-THOR

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.067946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:ee8289d5fd75f16f5d0cbcdabeb344ff0a9583545662d9280ddcbb73998bad92

Observation 34bb5d04-6765-4478-b729-3e313fa17c6c · outbound

This paper cites These factors include: (a) Model forward pass when computing agent rollouts.

AI2-THOR: An Interactive 3D Environment for Visual AI These factors include: (a) Model forward pass when computing agent rollouts

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T05:24:31.080353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T05:24:28.744397Z digest=sha256:ba7a75d098d665647421e381a812656f0ad58def37c6455a2ee2c00fa3127ca4

Pith citing papers

Observation bfe0c636-9ec7-4b4b-86f6-43ae0b28bf27 · inbound

On Evaluation of Embodied Navigation Agents cites this paper.

On Evaluation of Embodied Navigation Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-13T22:41:57.335040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T22:41:57.280449Z digest=sha256:e89b0f34de975b54596ca85052e216ceb90db66f9295c3b3e3c8b47fecab987b

Observation 2690d810-9cb8-449e-afe1-d3651b1b0b02 · inbound

Learning Grasp Affordance Reasoning through Semantic Relations cites this paper.

Learning Grasp Affordance Reasoning through Semantic Relations AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-25T17:31:06.190816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-25T17:28:43.255094Z digest=sha256:1e46ba98b6cc6986e3e6efe4975242112694a91df183ee5a8b763471d068e286

Observation 22eba7d0-8e09-4f3d-b5dd-7eba458188d0 · inbound

CraftAssist: A Framework for Dialogue-enabled Interactive Agents cites this paper.

CraftAssist: A Framework for Dialogue-enabled Interactive Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-24T19:09:50.152441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-24T19:09:00.377577Z digest=sha256:6be7f9bc3aa5ea190ce1d40099c877abc035600424e408f43a2f50b206e1a227

Observation a19ba7ae-3005-47be-a9e8-c4ddae5cc0a6 · inbound

Why Build an Assistant in Minecraft? cites this paper.

Why Build an Assistant in Minecraft? AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-24T18:24:48.477480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-24T18:20:58.081505Z digest=sha256:54c49649b3bc61461368fb6a712f91a7b78a99acf5967760d515c9e2c361ad9c

Observation c35da86b-e42e-497d-9f20-2ca7a81bc6a2 · inbound

Walking with MIND: Mental Imagery eNhanceD Embodied QA cites this paper.

Walking with MIND: Mental Imagery eNhanceD Embodied QA AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T15:15:36.801423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T15:15:36.801423Z digest=sha256:fd54eb8fe5c04ed4e1224f83fbfa0a28cfe236bdb28e191a7b709ddaf15027b2

Observation 004cd057-1325-4ba0-ab44-658b0645589f · inbound

VUSFA:Variational Universal Successor Features Approximator to Improve Transfer DRL for Target Driven Visual Navigation cites this paper.

VUSFA:Variational Universal Successor Features Approximator to Improve Transfer DRL for Target Driven Visual Navigation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-14T12:51:25.759225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:51:25.759225Z digest=sha256:3e321443502db55c34a92cb8375b2c7781d6e2c9ac94efc093fcd773cb5fb7eb

Observation 6eac0cdc-1adf-41fc-a40b-719ebe0bb742 · inbound

Domain Randomization and Pyramid Consistency: Simulation-to-Real Generalization without Accessing Target Domain Data cites this paper.

Domain Randomization and Pyramid Consistency: Simulation-to-Real Generalization without Accessing Target Domain Data AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T05:39:19.817301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T05:39:19.817301Z digest=sha256:06f39ba93f89d0f9a378028c6343813c1f1fdb1aa50d92a7695e5c59540c7177

Observation a30eda51-1b8f-4382-99b5-d12b335e7629 · inbound

Counterfactual Depth from a Single RGB Image cites this paper.

Counterfactual Depth from a Single RGB Image AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T05:39:19.364744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T05:39:19.364744Z digest=sha256:878609a6a778332e74625e76e4d89f0235dd8fee1ea394b3cc6d6ae35e5cae10

Observation 43b838bb-180f-4c4f-9298-ba30687ead45 · inbound

robosuite: A Modular Simulation Framework and Benchmark for Robot Learning cites this paper.

robosuite: A Modular Simulation Framework and Benchmark for Robot Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-12T22:37:13.462593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T22:37:13.401815Z digest=sha256:b0375e590bb0587a51b2c870cb0cee398783d4ff91ef19c34fd0697a10a989b8

Observation ae27d77f-82d5-41a7-b0ef-640e9442abdf · inbound

Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI cites this paper.

Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-14T18:24:58.031279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T18:24:57.956486Z digest=sha256:2767288aff238400bcf565f69dc034f188e8b1c6b86cd9f9b8aae5cb674460dc

Observation 05c84b09-1b43-4ebb-910d-b931dca8bee3 · inbound

Scaling Robot Learning with Semantically Imagined Experience cites this paper.

Scaling Robot Learning with Semantically Imagined Experience AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-17T18:59:10.408799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T18:59:10.352342Z digest=sha256:dffd856effa1d3394267f0ccb9aebbcb3870c9d019f3e1ba33d26f63a4ae806b

Observation d3981621-f233-4cb0-ab0b-b4e6d3e45482 · inbound

Voyager: An Open-Ended Embodied Agent with Large Language Models cites this paper.

Voyager: An Open-Ended Embodied Agent with Large Language Models AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:24:31.084211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T13:11:40.995345Z digest=sha256:777f8356ddffbd4ffb942a05682ee81ed79f57bd80ea8db153e41910ad8333b5

Observation fcd9de80-60b2-49a3-b441-02f5aff29195 · inbound

Agent AI: Surveying the Horizons of Multimodal Interaction cites this paper.

Agent AI: Surveying the Horizons of Multimodal Interaction AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T14:25:59.189438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T14:25:58.876978Z digest=sha256:f19312dce4a39845c1b9e3f39326506dea7b369ac845c60fa4c82a31cf3098bc

Observation 74d7dfa0-0376-41a5-99aa-105225633984 · inbound

A Survey on Vision-Language-Action Models for Embodied AI cites this paper.

A Survey on Vision-Language-Action Models for Embodied AI AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 176

Resolution
verified exact
local_arxiv, observed 2026-05-24T01:25:54.317439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-24T01:25:10.150459Z digest=sha256:c10574c49c7dcfea19719f81593387d250a1cfced9181f62a1c88427b9093fbf

Observation 239b312b-1165-4ec4-b1eb-5a1d79f2d5d6 · inbound

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots cites this paper.

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-12T23:46:30.274608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T23:46:30.198761Z digest=sha256:f68663ba62c7d83ac8285c4cf724aa43d41487639182caf9c8ecc9a89c4532b6

Observation e17293a4-93aa-42b5-8cf0-622d757b3423 · inbound

ClevrSkills: Compositional Language and Visual Reasoning in Robotics cites this paper.

ClevrSkills: Compositional Language and Visual Reasoning in Robotics AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T21:10:46.183052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:10:46.183052Z digest=sha256:21c032fb5d95fbe7c414784483c694612147cf7f3471c8f53b916ad7959b797d

Observation bc355c25-811d-4d28-9f59-6bfd8e3e914b · inbound

I Can Tell What I am Doing: Toward Real-World Natural Language Grounding of Robot Experiences cites this paper.

I Can Tell What I am Doing: Toward Real-World Natural Language Grounding of Robot Experiences AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T17:06:34.166406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:06:34.166406Z digest=sha256:b5aee83a24028841c27c283a991817d96f49e4939f76b987546e5f71d12ab344

Observation ddc1ef5c-9df6-4241-a33f-432e321158ad · inbound

ROOT: VLM based System for Indoor Scene Understanding and Beyond cites this paper.

ROOT: VLM based System for Indoor Scene Understanding and Beyond AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:02:18.629784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:02:18.629784Z digest=sha256:3a2b11f64465193faf5c75bdcd3788a725c0e6d2816029af66fc1c6c4aa42b03

Observation 1495f859-8e6e-482d-9c80-cc8ae08a62bb · inbound

CityWalker: Learning Embodied Urban Navigation from Web-Scale Videos cites this paper.

CityWalker: Learning Embodied Urban Navigation from Web-Scale Videos AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T11:52:48.074549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:52:48.074549Z digest=sha256:fac49ebd451206bf4f151f737b6f1b29fcff05da76f54f08f80387f047197a2e

Observation 51bb2fdd-3d6a-4871-a4e7-7c4ded61b03b · inbound

Online Knowledge Integration for 3D Semantic Mapping: A Survey cites this paper.

Online Knowledge Integration for 3D Semantic Mapping: A Survey AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:14.362819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:14.362819Z digest=sha256:946e32a4d11c7cc92fc0cf2d99c1df617b7f8899a1860d0b9ace3eecd8cedc0a

Observation 765bf078-6c0e-4aeb-8ca6-43bfc85cbba4 · inbound

BYE: Build Your Encoder with One Sequence of Exploration Data for Long-Term Dynamic Scene Understanding cites this paper.

BYE: Build Your Encoder with One Sequence of Exploration Data for Long-Term Dynamic Scene Understanding AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T23:32:13.149398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:32:13.149398Z digest=sha256:e0aa15e6682b62af3caab8f2f5f33d7e8e67de13c686f38c0172fc755bf59047

Observation 9b6acf39-ecee-4f1d-9e18-ccd3b154a435 · inbound

SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts cites this paper.

SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T20:40:39.582082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:40:39.582082Z digest=sha256:a66d7ef2800982d3484d7e415a12b013c9e596259884125e2e6b66f0067b5048

Observation 80be0ad9-0899-40eb-8f5a-61138d6a41fd · inbound

TB-HSU: Hierarchical 3D Scene Understanding with Contextual Affordances cites this paper.

TB-HSU: Hierarchical 3D Scene Understanding with Contextual Affordances AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T20:36:59.096170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:36:59.096170Z digest=sha256:217d139188bf469b224c015d6c3a836f5eb456ac3df0040260318a7d56452261

Observation 50175490-f70b-4544-8417-83e4e04e0a24 · inbound

TANGO: Training-free Embodied AI Agents for Open-world Tasks cites this paper.

TANGO: Training-free Embodied AI Agents for Open-world Tasks AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T21:22:36.624899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:22:36.624899Z digest=sha256:2afa764e1c28ddca9044849d21cccdbfd952c620875940a6d7cd42b197c1c0a2

Observation 3e10368a-630f-47ef-b5ea-756ea3814e72 · inbound

Efficient Policy Adaptation with Contrastive Prompt Ensemble for Embodied Agents cites this paper.

Efficient Policy Adaptation with Contrastive Prompt Ensemble for Embodied Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T14:57:26.640943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:57:26.640943Z digest=sha256:8e9f81046fa27bbe9fbcd4dc4629b4e0118dbc5915be835c4923cef8a282e35e

Observation 569cd454-3df1-445a-97ad-79447ed790ec · inbound

ManiSkill-HAB: A Benchmark for Low-Level Manipulation in Home Rearrangement Tasks cites this paper.

ManiSkill-HAB: A Benchmark for Low-Level Manipulation in Home Rearrangement Tasks AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:04:14.771119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:04:14.771119Z digest=sha256:89e4093e32f8da8dc8bf085939c8dba810170cd88997ff280067fe8820c36e4b

Observation 2c3728d9-80b9-46b8-b077-0a1a73c400dc · inbound

The One RING: a Robotic Indoor Navigation Generalist cites this paper.

The One RING: a Robotic Indoor Navigation Generalist AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T12:20:28.118313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:20:28.118313Z digest=sha256:1fa7ca88f20425fbb6da63fc0cbaf84f52586d234512a1e7445f69f5617f6abd

Observation 39bbe53f-2065-4bbc-9169-8fa60c692d74 · inbound

Multi-Modal Grounded Planning and Efficient Replanning For Learning Embodied Agents with A Few Examples cites this paper.

Multi-Modal Grounded Planning and Efficient Replanning For Learning Embodied Agents with A Few Examples AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T05:42:03.784007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:42:03.784007Z digest=sha256:66ea521b2dc9c8308073d9ccb34deffb745280cefb6614d3e8e06b024ae46c72

Observation 35066bc1-a63c-4196-8c90-e35bbedc7331 · inbound

Embodied Image Quality Assessment for Robotic Intelligence cites this paper.

Embodied Image Quality Assessment for Robotic Intelligence AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T04:35:18.676758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:35:18.676758Z digest=sha256:dc5b2ca825945c04a980c9aad181f3ee78ad95d3062948e582283ca964cf2a23

Observation ebc91bea-74a8-48c3-8d77-baf0e7c38a5e · inbound

Mobile Robots through Task-Based Human Instructions using Incremental Curriculum Learning cites this paper.

Mobile Robots through Task-Based Human Instructions using Incremental Curriculum Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T00:55:36.072193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:55:36.072193Z digest=sha256:d6e92def9c3faeea6b0f013e9530ac990c815d806f1379c9a5db964017b53e07

Observation 9dc14958-50c8-42bb-a725-d4710cab2b23 · inbound

UnrealZoo: Enriching Photo-realistic Virtual Worlds for Embodied AI cites this paper.

UnrealZoo: Enriching Photo-realistic Virtual Worlds for Embodied AI AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T23:08:23.012300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:08:23.012300Z digest=sha256:0ad8ef1bfec918ec5a9bc2606d9bce0fc0acbd3248eac713d0933a74780c7408

Observation 6794fea9-146b-409e-8b10-5bf9c89e740f · inbound

A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges cites this paper.

A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-10T22:17:23.201554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:17:23.201554Z digest=sha256:e3c8205cf4faf3bba8043a2b11a6604e2f531673dca6e4a810f5cb012cc686e3

Observation 456dec47-d870-4d6d-accf-7a4808134c2d · inbound

Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches cites this paper.

Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 226

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:12.903085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:12.903085Z digest=sha256:da7a6df9ee40540099ba90ec81c515a9803abfee421561cdf0f1bdb31d5487a4

Observation 4b4d039f-8fa7-49c7-af01-532f6f7b4cd7 · inbound

ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark cites this paper.

ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:25:47.053331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:25:47.053331Z digest=sha256:7d5e5db6d38fd03be246ba46904a04aa931eb3b8427745c3ba78fbc31f91502d

Observation 11f9d73c-fe41-4568-8245-86e8ee8fff3b · inbound

LongViTU: Instruction Tuning for Long-Form Video Understanding cites this paper.

LongViTU: Instruction Tuning for Long-Form Video Understanding AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:57.907464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:57.907464Z digest=sha256:f8b9746ef164fc3700de922d80d09e8568bd273a7a6b7ba7ce1d6c489bf67894

Observation bbed4e1e-e391-49e0-a1e6-972445575f29 · inbound

EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents cites this paper.

EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 373

Resolution
unresolved
no resolver link, observed 2026-08-10T17:51:10.556523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:51:10.556523Z digest=sha256:029206011c36e9ab5af5c4962a021a566de7eefa99fb3e6b8fb7fda4d7c05a06

Observation 30a559e7-2bf1-478c-853b-a462c84f49e3 · inbound

Embodied Intelligence for 3D Understanding: A Survey on 3D Scene Question Answering cites this paper.

Embodied Intelligence for 3D Understanding: A Survey on 3D Scene Question Answering AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T19:24:17.640967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:24:17.640967Z digest=sha256:3b08331cfe34b3013d09f5981d00ebe5f8b57563a33f226421ae02cfb200ded1

Observation 969c0735-d55a-4d67-a168-d874de1b3d23 · inbound

Learning to Plan with Personalized Preferences cites this paper.

Learning to Plan with Personalized Preferences AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T17:32:27.904372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T17:32:27.904372Z digest=sha256:4d5b279ee1be409f23f9995a87a24fd66cf1498e78beba4a8d206047266e2113

Observation 4d6953f3-5ed4-4dc1-a153-19fafc36328e · inbound

AdaptBot: Combining LLM with Knowledge Graphs and Human Input for Generic-to-Specific Task Decomposition and Knowledge Refinement cites this paper.

AdaptBot: Combining LLM with Knowledge Graphs and Human Input for Generic-to-Specific Task Decomposition and Knowledge Refinement AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T13:31:47.214651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:31:47.214651Z digest=sha256:f5edf66380c5ce35192aee6fc3cbc463e5173a3a6ca673b428e9ea5f4f7d0846

Observation dfb9161b-1377-421c-bbe6-35c83f602c89 · inbound

LLM-Powered Decentralized Generative Agents with Adaptive Hierarchical Knowledge Graph for Cooperative Planning cites this paper.

LLM-Powered Decentralized Generative Agents with Adaptive Hierarchical Knowledge Graph for Cooperative Planning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-08T19:19:43.887704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:19:43.887704Z digest=sha256:ffd439b60d00f6e6b77e336f1cb170daa9dd19e8f4b2e7a34baf800ed7806735

Observation 8f15c223-c8cc-4613-afd3-7072217fe6a7 · inbound

When Incentives Backfire, Data Stops Being Human cites this paper.

When Incentives Backfire, Data Stops Being Human AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T11:47:48.987816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:47:48.987816Z digest=sha256:d02fbaf3eb7b21091506f6fe024d6894ab8ea83914e1039f37fff9e81249c6a1

Observation 4150904b-22e9-4022-a2d4-bacf6b558c74 · inbound

Collision-Aware Object-Goal Visual Navigation via Two-Stage Deep Reinforcement Learning cites this paper.

Collision-Aware Object-Goal Visual Navigation via Two-Stage Deep Reinforcement Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-23T02:57:26.326481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T02:56:05.029236Z digest=sha256:c1f0f6a9ef8efae47b88fb43ca4cbb9e6abfd389108962b6d8548df793de9eb5

Observation befa3b3e-d1ef-4c9d-ba7b-702da23e90fb · inbound

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning cites this paper.

SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-05-23T01:32:22.461815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T01:27:33.123243Z digest=sha256:6134384ccf699b6399cd37bf9fda775c5e2e554405a85af65807f41be1d691a3

Observation 52791cb4-2960-4c46-a48e-6ca3d44af7f6 · inbound

Solving New Tasks by Adapting Internet Video Knowledge cites this paper.

Solving New Tasks by Adapting Internet Video Knowledge AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:33:00.779254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:33:00.779254Z digest=sha256:8a8eec11c51a4a7edc9fb367bc731775d430f700e0ca5a6f92bf3b9baa7020dd

Observation c611242e-b189-43b7-a533-4d88a3ec0684 · inbound

Multimodal Perception for Goal-oriented Navigation: A Survey cites this paper.

Multimodal Perception for Goal-oriented Navigation: A Survey AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T11:23:58.380803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:23:58.380803Z digest=sha256:a95e163d867834f31235d28e93e9fc51f1cc3f28cbe7b521f26e5068a4279185

Observation 94ae1922-af76-4704-8434-13d149ec429e · inbound

Demonstrating DVS: Dynamic Virtual-Real Simulation Platform for Mobile Robotic Tasks cites this paper.

Demonstrating DVS: Dynamic Virtual-Real Simulation Platform for Mobile Robotic Tasks AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:56.843366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:56.843366Z digest=sha256:30358e562e382040ecb52f639e8801fb541e976f5de7e43fc8dee4eb1af7180b

Observation 2e092315-257f-4cad-bee5-5ac852ae17db · inbound

A Survey of Interactive Generative Video cites this paper.

A Survey of Interactive Generative Video AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 164

Resolution
unresolved
no resolver link, observed 2026-08-16T04:56:01.064196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:56:01.064196Z digest=sha256:cc78e9483d015339df8043be783cf81e72626b4cef7d82a49f18a807bdb31ec5

Observation 0723d6bd-77d0-481a-9c65-dd0c156d6dbc · inbound

Towards Autonomous Micromobility through Scalable Urban Simulation cites this paper.

Towards Autonomous Micromobility through Scalable Urban Simulation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T04:42:09.980699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:42:09.980699Z digest=sha256:c13f7a1d623d368376ca87adaf8ccc1b1123f47aee5d43f1b35ebc693649a5ed

Observation 4948ebea-2b16-4b5f-8133-1c9b30aa2375 · inbound

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI cites this paper.

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:40.513321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:46:40.513321Z digest=sha256:1020427389bb619ea225b31026f1bc97cc835320da0facce07ac57e5680a7e82

Observation 4072340f-ab85-4af5-8860-05fe200c0407 · inbound

MetaScenes: Towards Automated Replica Creation for Real-world 3D Scans cites this paper.

MetaScenes: Towards Automated Replica Creation for Real-world 3D Scans AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T00:58:28.319894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:58:28.319894Z digest=sha256:5ad529b681c8153c5c5a4e69d523344ae947552db587b83f602685e0062bde20

Observation bb0d66e1-0446-4196-a93a-3e97cd5a5478 · inbound

Scenethesis: A Language and Vision Agentic Framework for 3D Scene Generation cites this paper.

Scenethesis: A Language and Vision Agentic Framework for 3D Scene Generation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:46:00.508809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:46:00.508809Z digest=sha256:fd2f8af2ef26a2525033e46baea38ddf2585fee26a896ab463622dc5630a2eeb

Observation 88ac10ec-c4d1-4164-9f77-ca6f7e6d48b6 · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 220

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:13.060346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:13.060346Z digest=sha256:4d7688f1ad10bcb823986b730d9dc54dd546934013fb8fd0d0fcec5d9250553b

Observation d6c6045b-9ea5-4ef1-9d8a-49f48a9e2800 · inbound

ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation cites this paper.

ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T21:31:50.388247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:31:50.388247Z digest=sha256:c18d6d8c64fe23bd62b185aaaab54d32c43505715503fe7b4f9d99b2988ffd70

Observation 31ea7cd4-766b-4f18-aa56-dda60a9b95c9 · inbound

A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics cites this paper.

A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T20:33:16.229244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:33:16.229244Z digest=sha256:6bbf3e779cd290165104d6b73c4e4cea216bc8d49633ea72ef80934f2a4655dc

Observation d231788a-7b0b-4030-b383-4bd9994ffa09 · inbound

SayCoNav: Utilizing Large Language Models for Adaptive Collaboration in Decentralized Multi-Robot Navigation cites this paper.

SayCoNav: Utilizing Large Language Models for Adaptive Collaboration in Decentralized Multi-Robot Navigation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:13:45.403335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:13:45.403335Z digest=sha256:63e157063abdfb0c55538b2133071d3d63c800c13d4737f53d6eea2fc947ab1e

Observation 89fac9a5-be2f-424e-96fa-2ecab99ce51e · inbound

Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets cites this paper.

Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:12.433304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:12.433304Z digest=sha256:d730cd0104855955ee5b1567ea4d4da3a2e0be83056e7f7f731219a13918c07e

Observation beacf3ad-ed9c-478f-8e48-347edec1419e · inbound

SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes cites this paper.

SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:36.492540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:36.492540Z digest=sha256:06ae9a0fdcb5c3e57e62a66f866cdbc2528c9d1f43d5bfcb4a90a0ab93f141f9

Observation f9e58caf-ea94-495d-8865-a4b90c8a8882 · inbound

Agentic 3D Scene Generation with Spatially Contextualized VLMs cites this paper.

Agentic 3D Scene Generation with Spatially Contextualized VLMs AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:29.202280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:29.202280Z digest=sha256:bb26d59778d813b7372a2779fa84990f73a0ccc8f8cbb73cbae808e711aef3a9

Observation 6db2704a-a6d9-4e34-8747-f03ec0022335 · inbound

Agent-Environment Alignment via Automated Interface Generation cites this paper.

Agent-Environment Alignment via Automated Interface Generation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:31.886169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:31.886169Z digest=sha256:4d0fc1fc9513d378a2f6199a7a055e6bfdbd1fff96596fc5f0c21772fe5e21f0

Observation 59e6988c-2cf4-4ed4-b91d-b3763ddf060e · inbound

Reinforced Reasoning for Embodied Planning cites this paper.

Reinforced Reasoning for Embodied Planning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:23.067075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:22:23.067075Z digest=sha256:1e03d4ed69bef447b9f9eae4c549d43074b67b4601bf251dd6fe52814d7ac241

Observation 74c0df06-47c1-4d1d-b966-e713bb109112 · inbound

SEMNAV: Enhancing Visual Semantic Navigation in Robotics through Semantic Segmentation cites this paper.

SEMNAV: Enhancing Visual Semantic Navigation in Robotics through Semantic Segmentation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:40:54.050191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T01:36:25.526965Z digest=sha256:4d766eca5ec374242711b02a948a1de4bb3e0a3cf35d0c53fd19abe890218328

Observation b4094c0f-eece-4171-b7c4-6a40c3ad120c · inbound

NTIRE 2025 Challenge on HR Depth from Images of Specular and Transparent Surfaces cites this paper.

NTIRE 2025 Challenge on HR Depth from Images of Specular and Transparent Surfaces AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:24:24.937823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:24:24.937823Z digest=sha256:20af786d1492072e1a5b13ceed1e256f6406331ef073c86bae8be88e7f9e009c

Observation 73dd62dc-55bc-43b5-974d-2b70083d1040 · inbound

Multimodal Spatial Language Maps for Robot Navigation and Manipulation cites this paper.

Multimodal Spatial Language Maps for Robot Navigation and Manipulation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:46.515374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:53:46.515374Z digest=sha256:cef1492eb1f0507b2d27ced45f2b0e07d7fbafc02a26a8bc5ebcec6b133116bf

Observation 8966c1d6-751b-461f-8ffb-5fcf74c4c5b1 · inbound

IntPhys 2: Benchmarking Intuitive Physics Understanding In Complex Synthetic Environments cites this paper.

IntPhys 2: Benchmarking Intuitive Physics Understanding In Complex Synthetic Environments AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:44:03.106298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:44:03.106298Z digest=sha256:0627d89c1093273b8e0a57d422c646f92bc3679c6f4dde7cadc2fcc4d363a819

Observation 65cf1253-22de-4e62-8a98-264988fc72db · inbound

GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation cites this paper.

GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:17:09.324886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:17:09.324886Z digest=sha256:7c9039d9321cf40b7059f6507357a683a1a6dde44532c2f96835e2b8408dbf2b

Observation 86805c05-b5b2-4265-b8e7-87318e94d8dd · inbound

Part$^{2}$GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting cites this paper.

Part$^{2}$GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:17:11.015625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T08:14:28.190409Z digest=sha256:1a4c8dfe9efd7974c303b1f21eb80e4e4307264f70959f27876d15fb18a1b8a6

Observation e55f7c6f-893b-4940-8bd9-6c465762e31c · inbound

From 2D to 3D Cognition: A Brief Survey of General World Models cites this paper.

From 2D to 3D Cognition: A Brief Survey of General World Models AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T22:57:41.711826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:57:41.711826Z digest=sha256:7d46d83fe01665b474867e895dd60a7a129a328e9e811d3f4b1496ba0674f1a6

Observation 9284a182-b22c-4f06-9947-62a1933ab909 · inbound

Embodied Domain Adaptation for Object Detection cites this paper.

Embodied Domain Adaptation for Object Detection AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:09.752738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:09.752738Z digest=sha256:8f0d6b95efee9f2035875b53ee269099e17d16fff9fb57a84f5cebb616345e78

Observation 480211d8-933e-4bba-ad99-78aa383df567 · inbound

MDC-R: The Minecraft Dialogue Corpus with Reference cites this paper.

MDC-R: The Minecraft Dialogue Corpus with Reference AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:15:55.983398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:15:55.983398Z digest=sha256:2dccc58ff473810520261f29ee2ee71077f9dc35fd11671aea0144172b8f455a

Observation 45924ed9-0e18-4060-a563-080f6458ba87 · inbound

RoboBrain 2.0 Technical Report cites this paper.

RoboBrain 2.0 Technical Report AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:47:28.735777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:47:28.735777Z digest=sha256:bb4aca77e53c1cb6c927c97ba82c8146cf4ba663d1ea20dd6dc6abeab0994fac

Observation da88b9de-cdf6-4229-90d6-35f5bd6350e8 · inbound

What to Do Next? Memorizing skills from Egocentric Instructional Video cites this paper.

What to Do Next? Memorizing skills from Egocentric Instructional Video AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:08.880761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:03:08.880761Z digest=sha256:96b752e6b94ef6c8db167f3a624f8466dabc899693e942397a54dae7a34ff730

Observation 20f1bb53-2066-4dfb-99de-171ef71815aa · inbound

Grounded Gesture Generation: Language, Motion, and Space cites this paper.

Grounded Gesture Generation: Language, Motion, and Space AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:05.106839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:05.106839Z digest=sha256:d4f0f9a69d819a241af7cfb2ff1eba37945c3ba45e3ba83fed089d92fd916412

Observation 9002432d-a4cb-4570-8a56-14c69c567873 · inbound

Conditional Multi-Stage Failure Recovery for Embodied Agents cites this paper.

Conditional Multi-Stage Failure Recovery for Embodied Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:24.963271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:24.963271Z digest=sha256:835da7b5a70fdb3bd97590cd8e093b9c530b15954f151530b0a94c58a3e3e246

Observation 3d322b2e-9b0e-4be9-b880-f216c0232d5e · inbound

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley cites this paper.

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:21.582294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:21.582294Z digest=sha256:930573f8a4475d05bf7695cd90f83484ded75fd2ffc0150408da45e6ee908947

Observation dd1b8742-3c0b-4850-ac32-dc19d36c1a78 · inbound

Continual Reinforcement Learning by Planning with Online World Models cites this paper.

Continual Reinforcement Learning by Planning with Online World Models AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:16.832020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:10:16.832020Z digest=sha256:e2a2aad68b2aed4ef1c1e7ce175e7e3462867c3db7694f2eac9ebe8680051b4a

Observation 4138687e-0eca-461f-b720-3c91f2cfbb02 · inbound

Foundation Model Driven Robotics: A Comprehensive Review cites this paper.

Foundation Model Driven Robotics: A Comprehensive Review AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-06T17:43:53.349863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:43:53.349863Z digest=sha256:e2bc40671a7be45bfc9e049583c58b468349415ed14258831848d6eeb45abcde

Observation 76fb6c6d-b446-41dc-bd32-301e38e460ed · inbound

CogDDN: A Cognitive Demand-Driven Navigation with Decision Optimization and Dual-Process Thinking cites this paper.

CogDDN: A Cognitive Demand-Driven Navigation with Decision Optimization and Dual-Process Thinking AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:18:26.474982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:18:26.474982Z digest=sha256:542a55e80163c8e9452093fa889b8bed172fc75ca9961c255bee4c789c224516

Observation 417b62cf-ba0b-45e8-b5cd-4e903fe9777e · inbound

Generating Actionable Robot Knowledge Bases by Combining 3D Scene Graphs with Robot Ontologies cites this paper.

Generating Actionable Robot Knowledge Bases by Combining 3D Scene Graphs with Robot Ontologies AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:05:53.169401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:05:53.169401Z digest=sha256:6d8ba7cd206ec419ff4e6af2ca88af1dc3b5e6fffc427c45156d44177ce88037

Observation 2ba95108-0d47-42a1-81ff-89d81045138a · inbound

MoDeSuite: Robot Learning Task Suite for Benchmarking Mobile Manipulation with Deformable Objects cites this paper.

MoDeSuite: Robot Learning Task Suite for Benchmarking Mobile Manipulation with Deformable Objects AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T12:25:48.744855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:25:48.744855Z digest=sha256:0c3511571fe20a1cfff887a1626f68b061215d3ac371da615f8b2a3ce676a7cc

Observation d253e52c-c7ad-44df-ba48-dfc4720cb8b6 · inbound

Modality-Aware Feature Matching in Visual and Vision-Language Applications: A Comprehensive Survey cites this paper.

Modality-Aware Feature Matching in Visual and Vision-Language Applications: A Comprehensive Survey AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-06T11:20:06.683880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:20:06.683880Z digest=sha256:6785ada2d75654b111d409dfec0a3bf21699d154f6e96df270cb6d06ad514253

Observation 38aa89cb-9f60-45ab-9602-498fad752492 · inbound

UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents cites this paper.

UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T10:15:11.357220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:15:11.357220Z digest=sha256:b5dd53b2ffe05dfe98a5b89b16c83d1068f0e6b87418d3566aec36702843472d

Observation ae86a301-103b-4f4f-bd50-1bb35b4cb51c · inbound

Sari Sandbox: A Virtual Retail Store Environment for Embodied AI Agents cites this paper.

Sari Sandbox: A Virtual Retail Store Environment for Embodied AI Agents AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T10:12:28.856846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:12:28.856846Z digest=sha256:26db6b7f75e315954420daaa9ef56859b02312053037cf5bae19c23714463a41

Observation f47733b2-a974-4525-aa58-312900138531 · inbound

DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning cites this paper.

DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:27:34.238419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:27:34.238419Z digest=sha256:09be847813f8c7132ca40882c4c4448701725646757d840069338bd4c4eb51c7

Observation 78d2c553-7450-48de-9bfc-cc78f2d38056 · inbound

Integrating Vision Foundation Models with Reinforcement Learning for Enhanced Object Interaction cites this paper.

Integrating Vision Foundation Models with Reinforcement Learning for Enhanced Object Interaction AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:09:55.608595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:09:55.608595Z digest=sha256:a892990a660ecb3d4949b39c5abdfb48dcaec028dfdb06a8e2e4384c391b6d4d

Observation ee9d3479-84b1-49f9-9704-324457b388a9 · inbound

DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation cites this paper.

DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T21:05:31.831949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:05:31.831949Z digest=sha256:f82495890b8475efdba7415b2d3f6a6277e971ef3d4b7f0a586b3e4241ffed74

Observation 6424cdae-e634-487c-88fa-e16ac89e42f2 · inbound

Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent cites this paper.

Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:05:10.767543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:05:10.767543Z digest=sha256:f513e9fe85ad94bfbfa0faf464dcbf374c74b512781c4bbc461b749cf38a1a26

Observation e8061d6f-a4eb-4ef1-9b63-699a5b383caa · inbound

SPG: Style-Prompting Guidance for Style-Specific Content Creation cites this paper.

SPG: Style-Prompting Guidance for Style-Specific Content Creation AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T19:54:56.935817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:54:56.935817Z digest=sha256:3f7af1e29c990b1fc19a18ee010e495affa557ad09b9e09c4aae9c251d67fb49

Observation c6dcdfe3-7ea1-4943-82cd-0c3e62ee678a · inbound

Integration of Robot and Scene Kinematics for Sequential Mobile Manipulation Planning cites this paper.

Integration of Robot and Scene Kinematics for Sequential Mobile Manipulation Planning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T16:58:34.168782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:58:34.168782Z digest=sha256:729591d365a9fade8ee5e8b66a112b44316bee28656f22cf416151390409f699

Observation 93ecc3fa-19b8-45a6-8695-d234e1dace8d · inbound

Constrained Decoding for Safe Robot Navigation Foundation Models cites this paper.

Constrained Decoding for Safe Robot Navigation Foundation Models AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-18T19:11:47.478244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T19:07:18.832187Z digest=sha256:71a37cbaf7ea8a3a46fa34c19f12a4431879a1b07aefd74c2e12eb2fe695c4d4

Observation 321e73f9-a5cb-4c18-9f98-8102aeccbeeb · inbound

Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture cites this paper.

Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T11:40:22.265805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:40:22.265805Z digest=sha256:d5afbccc7830198902aa9fd05ae743cb917a7b67fc7f1a6320b439e2b5f58bad

Observation e1dc4d94-a903-4afe-bc49-7d95ce860bad · inbound

ANNIE: Be Careful of Your Robots cites this paper.

ANNIE: Be Careful of Your Robots AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T11:00:38.166570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:00:38.166570Z digest=sha256:2bda6c734922b85e0119ecd5ae01ff3aa8d4e786175918f1d0319d9fa4efb16a

Observation 5a338012-32ac-4ad5-a0b2-ac783f65d162 · inbound

Mind Meets Space: Rethinking Agentic Spatial Intelligence from a Neuroscience-inspired Perspective cites this paper.

Mind Meets Space: Rethinking Agentic Spatial Intelligence from a Neuroscience-inspired Perspective AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-04T19:39:06.723001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:39:06.723001Z digest=sha256:b0d4225ff47d8571f3352e783bb93a8235844b523f8a69edb5ed18b2749bbce2

Observation 14064f24-6d38-4391-b8ab-ea2fa9233a86 · inbound

Curriculum-Based Multi-Tier Semantic Exploration via Deep Reinforcement Learning cites this paper.

Curriculum-Based Multi-Tier Semantic Exploration via Deep Reinforcement Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T19:17:00.579682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:17:00.579682Z digest=sha256:54c06bba16d6d7ff6f543e3c7475d049dbf3f1cfa95cdf1df684ed5efdfd2844

Observation f1bf732e-b3f5-4b80-86af-5bebc2436fcd · inbound

InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts cites this paper.

InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-18T17:11:40.134078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T17:11:26.524195Z digest=sha256:2b08d88ffcb6855b2b8b664cc1d7a866137bef01d2a949692563359b17db5867

Observation 696158e1-a3f7-44a2-a299-f5da49edcb44 · inbound

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI cites this paper.

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 241

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:01:14.296374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T09:56:36.716680Z digest=sha256:1e6ea9987a888d97f438e21eb871e9fc92b02e866235676ed812ea47ac487882

Observation 7894e5b7-0ba8-48a8-a727-e74e6c2bce8e · inbound

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning cites this paper.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.183849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.183849Z digest=sha256:f2be5ffeb853cb9eab88bea4374f060afbc42be48b839a6eb52e58d91935e8f9

Observation 7f87c461-47f7-486a-ae0d-dfa0dcf1e794 · inbound

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models cites this paper.

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T14:57:29.117151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:57:29.117151Z digest=sha256:c9c289983021db3dc6e379c4ce208f8bc21b22518484bdd3df88f900b91ebd9f

Observation fe4da3e3-1368-4eb5-bc67-624a44cb9f43 · inbound

AION: Aerial Indoor Object-Goal Navigation Using Dual-Policy Reinforcement Learning cites this paper.

AION: Aerial Indoor Object-Goal Navigation Using Dual-Policy Reinforcement Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T08:53:46.818451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:53:46.818451Z digest=sha256:29e3ea5eb267960670a9f0091693660817a17e0d74be7216f80e717481a0a505

Observation 87ead5a3-d291-42b5-9272-f9225fbed674 · inbound

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning cites this paper.

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:10:45.588045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T08:07:48.508022Z digest=sha256:3846f6e4ed748a324b9aa3c8dc4a5f0a0f1d6b941f9041324d3f1156738ca635

Observation 6162131b-afa5-4c08-b73c-de5dd85db89a · inbound

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs cites this paper.

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-16T06:00:40.510559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T06:00:02.043029Z digest=sha256:e32af1d093d0303df4e0080d6bcd04fb772c8b7305100876975675fa7c77f989