Pith. sign in

Paper Citation Record · LEDGER

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations

As of 21 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2412.11412.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11412 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:01:33.101578Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 88d7e4c3-a819-40c6-89bc-7e14d3ffb88a · outbound

This paper cites ARK- itscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations ARK- itscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.130075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.819069Z digest=sha256:eb6694543a612502b70da5d25e504622bfce53d8361ca210d9cd14b0f2bd1348

Observation 5deb3153-9e0e-48cd-a832-a81b6b5bce56 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.825902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.825902Z digest=sha256:9d5eed9d45df54ab347ffd974b9f57abf7d00c7e33b531a7a2aa5122a1798601

Observation 45870842-9e13-4590-b8ba-13c3663a0c1d · outbound

This paper cites Omni3d: A large benchmark and model for 3d object detection in the wild.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Omni3d: A large benchmark and model for 3d object detection in the wild

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.107460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.832395Z digest=sha256:066cdee5186976a499c4cb7669c934b120dd3d4a104bc8fddc311c0f1790c7fe

Observation 5441bd91-2ee2-4822-9941-6f2a2edf5ac8 · outbound

This paper cites Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.087970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.837973Z digest=sha256:79e01558d3f6007787fb7dbb706cfc4b4a53f59abfe041ac32da93c91f931561

Observation 1e58a82f-6466-4a47-ad05-da3854ba49e2 · outbound

This paper cites End-to- end object detection with transformers.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations End-to- end object detection with transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.843953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.843953Z digest=sha256:827cf11d7a855bd49453dd2bd5ebd3e045037034241fda4212cd858cc994d6c3

Observation f0eac324-f7a8-4a70-8c99-c8d8b793f9b4 · outbound

This paper cites Exploring classification equilibrium in long-tailed object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Exploring classification equilibrium in long-tailed object detection

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.052947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.850680Z digest=sha256:d690ad7a1199a08da3b67a3c5274a18cdd56d3e2cd8c8ec9cc97ca44f4f26733

Observation aeb6ac87-6f10-461a-b8e0-22c7034c164b · outbound

This paper cites Dqs3d: Densely-matched quantization- aware semi-supervised 3d detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Dqs3d: Densely-matched quantization- aware semi-supervised 3d detection

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.033969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.856107Z digest=sha256:d3475c0f50bb9f6e8bf2e8768539c111132d26e0cf220701c3ffa47df2e31ee4

Observation 990c192f-eb67-4516-807f-431d14fa9c31 · outbound

This paper cites Fast r-cnn.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Fast r-cnn

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.015333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.862565Z digest=sha256:89c3d18bd74e0711e4c75548999eca023a6e2f32487a9597e2df94eebf876285

Observation c5d034e0-83f5-4a34-bd6b-d9030947d806 · outbound

This paper cites Rich feature hierarchies for accurate object detection and semantic segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Rich feature hierarchies for accurate object detection and semantic segmentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.998956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.868468Z digest=sha256:e5d09783cb04a6504ea744e5cfe550b99fadecf5e947b42e8240529d7da89449

Observation 710d132a-d2d7-49fb-a757-7b7180520ace · outbound

This paper cites Open-vocabulary object detection via vision and language knowledge distillation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary object detection via vision and language knowledge distillation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.980561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.873751Z digest=sha256:c0a657b993c9230ba64a18c9a59feac817312ac69093a328d1514ae4d8ab5687

Observation 99a5487d-c825-4040-9ecd-e184b426a038 · outbound

This paper cites LVIS: A dataset for large vocabulary instance segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations LVIS: A dataset for large vocabulary instance segmentation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.962656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.879582Z digest=sha256:d3487c804bf1930977399bab1ffe16345aeda54c250c44af79460332e5658401

Observation 59b1d90f-f408-4140-b337-526a3e9fe003 · outbound

This paper cites Mask r-cnn.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Mask r-cnn

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.944279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.885508Z digest=sha256:9f3cd9181757037132685a78d6f8b41c9ffbb466cb42424c329addcbdf263368

Observation dd0d3eb3-9740-45a9-90e9-c545304fbf58 · outbound

This paper cites Cooperative holistic scene understanding: Unifying 3d object, layout and camera pose estimation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Cooperative holistic scene understanding: Unifying 3d object, layout and camera pose estimation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.926961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.891321Z digest=sha256:391549fdd2b162728b7f60926ea030229b37eae890362eee9602a57b43c8c304

Observation 166f0d29-6c2e-441e-a923-87f8aa162fd8 · outbound

This paper cites 3d-relnet: Joint object and relational network for 3d prediction.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations 3d-relnet: Joint object and relational network for 3d prediction

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.909519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.897240Z digest=sha256:646c626e5d4dd4e51a3ea7194a057a31cf047959d38a93bb40ff1808b87cc34a

Observation 654a91fe-b3c2-4320-810f-cba6662aa775 · outbound

This paper cites Cornernet: Detecting objects as paired keypoints.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Cornernet: Detecting objects as paired keypoints

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.891078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.904736Z digest=sha256:5990a95564a3d46482ca9f50e5c41951fef12c6e269bbe90dd03476321988d07

Observation b1519e68-ff5f-4d98-a294-c8e93c924c1c · outbound

This paper cites Overcoming classifier im- balance for long-tail object detection with balanced group softmax.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Overcoming classifier im- balance for long-tail object detection with balanced group softmax

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.873986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.910845Z digest=sha256:94cdf5d854790252ff0c30dee2edad491885a3ed1d6ecfc8125154c95ec05f2b

Observation 53370ef3-ee1f-4b19-8a72-1da4c33ad235 · outbound

This paper cites Towards Unified 3D Object Detection via Algorithm and Data Unification.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Towards Unified 3D Object Detection via Algorithm and Data Unification

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-11T15:01:33.248430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.917088Z digest=sha256:08095b498b3bc112af021e843372a461e81f563672c11cfc2abf8cbfa8e63e53

Observation 97e0f6ff-fb58-40ea-ab56-4712a01e5ce1 · outbound

This paper cites Ssd: Single shot multibox detector.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Ssd: Single shot multibox detector

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.856016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.926872Z digest=sha256:5be5fcd021e2e9d65f7658466a7584e1a1ea91154f6389cb2e1056d95b774b71

Observation 36342aa1-3951-4467-8802-58fbcdd5f140 · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary point-cloud object detection without 3d an- notation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.838797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.931772Z digest=sha256:1c27292a8ca8f17545ebb35ac5cb9ff0513dc0449f066522c6acc979a87eee2d

Observation 25943ee2-4262-4070-9764-530e0ac5ee3c · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary point-cloud object detection without 3d an- notation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.821828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.936810Z digest=sha256:b8eefd8297ff8d314b26452c878f2d57e76a6e2a5cfe89e1e900602c85a0ab9c

Observation 491d8e27-e166-4fa6-8764-fc4de63ac762 · outbound

This paper cites Total3dunderstanding: Joint lay- out, object pose and mesh reconstruction for indoor scenes from a single image.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Total3dunderstanding: Joint lay- out, object pose and mesh reconstruction for indoor scenes from a single image

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.804084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.941913Z digest=sha256:463440a4ee2af9c91629f9fe53792a40f1ae48be3f6a65292d9f1e5a9f733bba

Observation 5e8ce9cd-2d90-46ca-b45f-0e1d706b16d8 · outbound

This paper cites On model calibration for long-tailed object detection and instance segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations On model calibration for long-tailed object detection and instance segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.785036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.946996Z digest=sha256:328fb43577f9fc7e88fdf7d5a15af5d14ed22e48e65b03ac50becd3602a6c216

Observation 6188f051-6e9c-4e9c-8eb0-636d3780fc1c · outbound

This paper cites High quality entity segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations High quality entity segmentation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.762241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.952133Z digest=sha256:1c5d3b20fd22ad300267801d5300f615458cabfa201704e017bbf6b06ce3a18b

Observation 440c4190-259e-4f95-bcff-a94ccf8a616a · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Learning Transferable Visual Models From Natural Language Supervision

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.956830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.956830Z digest=sha256:9883c8d5f5118718f08e1464a5dbb4cf6afe90be2ac27b8005a61d09a9b3dc28

Observation 486a7f24-64bb-4348-8eee-799f93b9bbab · outbound

This paper cites Improved visual-semantic alignment for zero-shot object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Improved visual-semantic alignment for zero-shot object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.744752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.962414Z digest=sha256:64d0746bbe8b6d1d2a8f8aa4ac5f6cea906a6d05591a32092f6225376fcb2a3c

Observation 424b6811-85c8-4ad4-a860-5c4da54398ad · outbound

This paper cites You Only Look Once: Unified, Real-Time Object Detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations You Only Look Once: Unified, Real-Time Object Detection

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:32.967400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:32.967400Z digest=sha256:378009d98deaa6cb77d88335a24a03d45d3d736f00517d84140d68c14cd4434b

Observation dc4f12af-3a8d-479a-a7a8-266e3a327367 · outbound

This paper cites Yolo9000: Better, faster, stronger.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Yolo9000: Better, faster, stronger

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.724336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.972620Z digest=sha256:16bd41c0fa90c8bce35ba920742b401cfe04def347abf1bfa24e46339e19933f

Observation 68565193-9ad0-4579-b4b8-11e148445e53 · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.705623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.977879Z digest=sha256:969e02eef95d805c82b420ab4c2e315732d13b058d4c2d32859a6ff5313eb27a

Observation 65d8cf1b-13cd-4010-8595-fdfdc80b6c0c · outbound

This paper cites Susskind.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Susskind

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.686990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.983136Z digest=sha256:e7ea7e86fd52dc10e82bc21cc4a474e24824b17b52f963efb77ccafca047f306

Observation 237754b2-08d2-4433-8b2c-fc09950fcb95 · outbound

This paper cites Language- grounded indoor 3d semantic segmentation in the wild.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Language- grounded indoor 3d semantic segmentation in the wild

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.667928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.989741Z digest=sha256:81e22870d982299a621ac922a5522174fdbacd1ff0f1960fc2f374457db064f0

Observation 9a20f1d5-44cb-400c-96d4-f15d7f3359de · outbound

This paper cites Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.648685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:32.997903Z digest=sha256:aba324890b8781fd53a394a3285036ab823137d48b22f89bf948ee86b95209ff

Observation e88424da-79a8-4aaf-80db-bd918550f311 · outbound

This paper cites Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.629442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.005910Z digest=sha256:1bb245991612cf61c8c8b3925bfacaea182da79df837c61c103830ced7142b9a

Observation 09201718-0f36-4011-a98a-dc2f9dccb921 · outbound

This paper cites Object detection with trans- formers: A review, 2023.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Object detection with trans- formers: A review, 2023

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.608207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.012374Z digest=sha256:390099d2e5e263c428432ebc6c63d95ee9e76305320a7fcc9aa0ba46f3821896

Observation 9eac7d22-94d9-46bb-af41-dce3d90e9998 · outbound

This paper cites an unresolved cited work.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:33.587612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.018628Z digest=sha256:a0fcec34e0b9dc77d762146e344a291c8dd6aa929d0e80eea9c8185e027bce41

Observation 8c365951-a0a7-46b0-af0b-821fb756bf85 · outbound

This paper cites Lichtenberg, and Jianxiong Xiao.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Lichtenberg, and Jianxiong Xiao

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.566083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.024300Z digest=sha256:8a163df8fce6c789e76eac553aaa9151ad86d8965ced655459c064f06033266b

Observation 95431c26-5f2f-42de-927c-354b01eabecf · outbound

This paper cites Equalization loss v2: A new gradient balance ap- proach for long-tailed object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Equalization loss v2: A new gradient balance ap- proach for long-tailed object detection

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.547484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.030727Z digest=sha256:60f65cc3967bdfad6da89a5e37ac7580b4dbc7551a6c65583a3ecfef22539c19

Observation e303a519-1db5-482f-8a30-314b9f5829a1 · outbound

This paper cites Equalization loss for long-tailed object recognition.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Equalization loss for long-tailed object recognition

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.529520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.036167Z digest=sha256:a26ecb99cf6dc34c7156f15bdcfd472e0e90693d0f58f3ee1cd9f7533d949acb

Observation 73c663c1-342d-4235-a9d8-3fb7f4a982ee · outbound

This paper cites Im- geonet: Image-induced geometry-aware voxel representation for multi-view 3d object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Im- geonet: Image-induced geometry-aware voxel representation for multi-view 3d object detection

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.507325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.042021Z digest=sha256:130d8e28700dd79dcff841a18b73135e315426c0667f3e13b74f7345c27e8f00

Observation 444d46ca-d067-48d7-9b5b-4df0c17251fd · outbound

This paper cites Efros, and Jitendra Malik.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Efros, and Jitendra Malik

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.485450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.047593Z digest=sha256:4c8ce35c02bd4ee411a79409f3e92b3a0e3eedea22c274b1f84eae9e6d13a996

Observation 7fae7716-30e6-4cc2-9446-230a528ea6be · outbound

This paper cites Seesaw loss for long- tailed instance segmentation.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Seesaw loss for long- tailed instance segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.461296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.053003Z digest=sha256:b5a70ddec62766886d8cf1c3a170b22dce2bc12060099c5e71f30b36b2b3789a

Observation f5e5559b-19c4-400c-a8a5-70fcd0e49400 · outbound

This paper cites Detecting 11k classes: Large scale object detection without fine-grained bounding boxes.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Detecting 11k classes: Large scale object detection without fine-grained bounding boxes

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.442801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.060006Z digest=sha256:7885026a2bc93f96d725a14e7f1ed8a25945521bba85f43bc52c803a5a0af6c2

Observation 32bc1dbe-0e64-45bc-90fe-d0747c3c6652 · outbound

This paper cites Open-vocabulary object detection using captions.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Open-vocabulary object detection using captions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.422681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.065394Z digest=sha256:565b410152762ac17b21c31e3801ac8c39ecd63aa4be69860f1d8b1a4cbbc6a5

Observation 8c77f4fb-a743-4bea-a29f-651a1107745d · outbound

This paper cites Distribution alignment: A unified framework for long-tail visual recognition.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Distribution alignment: A unified framework for long-tail visual recognition

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.398101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.072574Z digest=sha256:4320a87998aef65a286932a807388483b798a05f79c0a6ddd6aac7f56db1a8d8

Observation 244f5ff2-1e3e-4507-a658-852a106662d4 · outbound

This paper cites Detecting twenty-thousand classes using image-level supervision.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Detecting twenty-thousand classes using image-level supervision

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.377779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.078707Z digest=sha256:3ca41fbbe5f3e7f604fed5c21f64f19c50ca865ab49e12137f7a9dffb211c378

Observation 97d2a028-2783-4762-933f-9bf4dad93dd6 · outbound

This paper cites Objects as Points.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Objects as Points

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.085113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:01:33.085113Z digest=sha256:5f9fd2361e1a4ef3d204a718283dd68d72d165bbdba107dc6d498b5db7ac0e13

Observation 45c93069-dc93-4b80-a892-1c2003893245 · outbound

This paper cites On the continuity of rotation representations in neural networks.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations On the continuity of rotation representations in neural networks

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.355539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.091840Z digest=sha256:6a74ab2984a5b4edfac7a18cca06d085ee5c18358c5776418739573bcf5e57b4

Observation 3eca40a0-15ec-4b56-8658-495af94c19a5 · outbound

This paper cites Tame a wild camera: in-the-wild monocular camera calibra- tion.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Tame a wild camera: in-the-wild monocular camera calibra- tion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.332911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.097085Z digest=sha256:985977d150eea5b3bdcc947351a8c89bea5fe68630f94a1c6b871036805e1a83

Observation 5578a354-832c-4560-a367-3464fc881645 · outbound

This paper cites Deformable {detr}: Deformable transform- ers for end-to-end object detection.

V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations Deformable {detr}: Deformable transform- ers for end-to-end object detection

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:33.307728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T15:01:33.101578Z digest=sha256:75b054b61ba7eb9446ad781f23aad2e00b55d462ae315d60ada9f9d440ca4106

Pith citing papers

No inbound Pith citation observations are available.