Pith. sign in

Paper Citation Record · LEDGER

Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 59 inbound Pith citation observations for arXiv:2409.18124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.18124 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 59 of 59 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:35:23.510421Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:46:41.338329Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ea05b4f9-d52e-4250-b9bc-72f41611cb75 · inbound

Bridging the Skeleton-Text Modality Gap: Diffusion-Powered Modality Alignment for Zero-shot Skeleton-based Action Recognition cites this paper.

Bridging the Skeleton-Text Modality Gap: Diffusion-Powered Modality Alignment for Zero-shot Skeleton-based Action Recognition Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T19:27:10.252315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:27:10.252315Z digest=sha256:8b925e30f559620984b1b83e7d17da4f2ab0a46d12b381280c10c9ccfb65726c

Observation be3c4eb4-4b42-4f73-95d1-46d8b8e92275 · inbound

C-DiffSET: Leveraging Latent Diffusion for SAR-to-EO Image Translation with Confidence-Guided Reliable Object Generation cites this paper.

C-DiffSET: Leveraging Latent Diffusion for SAR-to-EO Image Translation with Confidence-Guided Reliable Object Generation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T19:23:51.018426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:23:51.018426Z digest=sha256:e6ff6699b48f27ec7806d18292794d6cbceb7bbafe1e1b6a6d16381de2d5ec89

Observation 221dd6ea-73dd-41d7-8f11-64533b5411d6 · inbound

CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation cites this paper.

CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T13:18:59.377928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:18:59.377928Z digest=sha256:7a9d6e514c628405f8b0fd86297ffdfee006d9f9b85ecaa1584cabeb1146b2fe

Observation 9b02d17b-0e95-4da9-aa2b-5900149ae4f9 · inbound

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors cites this paper.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.443740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.443740Z digest=sha256:54089ebe97d0d767c21775bec4c5151f0a01553582b80ceb3d5b40669e35fe00

Observation 98e4ec60-5331-4b11-aa00-022c44bd9384 · inbound

DROID-Splat: Combining end-to-end SLAM with 3D Gaussian Splatting cites this paper.

DROID-Splat: Combining end-to-end SLAM with 3D Gaussian Splatting Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T11:57:07.054469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:57:07.054469Z digest=sha256:2bbf6afef384e21ee255ed36ee95a428bb44297fff8674cf1f2c2e3e9cea43eb

Observation 836fb508-6a0d-4ff3-8a7f-c24e733bc8cb · inbound

SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation cites this paper.

SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:28:35.221257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:28:35.221257Z digest=sha256:e2bb82fb09606d25f43c712b4cb6f2a686094a0c1c39a8c87884d7635ff3072a

Observation 862a587c-9c33-450c-b771-d1576e9dc811 · inbound

Video Depth without Video Models cites this paper.

Video Depth without Video Models Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:31:01.703813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:31:01.703813Z digest=sha256:88ade6a762ea5e122bf97f5dac155e10594dd07f1d4ca75f8bde0b6c0d4e9263

Observation bb6bbb50-d62b-4856-8fb8-b2425b000d3b · inbound

FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation cites this paper.

FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:13:20.265538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:13:20.265538Z digest=sha256:3c2fac861557c54e5f83e01e7133088011a04386d6a7756f70189d0ae61a8884

Observation cbc17bfc-b546-4d0b-a1c0-c7a01aaaa44f · inbound

Align3R: Aligned Monocular Depth Estimation for Dynamic Videos cites this paper.

Align3R: Aligned Monocular Depth Estimation for Dynamic Videos Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T22:54:11.371382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:54:11.371382Z digest=sha256:8751c2511babe7afa331f0cc7bcf38444014f0aa2634ef843a2e0c8ab7f5ddf2

Observation d1c2f536-f5bd-425c-8e7e-12b8011e5074 · inbound

How to Spin an Object: First, Get the Shape Right cites this paper.

How to Spin an Object: First, Get the Shape Right Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:55:28.259159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-23T06:52:59.386210Z digest=sha256:e65786ea75e570cd0a53bdf2f6a429439a016c8daa0b3980c22008a702d49901

Observation 7a2760fe-8600-42b9-ac77-13dd013e333b · inbound

EnvGS: Modeling View-Dependent Appearance with Environment Gaussian cites this paper.

EnvGS: Modeling View-Dependent Appearance with Environment Gaussian Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T11:35:34.768776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:35:34.768776Z digest=sha256:4347c735ee1288de419f246fe05667a34e2839315d6c20dc1cfaf51a797cf9b0

Observation 40e4a3eb-de97-4637-83ba-0b51d4943c4e · inbound

Detail-Preserving Latent Diffusion for Stable Shadow Removal cites this paper.

Detail-Preserving Latent Diffusion for Stable Shadow Removal Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T05:25:28.063408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:25:28.063408Z digest=sha256:e16172319a81eea7c7165d9b6d1244999338d819c6a011b5a631551dbd24d2e6

Observation b3e44472-b897-4049-af84-874aee2344eb · inbound

DepthMaster: Taming Diffusion Models for Monocular Depth Estimation cites this paper.

DepthMaster: Taming Diffusion Models for Monocular Depth Estimation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:12:38.878807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-23T06:08:14.988386Z digest=sha256:4b398af52e037b76ab56f809572d357ad18ba9e8b6716e5bbf5d87c53900e2ad

Observation 3b0a4cdd-7e3a-4945-852c-cf785ee4056f · inbound

TransPixeler: Advancing Text-to-Video Generation with Transparency cites this paper.

TransPixeler: Advancing Text-to-Video Generation with Transparency Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:13.497600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:13.497600Z digest=sha256:f77e957db4452e807f05b4424bca3061e36352d81997874a019dbb34e376fa98

Observation cb0743c8-e6aa-4979-b7b3-16d8b9fb7c90 · inbound

Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation cites this paper.

Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T16:30:33.284356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:30:33.284356Z digest=sha256:fe7cdf62976ae978da043fc2b4e8a8c657fa13da17d1e3f026ae037f28a38c5c

Observation 546d5089-38df-4a4a-9964-2d8508b8266c · inbound

The Fourth Monocular Depth Estimation Challenge cites this paper.

The Fourth Monocular Depth Estimation Challenge Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T10:35:23.510421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:35:23.510421Z digest=sha256:22c6518f20723d16ced6522998663080cc31022cb5f14c9fb5e4318f2ba4ba08

Observation 4dfbc172-a0a2-4479-b6fe-f3a31626b9dd · inbound

Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis cites this paper.

Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:39:13.988102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:39:13.988102Z digest=sha256:329b2bc5c16ddd2bbe271f4f0507d0312ff4dc4f5a932a328048434471255f25

Observation 78c0e7fe-87c1-417d-a699-88a2914654fe · inbound

Repurposing Marigold for Zero-Shot Metric Depth Estimation via Defocus Blur Cues cites this paper.

Repurposing Marigold for Zero-Shot Metric Depth Estimation via Defocus Blur Cues Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:13.871117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:13.871117Z digest=sha256:9e02c8e8adfa9ae7eab8217251879a51d7ae1d61123ca0550308f1ba4d8b6740

Observation 663a2a6a-235d-4991-940a-4d980609598d · inbound

Advancing high-fidelity 3D and Texture Generation with 2.5D latents cites this paper.

Advancing high-fidelity 3D and Texture Generation with 2.5D latents Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:44:51.575434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:44:51.575434Z digest=sha256:fa1826c494a4b03c0571d3532410699bdfe5b7d5138a3d7e0233b3a7496eb12c

Observation 180f7715-6b33-47ed-a81d-1f40606278b3 · inbound

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion cites this paper.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.779637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.779637Z digest=sha256:842f4fc14e59d1b94e00e0db7cc195dcfbba0a7f88c0486dab63b5ca2ae45708

Observation 94a7e30f-87e0-42a1-ae32-e8c86c1f456c · inbound

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models cites this paper.

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-07T11:36:13.586962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:36:13.586962Z digest=sha256:f4e3c149456f710b791785791ac4d5356d6e405206aa2455fe6db9769c9ca724

Observation 30127671-5022-4262-aebe-6db07d384c1d · inbound

Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment cites this paper.

Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:59.540350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:59.540350Z digest=sha256:1836a685c46fe89fd9b5ba4337d93535a97b45f1c8c4da189bc34a10775a95e8

Observation ab7224a7-6421-4a1f-8a7a-485117895efb · inbound

Aerial Multi-View Stereo via Adaptive Depth Range Inference and Normal Cues cites this paper.

Aerial Multi-View Stereo via Adaptive Depth Range Inference and Normal Cues Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:00.790687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:00.790687Z digest=sha256:06eaf13708c0b7043cb09eb15ec82da2698898a8579fee56eeb05d73e69565e3

Observation 975a96a4-5035-4f4a-b262-3b421c5bad19 · inbound

NTIRE 2025 Challenge on HR Depth from Images of Specular and Transparent Surfaces cites this paper.

NTIRE 2025 Challenge on HR Depth from Images of Specular and Transparent Surfaces Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:24:24.868069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:24:24.868069Z digest=sha256:a9d88fdefb6228d3f52c14256d58fa13d8e9cfe098abf51a231d8307d01e504f

Observation f964329c-5f30-4c3a-b325-0b2b08ec8436 · inbound

DreamCube: 3D Panorama Generation via Multi-plane Synchronization cites this paper.

DreamCube: 3D Panorama Generation via Multi-plane Synchronization Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:04.598811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:04.598811Z digest=sha256:7f5ba3e2b93c7a837cbfd0e5b10e8d52f48096e089b12c8578c423798d8976f8

Observation 0fe94305-14dd-46a3-af12-4cd0ecbc02c2 · inbound

DidSee: Diffusion-Based Depth Completion for Material-Agnostic Robotic Perception and Manipulation cites this paper.

DidSee: Diffusion-Based Depth Completion for Material-Agnostic Robotic Perception and Manipulation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:23.440926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:41:23.440926Z digest=sha256:75eb8afa467db3c299fddd0d46dc4b28d46344f4861ea5a7531cd339e4b0858f

Observation 79f90f1d-71c7-49ac-a5fc-81842e49e766 · inbound

FreeDNA: Endowing Domain Adaptation of Diffusion-Based Dense Prediction with Training-Free Domain Noise Alignment cites this paper.

FreeDNA: Endowing Domain Adaptation of Diffusion-Based Dense Prediction with Training-Free Domain Noise Alignment Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:46:19.666447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:46:19.666447Z digest=sha256:0538f2698ce785cf95058371afa7116316ffc7310c251dde316365d112daf699

Observation bdc63428-b18b-46ee-8907-2e5fb4a18f28 · inbound

Depth Anything at Any Condition cites this paper.

Depth Anything at Any Condition Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:52:02.473283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:52:02.473283Z digest=sha256:361081f9d0f8df41f7dd13f0f0907c96d7f1684dd606252f9f070ea066282f75

Observation 77d1c9a8-c3b2-400f-8b7d-474a01b13429 · inbound

Region-aware Depth Scale Adaptation with Sparse Measurements cites this paper.

Region-aware Depth Scale Adaptation with Sparse Measurements Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:53:57.619782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:53:57.619782Z digest=sha256:aab6abbd90a0cd261f4dff8ac11060ebedac87f639a9f145e1d4a4a47aa5f74b

Observation 14e7924d-0ebf-4ea3-9a6b-20e8df6ad62f · inbound

BridgeDepth: Bridging Monocular and Stereo Reasoning with Latent Alignment cites this paper.

BridgeDepth: Bridging Monocular and Stereo Reasoning with Latent Alignment Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T23:59:31.170775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:59:31.170775Z digest=sha256:1945c77adc9812b877ddccbf8a3bd96391e18b7ff59b7ad2afe656ef48cda870

Observation ba09153c-562b-4969-996b-e4f2244aee5f · inbound

Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering cites this paper.

Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:21:53.173724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T22:17:55.678629Z digest=sha256:7f389236eec07352846639ec1dfe7e1b328e9fde0f708fe8a8d433aafb1aa18a

Observation 294bfb0b-98c4-4ebd-8b6c-fbc5cdbb8e2b · inbound

Doctoral Thesis: Geometric Deep Learning For Camera Pose Prediction, Registration, Depth Estimation, and 3D Reconstruction cites this paper.

Doctoral Thesis: Geometric Deep Learning For Camera Pose Prediction, Registration, Depth Estimation, and 3D Reconstruction Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T12:10:09.101237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:10:09.101237Z digest=sha256:c840d6538c2aac8250ae212bdad6c51034ca321ed220b0c04a9fd9ecc06703c1

Observation e6328953-5fee-41c0-98a7-e3c8ab8b4e7d · inbound

LuxDiT: Lighting Estimation with Video Diffusion Transformer cites this paper.

LuxDiT: Lighting Estimation with Video Diffusion Transformer Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:52:02.161254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:52:02.161254Z digest=sha256:11d8924fc273fcaefbd97136c471e6fa8d94f1d547aca4cd8a33d0da7dfa4eba

Observation f68478f4-e7fa-4ae7-bab9-8b4252ec8047 · inbound

Lotus-2: Advancing Geometric Dense Prediction with Powerful Image Generative Model cites this paper.

Lotus-2: Advancing Geometric Dense Prediction with Powerful Image Generative Model Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:24:18.262242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T18:21:15.831853Z digest=sha256:ed3881d4c5e9d21f0a4753ee57d3586e628affd074776189debe8cebf5f5d316

Observation eb722850-cfb4-4e3c-b4d5-a36b1b4bb884 · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:07:29.852297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:f5e162c89aa30e7d60e36c9242309593a2b3d3333766cfaaab4f0b5024ffdc19

Observation ffc36240-4368-4101-8a46-59ca50366012 · inbound

Need for Speed: Zero-Shot Depth Completion with Single-Step Diffusion cites this paper.

Need for Speed: Zero-Shot Depth Completion with Single-Step Diffusion Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:55:53.217541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T13:52:01.152288Z digest=sha256:907ba177acac86da656ae74ef3cd7c686f7311f97527020c30e577b58798dbd4

Observation bdb528e6-dcb0-47c9-9391-642d56f8b5ab · inbound

CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation cites this paper.

CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:04.133501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T15:44:05.258927Z digest=sha256:5b623eaa58c666ca39b4608c12f956190f0c29c3e9a070516104f0ef8e6c4f6c

Observation 19c2f7bb-5f14-4fe3-a243-0a609848c9d3 · inbound

Image Generators are Generalist Vision Learners cites this paper.

Image Generators are Generalist Vision Learners Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:41:06.017039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T01:14:05.034951Z digest=sha256:f34618badb82e9332d78ee476df15a1cf9501a05ccd567364e2e4a2236375eac

Observation c3ca5f94-5074-4341-a1f6-eac5360915a1 · inbound

Image Generators are Generalist Vision Learners cites this paper.

Image Generators are Generalist Vision Learners Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.678896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T07:40:46.090808Z digest=sha256:a8b0ed424925354f9fc1b8ddadbacc37c932a007fa7dd7618a609759d2400e29

Observation 984b9857-8b68-48f0-b7d0-c9c3f276538b · inbound

Diffusion Model as a Generalist Segmentation Learner cites this paper.

Diffusion Model as a Generalist Segmentation Learner Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:41:15.741138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T04:37:20.434070Z digest=sha256:7b43c84386165cf61386adb29928427ecd43e4a1443444c92f21c1890f56d8d0

Observation 34dab2b2-aebd-4957-b7ff-8de55250ca12 · inbound

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors cites this paper.

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:26:07.729248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-09T20:05:21.723724Z digest=sha256:069ea90461ddcf0fd411b45d062b1e2959a22882f0daae14140ee0ebaada9f35

Observation 9e391fed-6492-4c83-bc20-d323d8418c98 · inbound

Open-Source Image Editing Models Are Zero-Shot Vision Learners cites this paper.

Open-Source Image Editing Models Are Zero-Shot Vision Learners Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:07.400642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T16:44:53.286118Z digest=sha256:dce73a26c7246ad259da3f644f76640a31ecd594f58a933f416c082ced0b6752

Observation 5d512a63-0a03-4fbb-8260-81e1ebcd0fed · inbound

The Midas Touch for Metric Depth cites this paper.

The Midas Touch for Metric Depth Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:37:03.314831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T01:36:59.161241Z digest=sha256:681a3496f166b5894b53cc5bed811c55db8f529be08973aca8d5778a0e729755

Observation 2ef2ae40-775a-43ca-88b9-56981f96d256 · inbound

TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking cites this paper.

TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:29:29.005530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T21:28:12.151547Z digest=sha256:e3365a84d56c4134c7a24dad0a2db85bfa1f4b712408de02cc9f91d4f7859bab

Observation 5388f600-5adf-4376-8323-96fff5808f5e · inbound

Towards Consistent Video Geometry Estimation cites this paper.

Towards Consistent Video Geometry Estimation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:16.082442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T08:03:13.579650Z digest=sha256:66023a32cce141ea7d01d71ca339d221e298756ba012cfa1824cda8e3456ed53

Observation 138a33bc-af3e-49a0-8eeb-0157f0167613 · inbound

Towards Consistent Video Geometry Estimation cites this paper.

Towards Consistent Video Geometry Estimation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T12:54:15.275849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:54:15.275849Z digest=sha256:2dd33c27c192907e736fd9d097a898bcad89377d38dbbf63a02936e0998f026f

Observation 06a8dbd4-bdf2-4305-98f9-52cfd303c153 · inbound

VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching cites this paper.

VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:12:46.852715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T23:10:29.250087Z digest=sha256:88693ae55ef89fe19a8a993bc7d3e04b131aca7374ee25e058195eb253bcc54a

Observation 2a5beb73-fb6e-4cf8-8a6e-98a4b7f00b1c · inbound

Modality Forcing for Scalable Spatial Generation cites this paper.

Modality Forcing for Scalable Spatial Generation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:08:32.545865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T06:45:00.481276Z digest=sha256:db8883d651fc3b15bfef15ae2b6cb63cbd4e3a7fa76f9ddde61341cf102445ce

Observation 2c8e649b-a637-4684-9505-38e6a5c8d5d4 · inbound

Error-Conditioned Neural Solvers cites this paper.

Error-Conditioned Neural Solvers Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:49:51.646539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T04:54:56.698035Z digest=sha256:cfefa2faf7b63b3a4244a0a3f107e54c79ece5c4fb6daefcccd75cd19d75298d

Observation ea5f76b0-e70c-4938-ac22-6c5dfd6195c6 · inbound

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World cites this paper.

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:54:20.923261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T06:46:21.941055Z digest=sha256:8509b7eecd7e7b9856cb187d7e85215528a27e09be2404e7763d0ac38d6db98f

Observation a63d663f-1557-4972-bcf7-cea9189a86b0 · inbound

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World cites this paper.

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T09:40:58.691757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:40:58.691757Z digest=sha256:57772d7ddb04ed599ed1ec35366e9f7927b2a00aaa444ab7761d9a38268796f4

Observation 04c684d3-ce77-409f-8447-a0d5569cd751 · inbound

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception cites this paper.

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:44:19.316666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T06:38:39.360472Z digest=sha256:2337eb6aab56aa9e318cfea20271e18fe6b61f366c6096e298ebd8b3305b87bd

Observation effedd29-3a4e-43dd-987b-60c71bad0d7b · inbound

MUSE: Unlocking Timestep as Native Task Steering for One-Step Dense Prediction cites this paper.

MUSE: Unlocking Timestep as Native Task Steering for One-Step Dense Prediction Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:24:19.415005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T06:17:53.096700Z digest=sha256:878204b52db1060e8890cc0147d2c5d5d286f4f118e366ab19ca552c57e7ee21

Observation 52100c7d-e22b-4f21-9aa9-eea9dc924f47 · inbound

PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation cites this paper.

PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:28.679469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-03T14:34:04.566992Z digest=sha256:45b4896391733fdcb34c0b07323200d6716ea8527f9756818088ef38af9215d5

Observation 7e5b9e9e-a18d-45ad-89b9-f08657697c4a · inbound

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device cites this paper.

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T01:46:41.339736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-10T01:38:11.997908Z digest=sha256:27d341d20ba7869dfaf10140b841cb8f64bcc8d70043e7612f3d1228fe596138

Observation 070102a5-a4ce-40e2-b2db-c17c74c887ee · inbound

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras cites this paper.

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:19.000775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:16:19.000775Z digest=sha256:7e37fe7afe8cb86e977db567f27ba4c62218078642163ad099bb172fae28b3b3

Observation 2da0bb54-fc9c-4a91-91db-cdad257ee101 · inbound

RoughNet: Mapping Arctic Sea Ice Roughness Using Diffusion-Based Super-Resolution of Satellite Imagery cites this paper.

RoughNet: Mapping Arctic Sea Ice Roughness Using Diffusion-Based Super-Resolution of Satellite Imagery Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T05:24:41.894294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:24:41.894294Z digest=sha256:7e985e4dad2b25ee9f16638c4f52d7ad6283f96ebd834364912d8f71b6b24358

Observation 7240dba7-4645-4e9e-b0f1-3261910edd03 · inbound

Unified Video Dense Prediction from Disjoint Data cites this paper.

Unified Video Dense Prediction from Disjoint Data Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T07:03:29.592808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:03:29.592808Z digest=sha256:d94311f46937073885eb9c6d71b7632ded4e5773ba40ceaad12b8f7c094b6acd

Observation 81cb6fef-adbb-4bea-986f-f2a1128c0060 · inbound

WildShadowRemover: In-the-Wild Video Shadow Removal via Detail-Preserving Video Diffusion Models cites this paper.

WildShadowRemover: In-the-Wild Video Shadow Removal via Detail-Preserving Video Diffusion Models Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-01T00:34:01.601724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:34:01.601724Z digest=sha256:e0fddb17d16c4ffdbc3f4458c4597c29afb91e7d3709b26b349ec3579b8683ea