Pith. sign in

Paper Citation Record · LEDGER

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion

As of 10 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.23085.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23085 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:59:28.175667Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T21:59:43.755956Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T14:21:06.813117Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved25
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49941d42-b991-47f8-8fde-a70a8c25d5a9 · outbound

This paper cites Photorealistic monocular 3d reconstruction of humans wear- ing clothing.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Photorealistic monocular 3d reconstruction of humans wear- ing clothing

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.780982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:20.930873Z digest=sha256:1addb3cca757c618aa0ba9af31bb7611b34f80cb301673f0827164eea176c2b7

Observation 367974e2-3c0b-494b-9592-b44c88e3edad · outbound

This paper cites Lumiere: A Space-Time Diffusion Model for Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Lumiere: A Space-Time Diffusion Model for Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.994210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.994210Z digest=sha256:0de3ba54a7aa1621247a66bb6fd0e103cbb8a4629840b121bab7bfa46dd7ac6d

Observation 34b33b62-3e13-48f7-a25c-febeff7487d5 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.107495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.107495Z digest=sha256:48227ca77b8f1f536c7bb31972d68f6f1d5c0acb2286ce730033896653ab3144

Observation 5eeb489c-94f3-4068-90fd-ef64ce9dc5bb · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.225920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.225920Z digest=sha256:eb1596b8b4ceee6a9da52f6306251a8fb25c608d6f508059abf8eaa1459e34f1

Observation 761a963f-c4dd-4e56-8b3a-b1be18517d16 · outbound

This paper cites Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth Pro: Sharp Monocular Metric Depth in Less Than a Second

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.343777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.343777Z digest=sha256:71f766a4e6c37eb3113f01225ca3bb13094986abf4b8e6772b40355257a5d8db

Observation 3f355d39-2a11-42f3-af29-3aed37b395c5 · outbound

This paper cites Video generation models as world simulators,.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Video generation models as world simulators,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.482823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.482823Z digest=sha256:411c1356fc5cb142824de15bf50facc0bda07c5457889b09da17dcd0d2300e6b

Observation 472d4005-bd5f-4931-80a7-7d08cb53fa49 · outbound

This paper cites High accuracy optical flow estimation based on a theory for warping.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion High accuracy optical flow estimation based on a theory for warping

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.611170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:21.585722Z digest=sha256:a2be52967adad0144baa35a97f2a015176fb5468dcc17eff29b793a0c2ce10d1

Observation e60237ce-cecf-49a6-9e0a-a7594c0d0c74 · outbound

This paper cites Stable- video: Text-driven consistency-aware diffusion video edit- ing.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Stable- video: Text-driven consistency-aware diffusion video edit- ing

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.442026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:21.742575Z digest=sha256:66f67c0a57a8c87c2bafe800d7d772c2041d1c0bd2b59fddba99b594c58edadd

Observation b1b1a472-c114-448c-b728-2b231a1625c6 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.840777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.840777Z digest=sha256:dde3ae7a685205495f80588325867c578a5aaa85a6ac0e729fc615659fc78242

Observation f5894e7a-d201-4265-923d-2cb0cc62fd58 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Videocrafter2: Overcoming data limitations for high-quality video diffusion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.235242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:21.956046Z digest=sha256:6c2fbfb2a141c9f4262daf4d6c637155f897052c62fa61dd8db5ea09edf1104b

Observation 4671f298-d307-4434-97d0-d02a079fa059 · outbound

This paper cites Self-supervised learning with geometric constraints in monocular video: Connecting flow, depth, and camera.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Self-supervised learning with geometric constraints in monocular video: Connecting flow, depth, and camera

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.104114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.077547Z digest=sha256:282268994b93d389bb5a10100c54a0b3339043841a016e19d314fea38cbd4235

Observation 29448567-24b0-4ad9-8531-348f540a34cc · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hi- erarchical transformers.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Cogview2: Faster and better text-to-image generation via hi- erarchical transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.826291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.208238Z digest=sha256:eb23cb1ffc8ab106e8302bec38ac7f1ccc5d898fbc03946e7d0eff11be59eb05

Observation a65020e1-9b5f-4763-be62-1b7958887b6b · outbound

This paper cites Depth map prediction from a single image using a multi-scale deep net- work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth map prediction from a single image using a multi-scale deep net- work

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.623717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.296951Z digest=sha256:b598daa6a4ca2e3cbe7dda5f664c22e9173afbd9f92f6f9d8269fc51adc9a498

Observation 330baaba-8581-4169-b767-c241eb71aedc · outbound

This paper cites Structure and content-guided video synthesis with diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Structure and content-guided video synthesis with diffusion models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.389610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.365317Z digest=sha256:0ed06ad7fb5475506467cd80ed00516ac50e59bf60a807c744117a378a784bcc

Observation 098ae50c-bc3d-4f47-a039-bf1d6b073d83 · outbound

This paper cites GeoWiz- ard: Unleashing the diffusion priors for 3d geometry estima- tion from a single image.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion GeoWiz- ard: Unleashing the diffusion priors for 3d geometry estima- tion from a single image

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.252183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.458937Z digest=sha256:b9a9b07ec4850a43930bbae407e83c1c9745547aff1e55861bac5e73957eef5b

Observation a57da66e-926d-49ab-91ad-d91136b814be · outbound

This paper cites Humans in 4D: Re- constructing and tracking humans with transformers.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Humans in 4D: Re- constructing and tracking humans with transformers

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.131541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.597289Z digest=sha256:9ec295cf40611b844ba2aecfb61fab146c69cd9a4acb5672955cfc3c5260101a

Observation a3564360-2e18-4929-ad3f-7e35594a9d46 · outbound

This paper cites High-fidelity 3d human digitization from single 2k resolution images.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion High-fidelity 3d human digitization from single 2k resolution images

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.024814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:22.692204Z digest=sha256:1f3b291898b6bc6719db690afe7e11b895e223832c45006ea4ed8bdd8a6fa22a

Observation 180f7715-6b33-47ed-a81d-1f40606278b3 · outbound

This paper cites Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.779637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.779637Z digest=sha256:f35a612d429d22a19a3e3f42c4d187106ecc7c045dae05505d200288ce2d20ee

Observation 862928a1-904f-4c57-acad-c5d91e36df0f · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.901882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.901882Z digest=sha256:07d29cf8dd8836d72325dde253fdfb7d3eecaac9e92d3d059c7dfc03eb8bb9ad

Observation 38c55ea5-a2ac-40b3-b5b4-e22544322618 · outbound

This paper cites Denoising diffu- sion probabilistic models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Denoising diffu- sion probabilistic models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.901721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:23.018296Z digest=sha256:f855b9aef30d31cc3f15095b1578bf05cd0ee0ca7011bbf44423ada218132db0

Observation ff12a275-d4bb-475c-b9d0-8c0273694b09 · outbound

This paper cites Video Diffusion Models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Video Diffusion Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.141059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.141059Z digest=sha256:9c8656ce786b57be7e63f69cba29d769645e6630171632ed6f76bcd499818d80

Observation 1d0655c6-1b64-4688-905f-55c08ea4372e · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.784968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:23.223405Z digest=sha256:553679e2e66220ebb2fcf671b42649fd2ed3eed9dc335194fc7959fbfc6bf0b4

Observation 4fe6de2f-4a3e-4eff-b9fd-a9a91d1f95da · outbound

This paper cites Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.363527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.363527Z digest=sha256:91ec64ff4a3f30b414352ebee0ce2dc421f6915f85cd3b0a14c128171cd6c7ab

Observation c9d314d8-7044-45ca-a973-8831b2525b77 · outbound

This paper cites DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.490758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.490758Z digest=sha256:f3575df5575c9be7877cb67f3668c794d56a731b39e992d598ada55bf3e4cf61

Observation 5f48e3d9-f51f-44ff-a2c0-eb692f198ae3 · outbound

This paper cites ARCH: Animatable reconstruction of clothed humans.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion ARCH: Animatable reconstruction of clothed humans

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.448315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:23.710031Z digest=sha256:e5ce38b3321124d8207883c9dd5957d5f494cb29aab8148d74dcfa72f3a6057a

Observation 3ce18d9f-1d23-4551-b295-da670be10b2d · outbound

This paper cites Flowformer: A transformer architecture for optical flow.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Flowformer: A transformer architecture for optical flow

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.215736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:23.844526Z digest=sha256:9936842d10de9fd9ba780e9408cbefe40a859244fb78a9abc88145cf764433eb

Observation 5e3d4528-7f7b-4b23-a769-4f5f35e85eb1 · outbound

This paper cites HumanRF: High-fidelity neural radiance fields for humans in motion.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion HumanRF: High-fidelity neural radiance fields for humans in motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.098110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:23.941445Z digest=sha256:fac416571e18cee10bfb5030e1720ae7946a9ebc1c192560aee69c3d1ec1d3dc

Observation 5edd3c3b-c62a-4836-bcb2-b0bc54bc82f6 · outbound

This paper cites Jafarian and H.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Jafarian and H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.950238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.038240Z digest=sha256:3661357486625d78de36d9c25d3250d534cf5e17a1e0f561678d077d70fbbecb

Observation 65ad9760-e6e4-4e9b-a895-13947f567f5e · outbound

This paper cites Learning high fidelity depths of dressed humans by watching social media dance videos.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Learning high fidelity depths of dressed humans by watching social media dance videos

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.671335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.158543Z digest=sha256:dda4e67463711dc984fbe1e4bc5533db2f08d1a311f4d57a1451c1b38de80f37

Observation 7ef05ac4-86e6-461b-9ca2-c88b95a586d6 · outbound

This paper cites Repurpos- ing diffusion-based image generators for monocular depth estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Repurpos- ing diffusion-based image generators for monocular depth estimation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.532733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.272479Z digest=sha256:f03341e3240842153e7f72876bb43f016677ec4b8a48d4031f2960c2197a93b0

Observation 1b188c7c-7d3d-441e-a9c9-9d1942134c17 · outbound

This paper cites Sapiens: Foundation for human vision mod- els.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Sapiens: Foundation for human vision mod- els

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.304940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.420233Z digest=sha256:c9daad15dde5e207c442f69bd0d4fd5a2d0494c4f1a80c7a002943ca58ee9b02

Observation ddab9752-2070-4210-beda-baa514076afb · outbound

This paper cites Auto-Encoding Variational Bayes.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Auto-Encoding Variational Bayes

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:24.519563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:24.519563Z digest=sha256:5a29ce7f577494644f1fd34ee69196b61cdddd338129171c700411aadf69d1cc

Observation 0eacca16-ca2e-49ed-9b1a-b7d76a46f35d · outbound

This paper cites Adam: A Method for Stochastic Optimization.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Adam: A Method for Stochastic Optimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:24.610161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:24.610161Z digest=sha256:dbeeab7f33e9cc8a37f7bc4de7277c547a2bcc63741981dde8fa823851a25cd3

Observation c02b8bbd-ff08-46f6-8dd8-0878a7f494a1 · outbound

This paper cites Robust consistent video depth estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Robust consistent video depth estimation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.168095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.706712Z digest=sha256:d0b3758ce3c38c8748bfc78007447397d1469d7f7229725d21cb2ab3e48f885a

Observation 9156ee3e-aa92-4401-ac86-7428000a56ec · outbound

This paper cites Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:24.778818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:24.778818Z digest=sha256:6a24bf4db57b6ddde4a8ec12c341082816f2d93342b28d90b79f0bd2efae5969

Observation aca88010-beee-4e50-8f11-1441357d59a2 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:35.989939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.879018Z digest=sha256:7a564a23d5da11aae889a843826ebd9003ada38385e1e0acc0259c4f4bc8d00f

Observation 83973606-f072-4174-b1ad-cadb94cd9cb4 · outbound

This paper cites Consistent video depth estimation.ACM Transactions on Graphics (TOG), 39(4), 2020.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Consistent video depth estimation.ACM Transactions on Graphics (TOG), 39(4), 2020

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:35.783342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:24.977142Z digest=sha256:d1180923b32289139e08d18e21b6fc248ffe9b8dee17f1d1691b5d92e479b568

Observation cbc13d4e-e859-4ed7-9445-24b28b22e216 · outbound

This paper cites Jewett, Simon Ven- shtain, Christopher Heilman, Yueh-Tung Chen, Sidi Fu, Mo- hamed Ezzeldin A.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Jewett, Simon Ven- shtain, Christopher Heilman, Yueh-Tung Chen, Sidi Fu, Mo- hamed Ezzeldin A

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:35.523612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.067043Z digest=sha256:30b7d6a093883c5c6af9d418cce01f17df76501016a8a147d7b855e65b44656d

Observation 867a13c8-b40c-4a9a-9d4a-e4b41e60cd48 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:35.350523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.114459Z digest=sha256:27bf4c7cf79f3b93a9ad04b2e9e35d88bddd5060b2ac368fccb567d6d60a1e66

Observation d4af9a8c-a7b2-47d6-97dc-f46fcd957eee · outbound

This paper cites UniDepth: Universal monocular metric depth estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion UniDepth: Universal monocular metric depth estimation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:35.140737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.183400Z digest=sha256:335acdf224ab0795b4af69ce1c9f24b6694f859c6c6aeafdc47681ed6a99ba94

Observation 2987710d-1973-4166-aa8a-dea10b3e31b0 · outbound

This paper cites SDXL: Improving latent diffusion models for high-resolution image synthesis.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion SDXL: Improving latent diffusion models for high-resolution image synthesis

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.925641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.231414Z digest=sha256:e6d6a0372144be4f0166a5c225331476fc9e7908343674634808126cd96a4b4b

Observation b8fc48b4-11c9-4495-9c15-5f78d9a96f20 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:25.277658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:25.277658Z digest=sha256:5a6dc0a22fa8e0c61c1835cf0c163290e3f7a58ff2d21b4529cdcabe70212871

Observation acec3087-4f57-44b1-b37e-effc5550f785 · outbound

This paper cites Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.651003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.340557Z digest=sha256:75a9ccc7d3c8ce974365d2c05609ef362ea2a580db2aaffa834321adbe134ed3

Observation 0ca20f74-3fb0-48dd-81a7-26e57ddd8f05 · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion High-resolution image syn- thesis with latent diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.524732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.393763Z digest=sha256:2a28342f5ac415f311e6894549007e12abb2e8c27bb3a8ceba62c560854b0a52

Observation 4f78e827-66ea-4b05-8b91-f2a5c145f163 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.315839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.463051Z digest=sha256:aac30e031063230628fce53ed18ffb135a9662ff26291026f88b405cd617d96b

Observation f860c84d-9f30-4a5c-aa7d-1be88a0bd5ae · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Photorealistic text-to-image diffusion models with deep language understanding

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.127227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.530959Z digest=sha256:01210e21ad122527b284a10edf28e03f7a77db424164ac43c2380a49e49744ba

Observation 862d2a01-f574-4b15-a755-2bb648e0e223 · outbound

This paper cites Pifu: Pixel-aligned implicit function for high-resolution clothed human digitiza- tion.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Pifu: Pixel-aligned implicit function for high-resolution clothed human digitiza- tion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.956629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.621719Z digest=sha256:e0fc4812187d8b1b64adefcc137ae7109708055d60bee9cf197eb72c2cc0b37b

Observation a0017d01-24ec-48c7-a2be-393dbf3feb68 · outbound

This paper cites Pifuhd: Multi-level pixel-aligned implicit function for high-resolution 3d human digitization.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Pifuhd: Multi-level pixel-aligned implicit function for high-resolution 3d human digitization

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.776720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.680799Z digest=sha256:5b4b5e8f91773e93eb981cc1e456610593b16509bc9b42f796a82f25297a4630

Observation 88961aa5-be61-49d3-ac5a-ce28397d65a6 · outbound

This paper cites Progressive distillation for fast sampling of diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Progressive distillation for fast sampling of diffusion models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.632255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.747364Z digest=sha256:afa11bc5378399050bdf730002d6d9b46edfecfd6880dead78ab74d4d8b140e5

Observation 98bcd4a8-55b4-47ab-b3eb-8e6b35804c24 · outbound

This paper cites X-Avatar: Ex- pressive human avatars.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion X-Avatar: Ex- pressive human avatars

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.395997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.834387Z digest=sha256:84e78dc01d794d5161e3d8016fe2e341f809523d9d2f8be30abcf5e2a6be12a4

Observation 1d8443e4-d83a-4c5a-85bd-205c0c06c2cc · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Deep unsupervised learning using nonequilibrium thermodynamics

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.116872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.919767Z digest=sha256:5db2a6f8e6a4602757edb963dcc35300d2e960c5db2c1ef2ef6af69f1f6a9fbe

Observation ff0dd391-6b33-425b-a5b4-453ced72beaf · outbound

This paper cites Diffusers: State-of-the-art diffu- sion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Diffusers: State-of-the-art diffu- sion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.825490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:25.970093Z digest=sha256:31863332c7e4cd41d9f39f0d111adca936ecd36f8b42723a1254d7220712209e

Observation 961a342c-79f9-4bc0-a26a-3397d5b7b854 · outbound

This paper cites Modelscope text-to-video technical report.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Modelscope text-to-video technical report

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.568013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.032794Z digest=sha256:ed4e68496b000891b193d0737973320b1f02eaac9a685faf0b16044828017bbd

Observation 58349eb9-8492-4cc6-9e65-5e30617d8781 · outbound

This paper cites 4D-DRESS: A 4d dataset of real-world human clothing with semantic annotations.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion 4D-DRESS: A 4d dataset of real-world human clothing with semantic annotations

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.308039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.071652Z digest=sha256:d888048f67c892f0f6087bf146604b39c9b988e92506211018fe85ba4c5dfd6b

Observation 1fa9faa1-24e3-409a-933b-90eb67cd7ed2 · outbound

This paper cites MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.160150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.160150Z digest=sha256:6e69b7c7d19c3956a5bd63bf56887f78240c86012bfe1457773cac58ad5aa6df

Observation be46c9aa-c0a1-438e-86d5-da82c6780477 · outbound

This paper cites Videocomposer: Compositional video synthesis with motion controllability.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Videocomposer: Compositional video synthesis with motion controllability

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.016949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.205797Z digest=sha256:2119e4a69daeaf749ae241b823b50abdd4bc302a6721101127a57be5b8f4928a

Observation dfeaf70c-1b4a-4219-b331-b41c37b6b0db · outbound

This paper cites Less is more: Consistent video depth estimation with masked frames modeling.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Less is more: Consistent video depth estimation with masked frames modeling

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:31.737807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.280443Z digest=sha256:5057989ef3526cba618fa55c0cb1b22db4862cd82b7d352c578632a0d845b668

Observation edd21a6f-b0df-4956-b331-04c9a679840d · outbound

This paper cites TRAM: Global trajectory and motion of 3d humans from in- the-wild videos.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion TRAM: Global trajectory and motion of 3d humans from in- the-wild videos

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:31.453658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.351944Z digest=sha256:f3488efb720fe1a3108dae7634b579d665b17da13f231d6b383d74dca4eb62a6

Observation 0d9b665b-ae44-48b1-9c76-2fcb8cd12e02 · outbound

This paper cites ICON: Implicit clothed humans obtained from nor- mals.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion ICON: Implicit clothed humans obtained from nor- mals

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:31.252652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.425490Z digest=sha256:d2b0f40f351ecb21df18fb8dd4e3dda053b8ca14b933519c7eb612d105680446

Observation dce71b8a-2f4a-43a4-9806-2ef4bf7cf110 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:31.012696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.491077Z digest=sha256:158efd72a5b3a32722eafee3e0d00c067c09dd00c3f99d0163cf0c0c823ea8dd

Observation fbfb37ba-2077-416a-9b13-f6c4e8fad643 · outbound

This paper cites Depth Any Video with Scalable Synthetic Data.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth Any Video with Scalable Synthetic Data

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.572903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.572903Z digest=sha256:e012f3391795a34b3606feef9af9c03fbdf57fd60e12944123a63e9b655e343a

Observation 907ed995-7673-44d8-b6c4-7662a0a9e360 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth anything: Unleashing the power of large-scale unlabeled data

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.804175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.634045Z digest=sha256:aee40b7d2e91481cd9d4c1a622daeb93e6bb6b7425bcc19447c6e751860f29dd

Observation 2fda88d2-6e6f-482a-ad69-38d26ce581be · outbound

This paper cites Depth Anything V2.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth Anything V2

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.713658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.713658Z digest=sha256:20e2072dfcc677d9edc4d40853a0f8c4049c8f20315e8d4d2214f57b18c35de1

Observation d8e321ed-141b-4bb9-b7bb-2a68a195f911 · outbound

This paper cites Metric3d: Towards zero-shot metric 3d prediction from a single image.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Metric3d: Towards zero-shot metric 3d prediction from a single image

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.592483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.790452Z digest=sha256:c45345278015e83bb35d47aa7f09511d5b9a6020420874ebd8d065560dfe4975

Observation 5e4588df-5554-48bc-9615-3cc5b313e01b · outbound

This paper cites Function4d: Real-time human vol- umetric capture from very sparse consumer rgbd sensors.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Function4d: Real-time human vol- umetric capture from very sparse consumer rgbd sensors

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.362228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.845197Z digest=sha256:dab8f110cfc6e226313eaa0d78bba533358e10f19ecc51eae8490654e7e38881

Observation 3cd973db-6364-460c-b8e1-6482f89fcd0b · outbound

This paper cites GLAMR: Global occlusion-aware human mesh recovery with dynamic cameras.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion GLAMR: Global occlusion-aware human mesh recovery with dynamic cameras

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.093061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:26.919244Z digest=sha256:ede76b2b893244dd436c76d6a5f7211d999fb39589905162924085a816991470

Observation d0ff54db-8de9-4f3a-984c-bd65f7edc42b · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Adding conditional control to text-to-image diffusion models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.972501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.972501Z digest=sha256:9a40cf0a697f2e75ccd41e52ff923dca0b2b68eac7f8b481d212eb3257024b57

Observation 5f650424-84f2-4430-97a0-a5a51ea0005b · outbound

This paper cites Adding conditional control to text-to-image diffusion models.ICCV,.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Adding conditional control to text-to-image diffusion models.ICCV,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:27.029876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:27.029876Z digest=sha256:f1903c62233b96fb23d68bdad7487c90e2e5a761186a13ea8236c21c1bd7e990

Observation 5697c2bc-bf4a-4209-8275-bec6e5e5dd87 · outbound

This paper cites IC-light github page, 2024.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion IC-light github page, 2024

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.954643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:27.085835Z digest=sha256:7fa5c1e88c5e5ce7d58ccafdcb340efac383e8f1ec348207a9fefa4259e94343

Observation c4cc89b7-8c1b-4a7f-a543-4af6c43f4f6a · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:27.197607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:27.197607Z digest=sha256:d51682e3382e4f2801537be76c85a8dbcfd3f12d4ee645dee73ad6a3e8951fa9

Observation 551c7ffe-eca2-4e39-9245-198e351bb707 · outbound

This paper cites Consistent depth of moving objects in video.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Consistent depth of moving objects in video

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.688301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:27.290642Z digest=sha256:92a35018b2d868919f60cbdcc3026345019f37f67c17be120878d2e0d5ddf222

Observation f2299730-b323-4bea-bfea-67b1ec70959b · outbound

This paper cites SIFU: Side- view conditioned implicit function for real-world usable clothed human reconstruction.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion SIFU: Side- view conditioned implicit function for real-world usable clothed human reconstruction

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.466161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:27.465300Z digest=sha256:eeb56133d26598b4a3d7b5475c8de90873c67b552bf23a47e6b00df8d61c7dc4

Observation 3f8c97b7-cf18-42a5-91c6-8db32c16e8c8 · outbound

This paper cites Bilateral refer- ence for high-resolution dichotomous image segmentation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Bilateral refer- ence for high-resolution dichotomous image segmentation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.220649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:27.622635Z digest=sha256:1b8f9a71c6bc589b1e4d170f8c3ad63709f0714e65c6c15eef1ae6872e8b0552

Observation 26fc7723-ac14-4cb9-9337-5a7c06e93f79 · outbound

This paper cites PaMIR: Parametric model-conditioned implicit representa- tion for image-based human reconstruction.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion PaMIR: Parametric model-conditioned implicit representa- tion for image-based human reconstruction

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.038110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:27.745112Z digest=sha256:99785dc51216f8a23cb10b64e85279f28eba15024792764fcac43f0036665566

Observation 4182cb87-e610-47f8-9822-d9ce8182df6b · outbound

This paper cites MagicVideo: Efficient Video Generation With Latent Diffusion Models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion MagicVideo: Efficient Video Generation With Latent Diffusion Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:27.891963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:27.891963Z digest=sha256:50daa73afe0f3b03f0b96943776188e76afcc09e25c6af57e9d610003a4a31c7

Observation 838da50b-0235-417d-b38b-897d9c48bdd3 · outbound

This paper cites “1q ř ktPK,dtPDt ››dt´ dgt t ››2 RMSEplogq: b 1řpKt““1q ř ktPK,dtPDt ››log dt´ log dgt t ››2 δă thr: 1řpKt““1q ř ktPK,dtPDt Kt.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion “1q ř ktPK,dtPDt ››dt´ dgt t ››2 RMSEplogq: b 1řpKt““1q ř ktPK,dtPDt ››log dt´ log dgt t ››2 δă thr: 1řpKt““1q ř ktPK,dtPDt Kt

Reference 77

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:59:28.872014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:28.006204Z digest=sha256:071c59b9d55f44aa9cb526c518e05b1788bd90f8f97d9660398c15e8341b55a1

Observation 1cc54783-db91-4d93-9ed1-e99a46b31234 · outbound

This paper cites 24 Figure S18.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion 24 Figure S18

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:28.563642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:28.175667Z digest=sha256:fe2ffd5211461d5b7f96c42b36e83bc1e7770f858c7128fa80b98c4de3a71fb5

Observation 3644b666-5d63-48a1-be02-28e9bf3e2bb4 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T12:59:37.639984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:59:23.588927Z digest=sha256:da2537648c854aaafda0618e8496cf007d36ef1dcd8875c7c89ff90a6820b2b9

Pith citing papers

Observation e42001aa-f210-455b-b801-30b8bb9a1ca1 · inbound

Sapiens2 cites this paper.

Sapiens2 GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:21:06.816106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T21:59:43.755956Z digest=sha256:b60a166e61fbe19039b9b096608a4f93e6bd48337750593bb5ec48512fac7644