Pith. sign in

Paper Citation Record · LEDGER

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

As of 7 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 2 inbound Pith citation observations for arXiv:2507.13344.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13344 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:30:34.764898Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T00:22:55.503532Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T00:25:32.931076Z

Reference resolution

86 of 86 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8571a637-e73d-4f7f-855d-d5a9b3b2307b · outbound

This paper cites an unresolved cited work.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.420799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.420799Z digest=sha256:34cbd292d769988c8867ba1347e5fafe658833de4bc02dfe6dbd3c8bf7399aca

Observation b9608fa7-51de-47d9-aa6d-230bc712c26f · outbound

This paper cites Creation of 3d human avatar using kinect.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Creation of 3d human avatar using kinect

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.425120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.425120Z digest=sha256:7573aa04ab8d99ec988f7507442d941c66949cb6b3d1130d676e1ce89c91d028

Observation 37df16ba-3007-4cda-8f54-942195d81f19 · outbound

This paper cites Tc4d: Trajectory-conditioned text-to-4d generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Tc4d: Trajectory-conditioned text-to-4d generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.429227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.429227Z digest=sha256:4834ffbd60375ff87573a4c14840fb0f5202079098ca5476a2176848601fd2c7

Observation f1515e09-b1b1-4f20-827c-c2fa1ab62330 · outbound

This paper cites 4d-fy: Text-to-4d generation using hybrid score dis- tillation sampling.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4d-fy: Text-to-4d generation using hybrid score dis- tillation sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.433775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.433775Z digest=sha256:60df1ac722381fe27b3bb21a81b1d3b5c17f07e7627cb18f1a898233ddc27e99

Observation 8e6c181b-52a0-45f0-8d1e-2bbfcfc8f81a · outbound

This paper cites Detailed full-body reconstructions of moving peo- ple from monocular rgb-d sequences.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Detailed full-body reconstructions of moving peo- ple from monocular rgb-d sequences

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.437662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.437662Z digest=sha256:ae6b50872cfdda61629cfc6fe3e15ee77d83a81988503c7e3625f78bdc65b3d1

Observation 3e90e633-72a8-449c-a8d7-26581edcb1e1 · outbound

This paper cites Video generation models as world simulators.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Video generation models as world simulators

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.441532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.441532Z digest=sha256:8f4bcacc32d8be054019b9fefacdf46f1d52a93ec150383f3e595ebe0d74811e

Observation 7f9d07fd-e859-4e60-9c03-e96119de7221 · outbound

This paper cites Hexplane: A fast representa- tion for dynamic scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Hexplane: A fast representa- tion for dynamic scenes

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.445677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.445677Z digest=sha256:5e8d8f2c251a6d5abe5c504517757679a54b9a0a960c01d13840e2ee1e0b9b5f

Observation 35acfce3-5e25-4342-a3e4-6988038785cf · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.742244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.450060Z digest=sha256:1ac99f82ee7dbaba29fc7f33dba65378bf9709dff7c2209075e3ca7bfd7f4a69

Observation 5ed6c430-2746-45c9-8dbe-79c1b92c16ce · outbound

This paper cites MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.453665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.453665Z digest=sha256:d1bfc931b702418c226c4237cd8b7c40bdb461ea4916e6bdaf0fd56caa262305

Observation 41e765d0-2b3f-47a7-a50c-f944d2ee17b8 · outbound

This paper cites DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.457845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.457845Z digest=sha256:acffbe6ee805885c3c2614811ca556b4152ae486591f0d3fa6438388f1f6e96e

Observation 3487a70a-6586-4f65-94e6-11311f655dce · outbound

This paper cites High-quality streamable free-viewpoint video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models High-quality streamable free-viewpoint video

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.463175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.463175Z digest=sha256:39e1d5e58a2f3f2eeb6dbe04639a69598226131e692add1e023f23f019df4a5c

Observation 57b4a544-a3ba-4c54-b5c4-de162356275a · outbound

This paper cites Objaverse: A Universe of Annotated 3D Objects.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Objaverse: A Universe of Annotated 3D Objects

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.468225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.468225Z digest=sha256:e3086de0233f07d16fa03be213474cb65c72ad962f5f83601a0385fd80553bb4

Observation 07c93530-0bef-4680-9950-d5a12ea4c3b3 · outbound

This paper cites Objaverse-XL: A Universe of 10M+ 3D Objects.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Objaverse-XL: A Universe of 10M+ 3D Objects

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.473031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.473031Z digest=sha256:fc363167a626588dcadcc4d2e6c2d7492b8941d4489654644de1221ce3d2724d

Observation 3c9e237d-1edb-4b47-b239-9cba292a755c · outbound

This paper cites 4d-rotor gaussian splatting: towards efficient novel view synthesis for dynamic scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4d-rotor gaussian splatting: towards efficient novel view synthesis for dynamic scenes

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.478269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.478269Z digest=sha256:048a0d8f2b89e80d912333f4ac11ecac13883184d6ba9d1a350d8d50eef08999

Observation 63044d39-968e-4b00-8c94-4cb1d275358f · outbound

This paper cites Fast dynamic radiance fields with time-aware neural vox- els.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Fast dynamic radiance fields with time-aware neural vox- els

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.482394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.482394Z digest=sha256:c99a1701d798f0c957e8ab94697e209e3bf4f3d35d0efcd829d751dff881ab9a

Observation ebab4608-131e-46c1-b912-45a500bcd16f · outbound

This paper cites K-planes: Explicit radiance fields in space, time, and appearance.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models K-planes: Explicit radiance fields in space, time, and appearance

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.486116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.486116Z digest=sha256:cb7507aa8a6e67d8975d779f309d4cacba60aeff9cc13920ffe67e8d1fe39ac5

Observation 412dcaad-27c0-451c-ae6d-b8eb5798775e · outbound

This paper cites Massively parallel multiview stereopsis by surface normal diffusion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Massively parallel multiview stereopsis by surface normal diffusion

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.702801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.489747Z digest=sha256:b0583a0ea9c5401d9c7839326b98cb645bd3b7b6b027f5e0fe2156ad65752e34

Observation e5a64fc4-8fd2-41f3-b092-fa094ce8fcc8 · outbound

This paper cites Srinivasan, Jonathan T.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Srinivasan, Jonathan T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.691475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.493299Z digest=sha256:befa5c7e52bb8a0d08dfe22328847a6acab2303aa676b38a7bacbf6fb7d17615

Observation 7398098b-d54f-40be-98ac-d56d3e690634 · outbound

This paper cites Studio production system for dynamic 3d con- tent.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Studio production system for dynamic 3d con- tent

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.680019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.497112Z digest=sha256:7c282920635c8b172c53ab2b36df2f62d62ec5f15e0e71ee49c8c7f557660019

Observation 4a3206aa-25ad-4a9d-ba32-0d9fee30c9bf · outbound

This paper cites Viewdiff: 3d-consistent image generation with text-to-image models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Viewdiff: 3d-consistent image generation with text-to-image models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.500818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.500818Z digest=sha256:0321b804365af9ada9739ccb0c5bf48e7ed0341651e143c0131aca6d1d76089e

Observation 9ce25f9d-15f4-4aab-a6a4-28439adf213a · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.505019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.505019Z digest=sha256:f342279f0c17e93e9d786a418f1f060acfc417885cd85e83af7b13c7258bcc43

Observation c42cf2d4-c328-43b1-b609-b938e70a2715 · outbound

This paper cites Gauhuman: Articulated gaus- sian splatting from monocular human videos.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Gauhuman: Articulated gaus- sian splatting from monocular human videos

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.658452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.509099Z digest=sha256:4d89005a35bb9bfc967bba391cd3cc6ec0145eaeb611cdef8f34634acc6712fe

Observation 82d4f14e-363a-4454-b731-3c1f90dfd63c · outbound

This paper cites Humanrf: High-fidelity neural radiance fields for humans in motion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Humanrf: High-fidelity neural radiance fields for humans in motion

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.641401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.513178Z digest=sha256:0421eed30708a8eba07d5ff003affc3847ee48411b7ce6d3157f90e1aee3d18b

Observation c329a7ff-e1fd-4c29-91fd-4d5bde5f8707 · outbound

This paper cites Pl ¨ucker coordinates for lines in the space.Prob- lem Solver Techniques for Applied Computer Science, Com- S-477/577 Course Handout, 3, 2020.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Pl ¨ucker coordinates for lines in the space.Prob- lem Solver Techniques for Applied Computer Science, Com- S-477/577 Course Handout, 3, 2020

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.628383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.517297Z digest=sha256:6042a3778a55247041bcc09bbbae2b9de340997c3d2038447da8c9670afa34e1

Observation c1115a64-0404-4544-b732-75954c1545f3 · outbound

This paper cites Virtualized reality: Constructing virtual worlds from real scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Virtualized reality: Constructing virtual worlds from real scenes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.616211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.521154Z digest=sha256:1e793860c63e47d45fbb1dfa71dee5b41de910593c9a982d7b5ea033ea858caa

Observation 819f7afa-a1b7-40c0-b24d-b68a3d82104b · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 3d gaussian splatting for real-time radiance field rendering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.603769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.524599Z digest=sha256:e7254ade9b7057e991cacdee9866a37aa394019db5bc0e0554722d3c46dd7737

Observation 6ef0743c-0968-4fe6-b29d-cadef456ac28 · outbound

This paper cites Sapiens: Foundation for Human Vision Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Sapiens: Foundation for Human Vision Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.528734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.528734Z digest=sha256:c1c081a65192f0dae3267e47c06d6880a1e8dd6e3b8a418200fa9da11cc773c6

Observation 00ad95cc-d284-4256-8955-d3ed759e87b9 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Adam: A Method for Stochastic Optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.532489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.532489Z digest=sha256:86e2068167a0849a62895bed55365ee92b9cee310f409bcc801569d626f4e341

Observation 6e6b6568-9bab-4190-9aec-bc1f1ead95c3 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.536274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.536274Z digest=sha256:d7bee4ff602460a4ed7f6fb03f6fb4413923fc143c7276980ec2ccc4f2977237

Observation 770f6a6c-308c-4c7a-973c-45dc1bdeb87a · outbound

This paper cites A theory of shape by space carving.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models A theory of shape by space carving

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.590851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.540369Z digest=sha256:850c023a870eea73ce9842eb3f2e55e73dca8e232663269d33a72f3c7669bf44

Observation aa7ce023-2622-481b-b80b-4f6d67963812 · outbound

This paper cites Vivid-zoo: Multi-view video generation with diffusion model.Advances in Neural Information Processing Systems, 37:62189–62222,.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Vivid-zoo: Multi-view video generation with diffusion model.Advances in Neural Information Processing Systems, 37:62189–62222,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.577477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.543908Z digest=sha256:4c87fa9c8ba28a6b56cd03930d78a7ca601ce143dcce36ebdb049a1cd2e10656

Observation 29b26c72-bd8d-49b8-b107-93cf2560d301 · outbound

This paper cites Neural 3d video synthesis from multi-view video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Neural 3d video synthesis from multi-view video

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.547337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.547337Z digest=sha256:245d45709db2c3ffee3ab7d2118d5eb781b8cd083f2850d91f81321997528778

Observation 84bb4f8a-9083-4fc6-87f4-9c849252f5c4 · outbound

This paper cites Ani- matable gaussians: Learning pose-dependent gaussian maps for high-fidelity human avatar modeling.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Ani- matable gaussians: Learning pose-dependent gaussian maps for high-fidelity human avatar modeling

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.557675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.551854Z digest=sha256:4560607c7b1c1d5f845da1b5d5eede9be5f78c5a0cd050cf8799df139918b52e

Observation cd2ac44b-3338-4bf8-ae12-f587abed1008 · outbound

This paper cites Efficient neural radiance fields for interactive free-viewpoint video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Efficient neural radiance fields for interactive free-viewpoint video

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.545973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.555847Z digest=sha256:061d456851eaabcef15b001115fe6f27d42b66d510e94d577d04d9b19591622f

Observation 7df6b6d9-b38f-4d19-bda8-baf082d481d9 · outbound

This paper cites Real-time high-resolution background matting.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Real-time high-resolution background matting

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.534476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.560525Z digest=sha256:efa22340c23b13750bd6a83314335c3a2f07179c8a3d3b6235dd3593567e53d3

Observation 4e1b59c8-bc22-4328-8b9e-4d310fc6d246 · outbound

This paper cites Raft-stereo: Multilevel recurrent field transforms for stereo matching.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Raft-stereo: Multilevel recurrent field transforms for stereo matching

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.564842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.564842Z digest=sha256:7e7eaaf1bb63a75d5087c9348fd7e81017ce057f4d9e3e9598da7547e24112f2

Observation eca86fea-97ca-430e-a778-5bd85a7cc633 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object, 2023.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Zero-1-to-3: Zero-shot one image to 3d object, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.514838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.569260Z digest=sha256:7efa12e67cb827121612c9c7e8269e19faea7c04957f09c79458ef46277ec8fb

Observation ca18d8f5-57fc-4cae-a43b-475f194e0e6e · outbound

This paper cites Mvsgaussian: Fast generalizable gaussian splatting recon- struction from multi-view stereo.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Mvsgaussian: Fast generalizable gaussian splatting recon- struction from multi-view stereo

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.502262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.574009Z digest=sha256:1dd535c272b0dce5ce78a52889a8fd618c628a7d73305721dbd46c302c059949

Observation b85b8ff9-775d-4be5-a5dc-5a943bc29cae · outbound

This paper cites SyncDreamer: Generating Multiview-consistent Images from a Single-view Image.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.578276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.578276Z digest=sha256:2f3d1ab4782d5b39c4a7b0a2b51bb75f93844a178636bffe815458de0617ad70

Observation 6df977f3-65a2-4882-9929-0db30d540cd6 · outbound

This paper cites Smpl: A skinned multi- person linear model.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Smpl: A skinned multi- person linear model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.582315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.582315Z digest=sha256:7cbf1a16e6e0ee0cda4de9d7290953ca233df579746f6386ae7a7501fc55f017

Observation 6df2aee7-a11a-408d-86ed-6d762fd7be18 · outbound

This paper cites DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.586204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.586204Z digest=sha256:fe9f08579c9891995c8adb61f6382efb20855ab1394b89da9fb58ff7bad2d895

Observation 3da23ece-d28f-44a0-b728-6337a9acfe48 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.590395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.590395Z digest=sha256:b292f9adf5c55ed343c41991b954f9f9555b7249a8dc808d8e749c4c68eacdfa

Observation e8523673-e024-4775-b9a8-779c17c5f045 · outbound

This paper cites Dynamicfusion: Reconstruction and tracking of non-rigid scenes in real-time.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Dynamicfusion: Reconstruction and tracking of non-rigid scenes in real-time

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.475108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.594323Z digest=sha256:1f1aaae231ea8105d5b09306ef3fc682a1f23c42ee7b8cc39bff6320ae8bb4d6

Observation d8eee92d-9d03-4e63-8bf0-ac7a052f6fc8 · outbound

This paper cites Effi- cient4d: Fast dynamic 3d object generation from a single- view video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Effi- cient4d: Fast dynamic 3d object generation from a single- view video

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.598925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.598925Z digest=sha256:10975ab4e232e692742a6e6fc5f037f78670fa31742b81a1147d77d96b9cff05

Observation 2b4724ea-e076-402b-a305-ebb61c952642 · outbound

This paper cites Barron, Sofien Bouaziz, Dan B Goldman, Steven M.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Barron, Sofien Bouaziz, Dan B Goldman, Steven M

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.602436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.602436Z digest=sha256:765dfa8c57a43c23a6196e2218ac03eb69dc71731a07d818bbcec041aa7d3ac7

Observation fbe7732f-c675-4c74-ba8d-535217bf7715 · outbound

This paper cites Barron, Sofien Bouaziz, Dan B Goldman, Ricardo Martin- Brualla, and Steven M.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Barron, Sofien Bouaziz, Dan B Goldman, Ricardo Martin- Brualla, and Steven M

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.456265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.605845Z digest=sha256:80db2a260185295f0feca72717b1ee494cb1d50dbcef98d284de480434de18b5

Observation 0d08a2b5-a4f7-48c2-9398-b13573a35f82 · outbound

This paper cites Ani- matable neural radiance fields for modeling dynamic human bodies.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Ani- matable neural radiance fields for modeling dynamic human bodies

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.444141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.609625Z digest=sha256:f0e39c3d828ac7755a626f5e11331b5d3f214c10e562c923dbab40df72a92a01

Observation da7146fd-6d79-4243-85e4-db4e97ea0cba · outbound

This paper cites Neural body: Implicit neural representations with structured latent codes for novel view synthesis of dynamic humans.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Neural body: Implicit neural representations with structured latent codes for novel view synthesis of dynamic humans

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.613234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.613234Z digest=sha256:50ab1eedc534f75d378da2415601b1ace23fd64a4ffe291117eae292ead57c1c

Observation d75891c5-45a7-4007-b204-f625daac911a · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DreamFusion: Text-to-3D using 2D Diffusion

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.616681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.616681Z digest=sha256:42b648212a3ed294d3025f850470be91b60df7b7355adbcebc0beb089b9dc8c4

Observation 6c6535d4-9fcf-44f8-896c-5510a5e0cb09 · outbound

This paper cites D-NeRF: Neural Radiance Fields for Dynamic Scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models D-NeRF: Neural Radiance Fields for Dynamic Scenes

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.424968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.620537Z digest=sha256:9f42bbae141d2f597d28be48f734738b7e6ee41e4a9a12d0b7f322f450b727a4

Observation d88181dc-54f9-44c7-ba87-cac7c5bec71f · outbound

This paper cites DreamGaussian4D: Generative 4D Gaussian Splatting.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DreamGaussian4D: Generative 4D Gaussian Splatting

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.624666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.624666Z digest=sha256:c2c60eb1f06698e6ceaeaad75f86b26b14251612a0f1dc81ab944fc718a3b3a0

Observation 6fb2c1b8-9158-42b5-873e-d34c29d1f12b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.410653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.628519Z digest=sha256:14f6d566b44ddce2d49e344ec2e310b0b2367fc75deca22e19d2aa11ef3a0a7a

Observation 7c2900e9-3051-4dd5-8801-8eeac70d29f8 · outbound

This paper cites Structure-from-motion revisited.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Structure-from-motion revisited

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.397435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.632269Z digest=sha256:a1086cee0c686f9cc5a0ab9cf33fa30e9a356fb55d534e542773ebde15dd6917

Observation 85467ce6-6700-4754-a610-6ba810eae861 · outbound

This paper cites Pixelwise view selection for un- structured multi-view stereo.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Pixelwise view selection for un- structured multi-view stereo

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.386295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.636478Z digest=sha256:d8c91f264a9ffc2141768a3e445b86b20f48733f592c2b69a07e0377f557dad0

Observation 109f7b72-5c41-49ff-8fed-6b6d6dc0bb7f · outbound

This paper cites Tensor4d: Efficient neural 4d decomposition for high-fidelity dynamic reconstruction and rendering.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Tensor4d: Efficient neural 4d decomposition for high-fidelity dynamic reconstruction and rendering

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.640007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.640007Z digest=sha256:db884126249fd0cb955166e3470af314dc7ba0b895bcf664d6c5c6e5343f68f1

Observation 36f7a853-d23e-4e47-96e4-e7f94e51d0c4 · outbound

This paper cites Rapid avatar capture and simulation using commodity depth sensors.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Rapid avatar capture and simulation using commodity depth sensors

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.367216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.643880Z digest=sha256:c4e80923e70d516c49e1781aadc4a40eca3f26122ed3e47bdc6dcbca4e4600cb

Observation c8afb796-2adf-4e6a-8e56-566a988e6d4c · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models MVDream: Multi-view Diffusion for 3D Generation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.647825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.647825Z digest=sha256:a88928487e3c658c3bbdd5f003e0ccb06fc8e4f921b63086c5cbf33f41ead3b6

Observation d67ffa65-560f-45f4-8755-0c07a3b4cf69 · outbound

This paper cites Text-To-4D Dynamic Scene Generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Text-To-4D Dynamic Scene Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.652426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.652426Z digest=sha256:24d8e1d3f6ee1ab931f37847b89bf5fe9d9cbe21d4483bf8d84e0cdc29b2f489

Observation e81720c4-0c83-4348-a42b-d4935844e092 · outbound

This paper cites Virtual view synthesis of people from multiple view video sequences.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Virtual view synthesis of people from multiple view video sequences

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.355417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.656745Z digest=sha256:a3992cd1ff5a082662e05d8f324c0f445b3925d9c05c570e637c1a4d1d3c3a83

Observation 5c534f71-c523-48a8-a77a-d6b5d7159a72 · outbound

This paper cites Surface capture for performance-based animation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Surface capture for performance-based animation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.343986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.660850Z digest=sha256:baaea37ef5221d99b0bbcd7eaaafd49098e622d634a8db43434d05e360f88cb2

Observation fb674737-27b8-4ded-ae19-96bf00061333 · outbound

This paper cites Scanning 3d full human bodies using kinects.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Scanning 3d full human bodies using kinects

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.332853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.664462Z digest=sha256:6e3cfed87fd13111615285e764a4077294e0fb417fc0e12039dfba40844fb9af

Observation 37ba4bb2-3918-4330-87f1-b3a0e036720c · outbound

This paper cites Diffusers: State-of-the-art diffu- sion models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Diffusers: State-of-the-art diffu- sion models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.668115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.668115Z digest=sha256:b0ee1fa1fe7829e80984a1c2e285082568e9d4159aaff82b1ee41a0afd9873f8

Observation 34eab7ac-c5f1-4ba1-99f3-005bee505737 · outbound

This paper cites Fourier plenoctrees for dynamic radiance field ren- dering in real-time.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Fourier plenoctrees for dynamic radiance field ren- dering in real-time

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.312840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.672266Z digest=sha256:fd7ce3b74aaea26a26c31fff5ec8246506dbb8b260cafc8570b63b0d62498d8e

Observation adc958c2-78f9-4542-8d6a-645e4139ff0d · outbound

This paper cites ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.675972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.675972Z digest=sha256:ed70414ba57fba0b9c083182dd87353efcd1ba82e47fd084ed83438b67cea69c

Observation 59f7878a-1b95-44ba-a5b8-a5049e69896e · outbound

This paper cites Controlling Space and Time with Diffusion Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Controlling Space and Time with Diffusion Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.679622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.679622Z digest=sha256:a3b008dae65ebbcc6d1abb4af686b08dc4f0cee7102e0abc87162df5a289cff4

Observation bc604099-22a5-4753-8bc1-3224b8a45be5 · outbound

This paper cites Hu- mannerf: Free-viewpoint rendering of moving people from monocular video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Hu- mannerf: Free-viewpoint rendering of moving people from monocular video

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.683825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.683825Z digest=sha256:83692b684e028edc2b07ab22db4845576e8c04ac3749aa4f29faf4cc03625e72

Observation 08c4c103-43fd-4a18-884c-7c2d26f98bb0 · outbound

This paper cites 4d gaussian splatting for real-time dynamic scene render- ing.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4d gaussian splatting for real-time dynamic scene render- ing

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.294926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.688073Z digest=sha256:91f5435bc97c9ce1110cbdb714e7d224fff8cbbd2c156f68f432d553a68b0e6c

Observation 791357f3-5e62-4e85-b0e5-dd097bf68ffb · outbound

This paper cites CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.692191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.692191Z digest=sha256:8db5eaa1e2d0b9409a4f1a6410692bb1dca9202fdd93d6cf855b9fac8d971f26

Observation 185d1a01-d5b7-4354-a110-2222020e9d07 · outbound

This paper cites SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.696845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.696845Z digest=sha256:176d7dcfbd331b42335d18f2129125cf4bca8cbb2b00b4cc1ee1d76517a9bb6f

Observation 39c07d01-d3e0-4685-b97b-b45c7187bd69 · outbound

This paper cites Easyvolcap: Accelerating neural volumetric video research.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Easyvolcap: Accelerating neural volumetric video research

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.283732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.700951Z digest=sha256:8158eb0b17647e475e63c16b316d787e04589bbacdb8b18f4be74073ce135120

Observation a70031ba-c695-4a77-9dfc-fa5e4c2d674a · outbound

This paper cites Relightable and animatable neural avatar from sparse-view video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Relightable and animatable neural avatar from sparse-view video

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.704528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.704528Z digest=sha256:58fb56857aa44adb8343950c833f5c58909768dda13a764ef5047df1a0ae009b

Observation 3a52ecad-f94e-4650-b2d9-c9b2d96a5b26 · outbound

This paper cites 4k4d: Real-time 4d view synthesis at 4k resolution.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4k4d: Real-time 4d view synthesis at 4k resolution

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.264293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.708737Z digest=sha256:1dfc7565326d28ad0aceeb8ff252f573229205f3e7d544b04062db1faaef31ab

Observation b428c9b2-58f8-413f-9c12-45a569c91d1e · outbound

This paper cites Representing long volumet- ric video with temporal gaussian hierarchy.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Representing long volumet- ric video with temporal gaussian hierarchy

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.252346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.712365Z digest=sha256:5dd2f03e47a86014885e062d3c98a9f091f294e727b6f2fd31789204e84a359e

Observation eb6042ec-0c17-4911-83d9-1702e3ca5368 · outbound

This paper cites Diffusion2: Dynamic 3d content generation via score composition of orthogonal diffusion models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Diffusion2: Dynamic 3d content generation via score composition of orthogonal diffusion models

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.240943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.715946Z digest=sha256:f9cf7a8afcb1681712b67b4bc59b2a7f8a346ee088be0fc5422b40b073418d55

Observation 185457a4-e58a-42e4-8b12-9ab04db6eff9 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.720259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.720259Z digest=sha256:e73f746f34ba39f10653f1f4610a47e0acaba97bdc2878b726d56dc5f8bca21c

Observation 4ed2c0d3-3c54-404f-bce2-82d27be50d6f · outbound

This paper cites Real- time photorealistic dynamic scene representation and render- ing with 4d gaussian splatting.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Real- time photorealistic dynamic scene representation and render- ing with 4d gaussian splatting

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.226367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.725313Z digest=sha256:6fa1f54ba43d9c8d710ccb4232b912dcd7337dc600b9ee46cf9405f616f20603

Observation 68e6eebc-89f4-4db3-a0b3-ed439169d524 · outbound

This paper cites Mvsnet: Depth inference for unstructured multi-view stereo.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Mvsnet: Depth inference for unstructured multi-view stereo

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.212285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.728762Z digest=sha256:9fa2c5e56c17d98d5109bc7259522c706f741107ec2c329594c37e9f4d388b96

Observation 3c1f007c-f168-4c06-9255-ed07bb53c4fd · outbound

This paper cites Recurrent mvsnet for high-resolution multi-view stereo depth inference.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Recurrent mvsnet for high-resolution multi-view stereo depth inference

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.200000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.732702Z digest=sha256:1359c6b9934deb7bd1be3e0107a56a10cea8610c605e1695452b33eb8a96dd26

Observation 88bc8735-bdfe-495f-9fb7-1a061878c8eb · outbound

This paper cites 4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.736671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.736671Z digest=sha256:2bf2931b24bb4d6059d0a33f44dd6b986a024a8bd5701d88a9977bf96eb52e31

Observation 3ddff97b-ce40-4da3-a71a-cb32d626f213 · outbound

This paper cites Stag4d: Spatial-temporal anchored generative 4d gaussians.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Stag4d: Spatial-temporal anchored generative 4d gaussians

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.188208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.741102Z digest=sha256:1ae8fc289b231784f53bea5c72149c26ef9982c4dcff49210dc02906c8a6ec96

Observation e2d7be86-4861-458a-85c0-c2dde556a466 · outbound

This paper cites 4diffusion: Multi-view video dif- fusion model for 4d generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4diffusion: Multi-view video dif- fusion model for 4d generation

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.175996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.745262Z digest=sha256:a04337f35a814380c8e9fe18c904cdeade5ab4088a7b055d20dc3df4dc39a7f2

Observation 61ec5bc2-708c-44ce-b4ce-4c3fa398894a · outbound

This paper cites Cameras as Rays: Pose Estimation via Ray Diffusion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Cameras as Rays: Pose Estimation via Ray Diffusion

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.749251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.749251Z digest=sha256:b1596656f308a25566fac1c0d5e28b036f525bf0f466993e810c94f6cca37242

Observation f84d5aed-3a19-4720-abed-d554226d3f24 · outbound

This paper cites Animate124: Animating One Image to 4D Dynamic Scene.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Animate124: Animating One Image to 4D Dynamic Scene

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.753710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.753710Z digest=sha256:b911ab454a09da62e81da65ecc635bfe442d88545636bc2f7078b75ba5c74a63

Observation 13578719-044e-47d8-8176-daf2a4d380f8 · outbound

This paper cites Bilateral refer- ence for high-resolution dichotomous image segmentation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Bilateral refer- ence for high-resolution dichotomous image segmentation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.757492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.757492Z digest=sha256:57843ebf0e403ff4a4111364f9447f8236831c98eff3a8179599e20603dbdef3

Observation 3231362c-a61f-4779-aab8-491d4ca0f60e · outbound

This paper cites Gps- gaussian: Generalizable pixel-wise 3d gaussian splatting for real-time human novel view synthesis.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Gps- gaussian: Generalizable pixel-wise 3d gaussian splatting for real-time human novel view synthesis

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.156419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.761142Z digest=sha256:0380c049bd2638e635d82cdde8c769bbe64a728551774490feac48efe4f3d578

Observation 57c8a991-f2bd-4308-b0ac-7061bb2d9260 · outbound

This paper cites A unified approach for text- and image-guided 4d scene generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models A unified approach for text- and image-guided 4d scene generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.140720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:30:34.764898Z digest=sha256:9207278bd7d93e91f69ef1db075e4aafee3e164d368632bcd858fcd62fa88b04

Pith citing papers

Observation 1cc16329-431c-4af8-8058-054d349c6e10 · inbound

Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges cites this paper.

Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:25:32.934331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T00:22:55.503532Z digest=sha256:99bf1e70e5ad9f639e6e376a7255e36475a66bc3f318c010cbc2f6b4b825a002

Observation 026f7edc-e9a8-445e-ad1e-15107d66761c · inbound

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras cites this paper.

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:43:18.026626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T23:41:24.423859Z digest=sha256:187f427570f5e2679ad78733e9be73bbccc713fa40e23b7b4d42a73a7d48a77a