Pith. sign in

Paper Citation Record · LEDGER

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents

As of 23 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 2 inbound Pith citation observations for arXiv:2507.17462.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.17462 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:53:38.496401Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T00:17:39.409452Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T00:25:09.923270Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ca5a907-6541-4edd-be98-86e26c94919c · outbound

This paper cites CACTI: A Framework for Scalable Multi-Task Multi-Scene Visual Imitation Learning.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents CACTI: A Framework for Scalable Multi-Task Multi-Scene Visual Imitation Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:31.294528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:31.294528Z digest=sha256:0d819356876ecf0af5524e55c30891437478d99e5631128b1f64c099c35e0ee5

Observation f0a1a531-f8d7-473f-a58b-97a808b6eff5 · outbound

This paper cites Scaling Robot Learning with Semantically Imagined Experience,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Scaling Robot Learning with Semantically Imagined Experience,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:46.735035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:31.361436Z digest=sha256:7d9ed80180077b403c0b77d84d10d906b7166a08c7a7008a285c24119f1e5169

Observation 8baeeaea-cb0f-4687-bc98-7a2727894b97 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:31.519794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:31.519794Z digest=sha256:c1b43ebb742fe0029b3a99814d7d42f3ecb13692fe8588812f1eecd21b586cc1

Observation 4be03cb1-ac15-4dd4-bb8c-6a463f158267 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents OpenVLA: An Open-Source Vision-Language-Action Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:31.611898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:31.611898Z digest=sha256:a5a3c3a0f5ed5e8a2b493856f459cc0d8c44438b0bf26643aee78d297d622795

Observation 04a71937-6bfb-4ee6-a0e0-869bd17dd796 · outbound

This paper cites MagicDrive: Street view generation with diverse 3d geometry control,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents MagicDrive: Street view generation with diverse 3d geometry control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:46.492290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:31.715847Z digest=sha256:009046ea8581cf15e27f3d80fac850fde1517ed29a354c19bf2674a1a4d9aa0c

Observation ad4d3104-934c-424c-a9f3-e97ebffce656 · outbound

This paper cites BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents BEVControl: Accurately Controlling Street-view Elements with Multi-perspective Consistency via BEV Sketch Layout

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:31.818402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:31.818402Z digest=sha256:cb088887bac8fd10000b742604e29211d707d513c56843baef68848f68d2d8e2

Observation 35592f22-efe7-47ff-92fd-c461a4bc71ed · outbound

This paper cites Street-view image generation from a bird’s-eye view layout,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Street-view image generation from a bird’s-eye view layout,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:46.224327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:31.939060Z digest=sha256:d8dd6c862ddff8e3aab1690f20a64940f72a8b6fafbe1ca9b63bf36f219d07c7

Observation 5ab518d0-e587-4160-9616-96361742f5ca · outbound

This paper cites Dragvideo: Interactive drag-style video editing,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Dragvideo: Interactive drag-style video editing,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:45.962585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:32.047960Z digest=sha256:c69daf7f32636442ce4e68a85f7544b1effca9285b3a3c269a26e464b6b1c208

Observation 26ed3f74-1906-4583-8acb-be8ce7bf7b73 · outbound

This paper cites Video-p2p: Video editing with cross-attention control,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Video-p2p: Video editing with cross-attention control,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:45.700618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:32.210605Z digest=sha256:45efcf68ac4fed05589c65041dbaccfc5f3c181b1a045d465e564a3430cae634

Observation 6c75b870-d27c-432c-bd5f-a7f2bfce9af3 · outbound

This paper cites Visual commonsense-aware representation network for video captioning,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Visual commonsense-aware representation network for video captioning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:45.383791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:32.355751Z digest=sha256:b22c5278e287767b960a276396adf0c78b5e211a83c6edd50d676a65ed31807d

Observation a17b54cc-3fe0-44e2-bea7-a313f0067b13 · outbound

This paper cites Diffusion model-based image editing: A survey,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Diffusion model-based image editing: A survey,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:45.135486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:32.484985Z digest=sha256:5efac51cf6b5d1f4d4cc3463e5217fed870ac5e50db8ff1823fad8bde5c2c044

Observation 858904f2-5e64-41b3-bb0c-63489df87bf6 · outbound

This paper cites Feditnet++: Few- shot editing of latent semantics in gan spaces with correlated attribute disentanglement,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Feditnet++: Few- shot editing of latent semantics in gan spaces with correlated attribute disentanglement,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:44.811434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:32.605859Z digest=sha256:29c1ecccb545e03b4756df12c7e3573faad087f4e77c2ad29ca8d4c72efd23b9

Observation de2cc8c5-818d-4812-82d7-c105f1d3360d · outbound

This paper cites Gaussctrl: Multi-view consistent text-driven 3d gaussian splatting edit- ing,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Gaussctrl: Multi-view consistent text-driven 3d gaussian splatting edit- ing,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:32.743292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:32.743292Z digest=sha256:7b0a7f4a4b3aa86c9dce774193b9b58884b64ade591476f5b21667fc7e50d925

Observation f33ea19d-8b14-464f-98c4-669954140c4f · outbound

This paper cites Efficient dynamic scene editing via 4d gaussian-based static-dynamic separation,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Efficient dynamic scene editing via 4d gaussian-based static-dynamic separation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:44.429108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:32.872349Z digest=sha256:a812f095fdba38fda3b5b3c6518e6e11b185ff06b03d000577b385eabf5f8b52

Observation 240dac32-42fa-4a83-95bf-d3b4d47eab96 · outbound

This paper cites Generating long videos of dynamic scenes,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Generating long videos of dynamic scenes,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:44.173267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:33.031861Z digest=sha256:5828bad5fb559289ba0ec8e857b5d4fe9e8a80c5a3a13e2a1ea72e45f771c00a

Observation b2344980-f9e9-4149-aac0-0f5ced496595 · outbound

This paper cites Storydiffusion: Consistent self-attention for long-range image and video generation,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Storydiffusion: Consistent self-attention for long-range image and video generation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:43.868743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:33.186072Z digest=sha256:91113ba4e63b42df86e1495b815da51b9bd888ec1597442a18605b1a104e1388

Observation fef81db5-5e53-49eb-85f1-025222e676f7 · outbound

This paper cites Evalcrafter: Benchmarking and evaluating large video generation models,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Evalcrafter: Benchmarking and evaluating large video generation models,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:43.588423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:33.288535Z digest=sha256:e89e00716091dd246a1f27c2ae32351408019e56a623666d5a191131586be01c

Observation d6105e95-80cd-4038-9ef7-ce23d0e3fe6c · outbound

This paper cites Towards long video understanding via fine-detailed video story generation,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Towards long video understanding via fine-detailed video story generation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:43.394123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:33.414870Z digest=sha256:dfbc7a4097c95a666323e95e3ce2e61344013525a3bfd111d6972ea94500cdf8

Observation b30c98a2-6965-4090-9e4b-9012ea9e36a0 · outbound

This paper cites Maskgwm: A generalizable driving world model with video mask reconstruction,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Maskgwm: A generalizable driving world model with video mask reconstruction,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:43.094305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:33.564446Z digest=sha256:6920ae8fab908fc5be01283998dd352a8f4b8030dce2d222c1e982113791d384

Observation 537a1492-6b2a-408f-880c-16d2b2a1bf8c · outbound

This paper cites A Survey on Long Video Generation: Challenges, Methods, and Prospects.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents A Survey on Long Video Generation: Challenges, Methods, and Prospects

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:33.702516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:33.702516Z digest=sha256:65efa4b4b9bbead27f04803ad939f30f1af08330f1d690d65b478a82e0d7b1bc

Observation 64896346-6477-4c74-a9cf-a30b6a39df53 · outbound

This paper cites Dall-e-bot: Introducing web- scale diffusion models to robotics,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Dall-e-bot: Introducing web- scale diffusion models to robotics,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:42.734812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:33.900587Z digest=sha256:3be17536b5cec54c343e2eae317c68d84fe331e97e9e6b2992c65a37d5d1f713

Observation ec05318f-4a89-4cac-a4b0-b51e7bc9083a · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:34.065833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:34.065833Z digest=sha256:4d9accb9bf2e68b25bd7938b6482e2db32dd3042656eb499e4a80512962a3621

Observation 7fe0b49a-3265-49fe-a6aa-54abf355c307 · outbound

This paper cites GenAug: Retargeting behaviors to unseen situations via Generative Augmentation.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents GenAug: Retargeting behaviors to unseen situations via Generative Augmentation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:34.194140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:34.194140Z digest=sha256:b459613d08482f60ca6320321b9b8605e236e8dc4026af0705ada2b64eca30bd

Observation e7569c55-ebc9-4a05-9667-47468ebe9be1 · outbound

This paper cites Semantically controllable augmentations for generalizable robot learning,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Semantically controllable augmentations for generalizable robot learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:34.333520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:34.333520Z digest=sha256:df6e581ed567f9a98d06d470a4c366acdb45dfe4fcd261cc6f10db8682fe47e9

Observation 3f9cb659-3d2d-4a89-9357-d4a00e51e36a · outbound

This paper cites Roboagent: Generalization and efficiency in robot manipu- lation via semantic augmentations and action chunking,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Roboagent: Generalization and efficiency in robot manipu- lation via semantic augmentations and action chunking,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:42.481792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:34.459950Z digest=sha256:c7d505da2c1a51e2aed387a363467e75a45277a113ed79bb58f31665006c5d4c

Observation f48df0bf-9d3b-4967-bb78-03bd870c2454 · outbound

This paper cites Segment anything,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Segment anything,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:34.586243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:34.586243Z digest=sha256:1d9a24349cafe93b976737940f1d3cbab85bba9c24e6bd68e6b2c507de2b0b7c

Observation e8b228d8-3a7f-428f-bc33-44f5c8b892d9 · outbound

This paper cites EnerVerse-AC: Envisioning Embodied Environments with Action Condition.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents EnerVerse-AC: Envisioning Embodied Environments with Action Condition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:34.708735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:34.708735Z digest=sha256:07a97fcfe01be64756d252bdd1984899f78af8c869c182d33d3aabf0cf10435e

Observation 27c73efe-2b8f-4912-8279-17ef1c9a34df · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Zero-1-to-3: Zero-shot one image to 3d object,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:42.247498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:34.849354Z digest=sha256:c9c266912c2abf2a73a8b3f201ccfe481501d4e9cb172d7886b33b29df0f42b8

Observation 0bce7733-0683-4bd4-b8a0-e45affa1d359 · outbound

This paper cites InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:34.980027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:34.980027Z digest=sha256:02878be128433f0f874dacd60744ff3f8927cbe941e37de6d001b275521963ab

Observation 1c2321a9-0043-42e0-8703-03a03eeb6977 · outbound

This paper cites 3D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents 3D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:35.168365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:35.168365Z digest=sha256:76ec9fa9236a7b5ae92b6f3c0a807d239c33a483495cd4a86528634985c1e6dc

Observation 2d9265b2-33db-4031-a976-f83f19887450 · outbound

This paper cites Dge: Direct gaussian 3d editing by consistent multi-view editing,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Dge: Direct gaussian 3d editing by consistent multi-view editing,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:42.001632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:35.339800Z digest=sha256:3d552642012a2012cd30c54b7f442fcb9ac14c758b35695eb172b8b520f4a385

Observation e8c810d4-e30d-4347-ae5c-2f106ca3aa20 · outbound

This paper cites Imfine: 3d inpainting via geometry-guided multi-view refinement,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Imfine: 3d inpainting via geometry-guided multi-view refinement,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:41.681720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:35.479138Z digest=sha256:f9d1eedfa264f6b3517a7a809c73432d2a613ba5fef4536433c816030b3588fc

Observation 822d2eb3-7671-400a-b957-cd6d7ef32a5d · outbound

This paper cites High- resolution image synthesis with latent diffusion models,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents High- resolution image synthesis with latent diffusion models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:35.647926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:35.647926Z digest=sha256:1f80fcffdeefbc28cddfaf945b2fcb77d0a4f2a78c11bf9c524110b69f01950d

Observation efaa2062-6f9e-416b-8b88-5f1e572d0c7a · outbound

This paper cites Towards language-driven video inpainting via multimodal large language models,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Towards language-driven video inpainting via multimodal large language models,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:41.373678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:35.820351Z digest=sha256:88fa0a7a629609dbb4b71756823d557871dfc619766f9352225aa62d28a19936

Observation 0d35a874-59eb-41ee-8366-83b33dab9d1c · outbound

This paper cites Brush2prompt: Contextual prompt generator for object inpainting,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Brush2prompt: Contextual prompt generator for object inpainting,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:41.065267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:35.926503Z digest=sha256:eda94bb2b9b7e1f5479eb071d19b1e1cdbe8b5e6b0fee1ab650b0b80a722fd2e

Observation 9e760139-0a9d-4a6d-9ac5-19d643a221d6 · outbound

This paper cites Sketch- guided image inpainting with partial discrete diffusion process,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Sketch- guided image inpainting with partial discrete diffusion process,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:40.819891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:36.052743Z digest=sha256:8ba16b001df66f2c574833541722537d3a823d987a101d2e014beb49d0f6888c

Observation 337a5171-982a-4350-842d-46761fed50a2 · outbound

This paper cites Few-shot image generation via style adaptation and content preservation,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Few-shot image generation via style adaptation and content preservation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:40.501753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:36.256605Z digest=sha256:09d02297fbf05042fb361800b837fd8a44dd1572c6318b6dbd9126e0e11c9251

Observation 2b5dfbe2-4686-4566-ac14-412938512fcb · outbound

This paper cites Learning transferable visual models from natural language supervision,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Learning transferable visual models from natural language supervision,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:36.417013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:36.417013Z digest=sha256:39ebbd3dbd114ac7d825bb83f260a28ecdddbe40b8a5dc610bf134f712336d98

Observation c1f65cef-63aa-4d4d-b954-2bf34b686539 · outbound

This paper cites Video diffusion models,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Video diffusion models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:36.559772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:36.559772Z digest=sha256:07c83e59a0105a37c9fbf8c329e93d53ebebdcce90699f8d268353c7df6e4cbe

Observation e59b2166-dec2-49ad-b81d-6e467a3a7ffb · outbound

This paper cites Enerverse: Envisioning embodied future space for robotics manipulation,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Enerverse: Envisioning embodied future space for robotics manipulation,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:36.744744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:36.744744Z digest=sha256:6b363210611aee21147ae0a2ce538338e68d6d1011586e29989ac99e1aab952b

Observation 02991ebf-d39f-4ce1-b936-b379fb0b1826 · outbound

This paper cites Epidiff: Enhancing multi-view synthesis via localized epipolar-constrained diffusion,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Epidiff: Enhancing multi-view synthesis via localized epipolar-constrained diffusion,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:40.178023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:36.907756Z digest=sha256:df324e784641401d05566f28b399fc54be66d45fc6b889ea669008e5ec5ec93b

Observation 4e7cb55a-f64f-4b26-b156-4c97ebcb9b65 · outbound

This paper cites Ar-diffusion: Asynchronous video generation with auto-regressive diffusion,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Ar-diffusion: Asynchronous video generation with auto-regressive diffusion,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:39.812003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:37.094974Z digest=sha256:c3ca62bf22f4631afa1d58da73f7d60c243d5d60cde6fdd7479f3657b5803939

Observation 475c8523-758a-435a-a4cc-b3631bffabf6 · outbound

This paper cites Progressive autoregressive video diffusion models,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Progressive autoregressive video diffusion models,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:39.489380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:37.295899Z digest=sha256:cd026caf0142c5e73c267e5eb538d2132b7441ad9827aa948e872375d307bc82

Observation 8a3c1759-7c0a-4c46-8052-9405e6541aa0 · outbound

This paper cites From slow bidirectional to fast autoregressive video diffusion models,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents From slow bidirectional to fast autoregressive video diffusion models,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:37.457637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:37.457637Z digest=sha256:99e517d3e531c009c8ce464f15523a59a3de7d4ed451216e0d72912713304525

Observation a207e586-b1bb-4667-bdb0-8ad1045082c9 · outbound

This paper cites Qwen2.5-VL Technical Report.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Qwen2.5-VL Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:37.599945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:37.599945Z digest=sha256:c20b8ff02555f4d02518e54a1700b3e25c69eb15a234cc7b6eb45ffe22968a40

Observation 72315267-80ee-42a1-b6e5-0598d6458ecc · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Step1X-Edit: A Practical Framework for General Image Editing

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:37.812216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:37.812216Z digest=sha256:70ccb6e88377f29dfa6b211c52d9099351b6eb2faefe80fa5173a0ccba07c68a

Observation 80f9a6bd-903f-46d4-82d4-275ddf05a4f6 · outbound

This paper cites Robotwin: Dual-arm robot benchmark with generative digital twins (early version),.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Robotwin: Dual-arm robot benchmark with generative digital twins (early version),

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:39.242888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:37.961048Z digest=sha256:8b7abf2cfd1435c630e712ae6f5351591481dfbbcd701c56384a25f81d6558a1

Observation 4b404dc2-4398-4132-8db1-459fd9fb345d · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:38.099075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:38.099075Z digest=sha256:95a32c382627f2336e7e8399c5b6ca9bdd95b90f9e7a29ceb46318f7ea0bbc5a

Observation cd782825-b3e2-4a52-a651-73a0e9b715f1 · outbound

This paper cites Schedule your edit: A simple yet effective diffu- sion noise schedule for image editing,.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Schedule your edit: A simple yet effective diffu- sion noise schedule for image editing,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:53:38.933412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T14:53:38.294649Z digest=sha256:4b9d360814940127993970dd9629bee9acee6e001f84b589aa660fc64304f62b

Observation b0ffbd48-09d1-4a61-98ee-25730bfa6101 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:38.496401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:38.496401Z digest=sha256:8d4c5f1dea16f5865a06b963408a5c2019b7a424b3cee5ea53fd8386651cf214

Pith citing papers

Observation 6fb304ea-b10f-4f1d-8fde-8c04a951d874 · inbound

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling cites this paper.

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:55:32.152811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-08T19:16:50.004359Z digest=sha256:a5e8bf3050e6089a9120f552447842e21b85c8693bd915e7927dd31528db368c

Observation 23da8182-92d7-4d43-a694-e877fb20b4a1 · inbound

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling cites this paper.

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling ERMV: Editing 4D Robotic Multi-view images to enhance embodied agents

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:25:09.924677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-01T00:17:39.409452Z digest=sha256:d1450728932f86d98d8af9c2d57743161548950ecd115c24d3cc9f7283261a2c