Pith. sign in

Paper Citation Record · LEDGER

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech

As of 17 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2412.11409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11409 v3

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:01:33.912139Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:34:07.907336Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T04:34:07.970750Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c705524e-f197-4a1d-939c-0c3fa5483dc6 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.739038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.739038Z digest=sha256:df92117e9edfc23ed64f055b13bc71e4dc330c15d3402e44dafe7ad99f9f5853

Observation 9e631196-c032-4965-bc13-04bbe0c34f52 · outbound

This paper cites write newline.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.744102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.744102Z digest=sha256:adb5ad048a89f6fc860d7a1a053e8020aacd63091db021453d1d668b17546097

Observation 5317d04d-bc0d-425d-83a1-b2e1a780c7e3 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.462181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.749585Z digest=sha256:19aacdc1397f6404e7e993c892b387127a77803d3a9af6cc142d687b6dd9eeaa

Observation ffb4a67a-d664-4158-8575-be662b556a44 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.446298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.754047Z digest=sha256:b2efb97f2d1aa8cec7dba9e53a501da5f3c22617dabdb799342a1dea07c330c5

Observation acb71e47-d6c8-4f24-bbca-d13450814309 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.431945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.759161Z digest=sha256:f894f8618777640ca214fa6b758e7260b2969aa652f59f8a9f3de37a4bbf7df4

Observation a79ebfa7-0d48-4c24-a165-d8486079fca1 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.416672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.763741Z digest=sha256:59ec9f70266d5fcb3a0fd03549d03a907121c570f0878d88143987c08b7ab5bf

Observation 67654fb8-28d2-4741-824e-0d5c02c4aead · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.399707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.767991Z digest=sha256:6ae557da058e5803d3cb966f32fab28d9deb5abc5794ab262c87cca9f72209c3

Observation daace53f-7531-49e7-a731-8ac77f4de9ac · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.384054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.772569Z digest=sha256:f737cbbb19101705e7650167f38baeb2375eebb0692564b6b22cc75af700cdec

Observation 6afcc842-ea5b-4c2d-8d9f-943ec20cf967 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.776974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.776974Z digest=sha256:15f9d24a9a18987f6456ace5dfe54553461a2066d47b0282011f36e8c6f9ceb0

Observation 1b34d21b-ea50-411c-9aee-745b07d2dab8 · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.780829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.780829Z digest=sha256:e2868f0c69bf5ef9b8d96249101733f07ca4d837f2cdb90c89eec419dd1859c7

Observation e9b42dad-0124-4f16-8954-0d31983c5599 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.358773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.786306Z digest=sha256:2477ff920f5cd8f072b8ed6ced3d9f2058af6f058ef4b6e0c30e5e35e98b268d

Observation 9ff317ad-9929-4402-a93f-f0be088005fe · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.791200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.791200Z digest=sha256:3dadaf0895072e526fd1591ec18788ce41b25ff9dfa59c864520e81f024e319f

Observation 5ff458eb-6d03-4358-9d4f-977c5d726824 · outbound

This paper cites Multi-Source Spatial Knowledge Understanding for Immersive Visual Text-to-Speech.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Multi-Source Spatial Knowledge Understanding for Immersive Visual Text-to-Speech

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.795919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.795919Z digest=sha256:c64f21a4bbaf59154b9a08c129060234e23646cfe706b00f92bc70e025e67b76

Observation 3947395a-d4de-42a0-bf82-50dbe3d93928 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.334532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.800794Z digest=sha256:9e6830ed7efcc9377f07607757f8afa4bb6035023614bbb44212344fb24a724f

Observation 2001f691-e00a-460b-b8fe-fb14d7720ec7 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.320257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.805342Z digest=sha256:9aacca4331c0dfdcb8695b28b56ff91b18c802970157fa74f942b6392cb42f04

Observation 1be592c6-f85d-4895-89d2-5d38ccae6684 · outbound

This paper cites I Want to Figure Things Out.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech I Want to Figure Things Out

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.306395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.810025Z digest=sha256:173f30fa24b8f3cc0d5d126de47e31ea60f70eadd0e86feb72994461327c8191

Observation 9574b814-ab1e-4dec-b20f-7b5d18629c11 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.292349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.814503Z digest=sha256:77c471a31a0b3445881786367c7e640858c23be5c2f100bd880dca4c130392fb

Observation dd355004-bb8f-4cbb-8192-a963e173c590 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.279446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.819181Z digest=sha256:df72fc397ba2e43302f9d148f42a53cfcb25e005185d3abb800900231f3cdf56

Observation 40ae4745-caec-445a-a88f-6ae1dd587d2c · outbound

This paper cites BigVGAN: A Universal Neural Vocoder with Large-Scale Training.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech BigVGAN: A Universal Neural Vocoder with Large-Scale Training

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.823743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.823743Z digest=sha256:fcd3c75d5485c2a4f128ddcf44dddf703656239b3e8ecaa62bf2f4f58b598aa7

Observation c0978bca-366d-49ac-801b-55b80b065418 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.266073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.829016Z digest=sha256:cf70d943194e2837769d251887a8a475f024f96b89c0cf9f8ea6396d8fd19243

Observation 70e70d2c-3577-409f-8edb-c6d83fec4a91 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.252060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.833955Z digest=sha256:72b9dffb949daafa38e5455bf700beea2c44f9b35817a146d81d73f54abb99e1

Observation 69526ea0-acd6-4c1a-b90b-a65febcd5a91 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.237380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.838309Z digest=sha256:3a96184f96646aaeaed32844944017ca60cfa1ac9049ceab8ba3927d1bd89d5c

Observation 7a359cd9-a89b-4151-b592-2b6075e3e73b · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.221868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.842440Z digest=sha256:2b9257f9f4e09be998af3296e1f0f4526e0a5b906fbc7bfaa6250237e5c6e324

Observation 95b7c3d1-10b7-4e3d-9774-a5ee7c04d747 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.206779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.846787Z digest=sha256:59b14e07ebaddebf4019a2e6991f1fa7f48bbd545e3b8d7b92acc5280f00f0cd

Observation 55ec6a08-cbb9-4afb-bae2-853d5c06997c · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.191700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.851111Z digest=sha256:5ccd36904c7bb0d42e5b038f1f01160858bb1eb77650f1cbb77b5dec22127382

Observation 0b52ec4d-6e79-4e64-b88e-82f8156726f0 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.176506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.855330Z digest=sha256:6d88916123d5e827bfd557150c6700eecb212a46ac8110362af164d42c1b211c

Observation 4ed2f487-db09-40f8-856e-fd0699c39ed2 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.162005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.859348Z digest=sha256:8c355d13d3230793289bedccb2622c490a770c44a2aa7fa708739c51c35927a8

Observation b48ec891-ff5a-4114-a4ff-f89f1b5c907f · outbound

This paper cites W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.864813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.864813Z digest=sha256:43d9a9865bbcbb6cb56640f8f890308b0e1f7c397bb8345dd2f30c71a6289b0b

Observation 1e94e569-b79c-4f41-b772-23178a59a521 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.138379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.868936Z digest=sha256:b2a3083c82eab8d078e59417b9e9e45cd7f24340790396a667555e58a9dd7b4b

Observation 640690fe-27a4-4af8-98ac-b369067a71bd · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.125191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.872844Z digest=sha256:c2a7254e477d01e345c964daa099d5d9ae9420dd8f34896fa1beef59c0a3f155

Observation 79c5ca32-baf5-4c51-a7dd-bb5b881743ab · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.112109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.878274Z digest=sha256:582d5f0e8bb744d39df6cddc542e5e677593b8b7b656562379d8bae14408f8ce

Observation 0c14f8f4-2d1c-4f34-80e8-7f37a6a47d4e · outbound

This paper cites N.; Tran, S.; Yao, B.; Chilimbi, T.; and Shah, M.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech N.; Tran, S.; Yao, B.; Chilimbi, T.; and Shah, M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.097133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.882704Z digest=sha256:cbbc9f71978b665a96a83214feca69ea36518ebf3005ac32829c9b36855aa888

Observation 38206b27-7f94-472a-9891-5a092783cb75 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.081238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.887301Z digest=sha256:742d77160542693358734eb6e5ba48c5f38463324192c9ee1d69c50401f4ffe1

Observation 1b240567-7cb5-4136-b312-0db0ea272faa · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T15:01:33.891600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:01:33.891600Z digest=sha256:0e56e59dae48e2d85847b061c575ff7d29b63687044905c22f26fff6a9f1e97a

Observation 17aeb1c9-3e11-4734-853c-aad21ad18bd7 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.066593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.897141Z digest=sha256:6debbd23eb1f64f8163d1587961eaa4d5060fa4d2cc18392e304481c21a5bad6

Observation 761eacb8-d295-4c79-bb62-0f9c58bee78e · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.052267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.902340Z digest=sha256:68ac74914ac416bb4358474e1b0df1726ffdecf136974c6b81eeccaf22ac8838

Observation bd7e2edf-23ba-465b-9cbc-7bd953f8cfcb · outbound

This paper cites C.; and YAN, S.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech C.; and YAN, S

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:01:34.036847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.907180Z digest=sha256:6803e98df6c019f5c1dd89a01641cfb9346e28ed2ab5df77b012f4535bbf136e

Observation 01a873bd-32f7-49ba-9806-04828d57aec4 · outbound

This paper cites an unresolved cited work.

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:01:34.020850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-11T15:01:33.912139Z digest=sha256:49bfe7eaf45f682caf790d1659061d5bc1f7194da5852f69c0c4b570b0e8f6b7

Pith citing papers

Observation 6ece83f0-cdae-48d8-b707-a9faaa591b34 · inbound

Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction cites this paper.

Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-11T04:34:07.977433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T04:34:07.907336Z digest=sha256:c2d971170a106e3ae93b6377576c5da3dd2f2870948a20ce701848e198159c77