Pith. sign in

Paper Citation Record · LEDGER

Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2311.05844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.05844 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:50:43.735052Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T08:03:13.799275Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e6f77063-e505-4adc-bede-bbd5ddb52ac3 · inbound

Emotional Face-to-Speech cites this paper.

Emotional Face-to-Speech Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T16:50:43.735052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T16:50:43.735052Z digest=sha256:244f633fe553c7e37ef2cef92f390ce2dcbf14a9f6996a5566946f214c70b8e3

Observation 2a049e5f-6024-4dd4-bd81-5397f65638c4 · inbound

Hierarchical Codec Diffusion for Video-to-Speech Generation cites this paper.

Hierarchical Codec Diffusion for Video-to-Speech Generation Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:12:26.280862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T08:12:01.260833Z digest=sha256:312c6ce1f97f44eb94502536a29c43a008d09fc3e997515bffdd0f8b43d630ec

Observation 5f749fed-affe-415a-b68e-ac1693998c52 · inbound

Archon: A Unified Multimodal Model for Holistic Digital Human Generation cites this paper.

Archon: A Unified Multimodal Model for Holistic Digital Human Generation Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.802213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:03:04.294439Z digest=sha256:7127b58a8d540949692daf999674c9208372254bd332ff1a1d879d88cb72a53a

Observation 45865aae-bcaa-474c-9591-061f402bae24 · inbound

Zero-Shot Face-to-Speech Synthesis via Latent Space Adaptation of a Style-Diffusion TTS Model cites this paper.

Zero-Shot Face-to-Speech Synthesis via Latent Space Adaptation of a Style-Diffusion TTS Model Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T22:26:14.487391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:26:14.487391Z digest=sha256:3c8cadee888cf76a4d474c92e6bf4e282d9758272bb04bfa9609a0df70be80bd