Pith. sign in

Paper Citation Record · LEDGER

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync

As of 8 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 0 inbound Pith citation observations for arXiv:2507.20452.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20452 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:39:36.267401Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

86 of 86 outbound references displayed

  • verified exact3
  • verified fuzzy46
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b5c61b92-42de-440b-8864-c2c13d694652 · outbound

This paper cites Facewarehouse: A 3d facial expression database for visual computing.IEEE Transactions on Visualization and Computer Graphics, 20(3):413–425, 2013.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Facewarehouse: A 3d facial expression database for visual computing.IEEE Transactions on Visualization and Computer Graphics, 20(3):413–425, 2013

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:27.844754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:27.844754Z digest=sha256:4389368140067d70850b7b5641ee3fdbbd6f332212184a416d94c6d7450e06bb

Observation e53ddc25-af8d-4e39-a107-3f167b852de0 · outbound

This paper cites Hiface: High-fidelity 3d face reconstruction by learning static and dynamic details.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Hiface: High-fidelity 3d face reconstruction by learning static and dynamic details

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:27.988641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:27.988641Z digest=sha256:23118f1458e82fa39a3f3b97da8e7155cdd9d560d83774b5815d8b27abf3dccf

Observation 9ccc7da9-a064-4f53-b334-f8df148f8e21 · outbound

This paper cites IQA-PyTorch: Pytorch toolbox for image qual- ity assessment.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync IQA-PyTorch: Pytorch toolbox for image qual- ity assessment

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.098360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.098360Z digest=sha256:d49c4450df1bd5fced54f490739a374a67c6d06f3824d5b3e9ec44fb46f44a5c

Observation 036fa333-8e8a-48f0-b837-092eeda73bf4 · outbound

This paper cites Topiq: A top-down approach from semantics to distortions for image quality assessment.IEEE Transactions on Image Processing, 2024.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Topiq: A top-down approach from semantics to distortions for image quality assessment.IEEE Transactions on Image Processing, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.199572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.199572Z digest=sha256:d4aa864389549e78deafaf784d7174cc83449e9bdaf03b8dabaaf0d15495a3bc

Observation 2b8a60f2-fb4c-42f6-9e76-3f61b8d69fcc · outbound

This paper cites VoxCeleb2: Deep Speaker Recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync VoxCeleb2: Deep Speaker Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.325314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.325314Z digest=sha256:96575d54df93ca5062ba6e08ca6e21fe218d8da9115cb698ab5a5678d592dd47

Observation 64300e98-df5e-45d8-a1ff-9ed4a51d992d · outbound

This paper cites Emoca: Emotion driven monoc- ular face capture and animation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Emoca: Emotion driven monoc- ular face capture and animation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.419724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.419724Z digest=sha256:1e87a93b10d852a791f1ecc5b15d1219e6cb51f45aa1216c0bb6c79d48587e7d

Observation 5e8311fb-0280-47f4-8c02-596ffa3c3ea7 · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Arcface: Additive angular margin loss for deep face recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:28.554745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:28.554745Z digest=sha256:466fd28d5ac0c3c9aa97bdd553e4a763ee32d7fd8c48274f8df8116987fb63ff

Observation a8416907-96b5-4a95-ad01-e07d2f657584 · outbound

This paper cites Ac- curate 3d face reconstruction with weakly-supervised learning: From single image to image set.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Ac- curate 3d face reconstruction with weakly-supervised learning: From single image to image set

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:52.154750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:28.634750Z digest=sha256:3f67ce2683ed4fcb0bde38304810421d17094484f78744dc22711c93a17126a0

Observation b6826680-cb5a-4022-be23-c926fbddc557 · outbound

This paper cites Headgan: One-shot neural head synthesis and editing.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Headgan: One-shot neural head synthesis and editing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:51.704889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:28.779920Z digest=sha256:e0c72be078e7adaf503672c8817f529a6806e0b27404751997335c66c791abda

Observation 9594e41b-880b-42f0-a799-2e44e584cdbe · outbound

This paper cites Free-headgan: Neural talking head synthesis with explicit gaze control.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Free-headgan: Neural talking head synthesis with explicit gaze control

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:51.294070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:28.914742Z digest=sha256:522ce9dd66411a9023d7ba99298abb4e1c311c2a328a06f8adda672e410b43a7

Observation 40b383c9-873e-45b6-b18e-906817f839bd · outbound

This paper cites Megaportraits: One-shot megapixel neural head avatars.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Megaportraits: One-shot megapixel neural head avatars

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:51.035436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:29.098001Z digest=sha256:82a29279073516aa14b5988210cf2ae06d584e18699eb619b6aa5d8ad0230f1a

Observation 35fb1ffb-7c72-4aa9-95bd-33bc82142019 · outbound

This paper cites Emoportraits: Emotion-enhanced multimodal one-shot head avatars.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Emoportraits: Emotion-enhanced multimodal one-shot head avatars

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:50.747678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:29.252262Z digest=sha256:c6c06f22a46248dbc9cc0f7a34d7211cd39c598fc6709b070311f1c6880f6ccb

Observation aa420262-44c8-4865-bc32-8fc217c28963 · outbound

This paper cites 3d morphable face models—past, present, and future.ACM Transactions on Graphics (ToG), 39(5):1–38, 2020.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync 3d morphable face models—past, present, and future.ACM Transactions on Graphics (ToG), 39(5):1–38, 2020

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:50.434842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:29.340935Z digest=sha256:432d5c99c7de5e3b35cf5c9a77afa9387c6f1cafb07f91f3e49b6ea646836a31

Observation 2f75bd06-e7b8-4aa1-9fc2-b4d209e1fabb · outbound

This paper cites Facial action coding system.Environmental Psy- chology & Nonverbal Behavior, 1978.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Facial action coding system.Environmental Psy- chology & Nonverbal Behavior, 1978

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:50.028325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:29.438106Z digest=sha256:81198d8eedc2f468d6525c26bbc19d1ba9a853643415c5f48e7417cd2ad3eb8c

Observation e23e40ce-d579-4825-974c-0d672e3df384 · outbound

This paper cites Looking to Listen at the Cocktail Party: A Speaker-Independent Audio-Visual Model for Speech Separation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Looking to Listen at the Cocktail Party: A Speaker-Independent Audio-Visual Model for Speech Separation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.510807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.510807Z digest=sha256:0b532490d8e3c4cc6e0d403893542b89c020247be101cef97f841675954d1e64

Observation 5c86ad74-8e25-477c-aeab-aef5108e5339 · outbound

This paper cites Learning an animatable detailed 3d face model from in-the-wild images.ACM Transactions on Graphics (ToG), 40(4):1–13, 2021.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Learning an animatable detailed 3d face model from in-the-wild images.ACM Transactions on Graphics (ToG), 40(4):1–13, 2021

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.590290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.590290Z digest=sha256:30891c069dc054a7d5bb01d92d7b8adb507395432675ee7a87551b6117d1aa62

Observation 833d90c2-3484-425f-a387-38cb244c7af6 · outbound

This paper cites Surface simplification using quadric error met- rics.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Surface simplification using quadric error met- rics

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:49.634750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:29.646577Z digest=sha256:d581fb150911afcacd62632c30aab8168ed7ce8b9ef526e1714bab75280bab38

Observation 2d03d31e-d535-4d28-a03a-7c05d0886086 · outbound

This paper cites Morphable face models-an open frame- work.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Morphable face models-an open frame- work

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:49.325837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:29.710651Z digest=sha256:f95b264b9f93cabfedd9291fab2ad813c7817425cea05bdb791b495a53258260

Observation 23d37766-fe89-4821-89f4-32a86b935009 · outbound

This paper cites Attention Mesh: High-fidelity Face Mesh Prediction in Real-time.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Attention Mesh: High-fidelity Face Mesh Prediction in Real-time

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.799219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.799219Z digest=sha256:b9693dcf6e5eab361631dba4f1783e3c57f22876f5b23d58670ac9bb5e1c9250

Observation 5ceb2151-0041-4694-aa51-e82987f7f473 · outbound

This paper cites LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.907882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.907882Z digest=sha256:36f437eebdc0a1c27c7c5988e4bd0b752e013891f1dd69c0451f89e6af71b755

Observation 5aaf2052-996d-4143-a032-daeb9c554c38 · outbound

This paper cites Deep residual learning for image recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Deep residual learning for image recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:29.997607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:29.997607Z digest=sha256:534b6fbd1a448e7ec98d104708802db0868c4a44b5285dab67edc30f60adca9b

Observation 10ebb122-024f-4579-a734-e317f77588ec · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.129703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.129703Z digest=sha256:ae44bf7e33ff071236c96c50a89e1584e0167b7f2a9f96437cf39677a22e4015

Observation 2f9c5b4b-a39c-4015-8bad-19eb3343bda5 · outbound

This paper cites Classifier-Free Diffusion Guidance.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Classifier-Free Diffusion Guidance

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.197962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.197962Z digest=sha256:3f251892bd095a4d68634b5e25518978949c99f2c6472fbde1170a26fd7e88fd

Observation dc607002-be97-4dde-8835-1fdd06b0a10f · outbound

This paper cites Denoising diffusion probabilistic models.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Denoising diffusion probabilistic models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.313052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.313052Z digest=sha256:d353877f00e8a600d12badb2cc489ab6357937f3b61a0d2d9a855093f2b1d517

Observation aa1f875f-e376-4ce1-a937-d31e3f6860e3 · outbound

This paper cites Image-to-image trans- lation with conditional adversarial networks.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Image-to-image trans- lation with conditional adversarial networks

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:48.769236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:30.376358Z digest=sha256:261c6c29db804762db65e72e7d797bd0c2d380c0297d4f72cbed41445d868876

Observation 91a3f7c3-d4f3-4483-990a-bb19a66d338a · outbound

This paper cites RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.454868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.454868Z digest=sha256:adb5771c03581f6fb8d93303a6be9820643e452c6d681153ce63d175a81f3195

Observation d28459ac-b837-43cc-bd0a-d8b7408469c6 · outbound

This paper cites Eamm: One-shot emotional talking face via audio-based emotion-aware motion model.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Eamm: One-shot emotional talking face via audio-based emotion-aware motion model

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:48.466735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:30.530150Z digest=sha256:dfd74f1c1fc4b946b49f9c29c4fb46d0fb5ccf7d7b5cce78053c761895b0545f

Observation 8c28182c-38c5-4419-ae04-ebc6fc929316 · outbound

This paper cites Loopy: Taming audio-driven portrait avatar with long-term motion dependency.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Loopy: Taming audio-driven portrait avatar with long-term motion dependency

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:48.172525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:30.643230Z digest=sha256:4cf22ac0313a62e4914ac3198113f69bbd9b2734762283e86553adf2844fd449

Observation 56aeb3b4-4077-4ecc-9755-50f082108855 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Adam: A Method for Stochastic Optimization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.716088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.716088Z digest=sha256:da32cf30c9181483a93389fda1e3f58dd6c2ad389d5e9ba13fc95ab95ff77352

Observation 97bd352f-9815-4354-8069-726945ac47f2 · outbound

This paper cites Photo-realistic single image super-resolution using a generative adversarial network.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Photo-realistic single image super-resolution using a generative adversarial network

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:47.884757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:30.809494Z digest=sha256:604c1828d4e69347aa00173ce32336b829714e5549ba4ca9514eb0044dc7e56d

Observation 3299ad7a-593c-4076-8a1e-68af829e8e84 · outbound

This paper cites LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:30.957571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:30.957571Z digest=sha256:04a2557a9798de0d037967e89668730d332eea5dbdf6fb796275df3498f3790a

Observation 09d231e1-49f5-4228-b14c-a788cc1481c1 · outbound

This paper cites Learning formation of physically-based face attributes.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Learning formation of physically-based face attributes

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:47.487963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.057220Z digest=sha256:4bfd6a6d3c7b5c7060e64809a311efa9a48427626329a1f5bb4316089f312312

Observation ed54530d-224a-4b51-8051-7b4468aa05e2 · outbound

This paper cites Geometric GAN.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Geometric GAN

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.123805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.123805Z digest=sha256:819482944b852c15977033d679520b9b7672a5b408a799aeb3380f6018486cc8

Observation 5abfb1b5-46c5-4e93-be8e-d653e54737b1 · outbound

This paper cites Robust high- resolution video matting with temporal guidance.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Robust high- resolution video matting with temporal guidance

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:47.275043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.204493Z digest=sha256:7345035de85c15c4ad77902c2d1278ea2b1386117926864906706777694f2705

Observation 447359e1-4e18-4db1-b8b2-17496f8b76da · outbound

This paper cites Anitalker: animate vivid and diverse talking faces through identity-decoupled facial motion encoding.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Anitalker: animate vivid and diverse talking faces through identity-decoupled facial motion encoding

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:46.944757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.283100Z digest=sha256:d463441b9dfc67319cea8223b8f423e83c1f41794aa76ad3b92308c1801202f9

Observation 916803fc-2b96-480e-86f2-02362b0caadb · outbound

This paper cites Decoupled Weight Decay Regularization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Decoupled Weight Decay Regularization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.347974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.347974Z digest=sha256:13d34517531438bb84cc70f9e0fd115098193b31c3bfca161640293841ae9163

Observation 6d010939-c2c8-4da8-be34-77c14a988298 · outbound

This paper cites Repaint: Inpainting using denoising diffusion probabilistic models.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Repaint: Inpainting using denoising diffusion probabilistic models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.456523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.456523Z digest=sha256:799615f15952002d3069e3e8c3f18e277eed4efab884e6a0ff903623ae4d5c81

Observation 59ece410-6f95-4a26-8001-f29fb2306383 · outbound

This paper cites Implicit warping for animation with image sets.Advances in Neural Information Processing Systems, 35:22438–22450, 2022.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Implicit warping for animation with image sets.Advances in Neural Information Processing Systems, 35:22438–22450, 2022

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:46.564761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.554628Z digest=sha256:422e48b147f8fbf0c4705868347d2989d3b028e95039ed6922fc65a95cb2d8a8

Observation 629c79a5-ad66-474c-8235-c0b5df06a66b · outbound

This paper cites Sidgan: High-resolution dubbed video generation via shift-invariant learning.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Sidgan: High-resolution dubbed video generation via shift-invariant learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:46.209299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.616512Z digest=sha256:155a7d1dac3d3ca4af6d28c34267d4f401e34570ef4d143ccad3d10d751382b9

Observation 78b3d442-d2a7-46ec-8f74-827b8593fadd · outbound

This paper cites SAiD: Speech-driven Blendshape Facial Animation with Diffusion.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync SAiD: Speech-driven Blendshape Facial Animation with Diffusion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:31.735563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:31.735563Z digest=sha256:eddbb724b3b3d3769fa256ba1aa9e3c508ebcfcb6c095a839fb6c6dd0300b0f7

Observation e12a04d5-9e4c-4f95-bbd4-fa6df4d89496 · outbound

This paper cites Synctalk- face: Talking face generation with precise lip-syncing via audio-lip memory.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Synctalk- face: Talking face generation with precise lip-syncing via audio-lip memory

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:45.930393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.837154Z digest=sha256:44de908a714d7ea635a0d41aaa56d9b482123b1443f5e845094f1d8ec33535d9

Observation 18b4a7fe-d2cd-440b-9152-21e9f6e1907d · outbound

This paper cites Interpretable Convolutional SyncNet.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Interpretable Convolutional SyncNet

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:39:37.325379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:31.975485Z digest=sha256:0ee59b156c0657d374518c59fa243c7d4b5d33c91eec4076251cfc0a60ab7a12

Observation 9b9c38e4-9715-443e-b3cd-74df5d3d15ed · outbound

This paper cites Semantic image synthesis with spatially-adaptive normalization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Semantic image synthesis with spatially-adaptive normalization

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:45.534830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:32.077299Z digest=sha256:0c69d0de4293bd54edb88e2e6595545127230f3376f3df5a6e132af1395a9808

Observation 7a8cfbf9-506e-4a66-a661-6413bbe2ba9e · outbound

This paper cites A 3d face model for pose and illumination invariant face recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync A 3d face model for pose and illumination invariant face recognition

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:45.197617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:32.253438Z digest=sha256:273d0a2bea52e8c8164d1457994d615e11fcc4a663c0fe2171ca592af40ddb52

Observation 43871328-a0f3-4e5d-a8e3-3299a07bf0da · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync A lip sync expert is all you need for speech to lip generation in the wild

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.384741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.384741Z digest=sha256:0ead33bb9f610bf7c4157f312a96aa03df53f7f3dad220070ee2ba4bb760af26

Observation 13cc7b11-b53d-476c-adfe-efcb38840494 · outbound

This paper cites Accelerating 3D Deep Learning with PyTorch3D.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Accelerating 3D Deep Learning with PyTorch3D

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.431687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.431687Z digest=sha256:3a80073729e2a039ecb5b5dc604e92f75bcabafd50cb7bbcc16f6fe247b4984e

Observation 04e3ab69-bc01-496f-805c-c7c2f6fad83c · outbound

This paper cites Pirenderer: Control- lable portrait image generation via semantic neural rendering.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Pirenderer: Control- lable portrait image generation via semantic neural rendering

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:44.816217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:32.553045Z digest=sha256:799305553c26329f82745ef3a2ea41b791ca52dc03a44cd20fa2f4241d2637c9

Observation f03f55cd-d0b9-40a6-9506-02867c514aad · outbound

This paper cites U-net: Convolutional net- works for biomedical image segmentation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync U-net: Convolutional net- works for biomedical image segmentation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.603221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.603221Z digest=sha256:0ee0bb2e22a1ef89784653693569472d77790e8339892815d036bc69dc6671cc

Observation 15b40bc3-5270-416b-8f8f-306d813fd4f6 · outbound

This paper cites Palette: Image-to-image diffusion mod- els.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Palette: Image-to-image diffusion mod- els

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:44.479606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:32.656474Z digest=sha256:4b2c5e5c1a0d50da0e5f70554718f2822904427dcdb6771716ff5d2a4f4dbc80

Observation a5e4fb18-61d8-4aa4-9951-19e4c71d618d · outbound

This paper cites Improved techniques for training gans.Advances in neural information processing systems, 29, 2016.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Improved techniques for training gans.Advances in neural information processing systems, 29, 2016

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:32.802566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:32.802566Z digest=sha256:4097e24f306a706c60bc4716adec06fffa273b47b00f9ae198a018f4491cefaf

Observation d96f5c64-a23c-48f1-b7d9-a8dcc8fa3ca1 · outbound

This paper cites pytorch-fid: FID Score for PyTorch.https://github.com/ mseitzer/pytorch-fid, August 2020.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync pytorch-fid: FID Score for PyTorch.https://github.com/ mseitzer/pytorch-fid, August 2020

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:44.054773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:32.952752Z digest=sha256:1c025225ad3216b55292d4ded6a174499549a0b17bf22fae2ffa46541f3ae62d

Observation 207ed7f0-e8c4-4a9a-a9ab-e46f2171d204 · outbound

This paper cites First order motion model for image animation.Advances in neural informa- tion processing systems, 32, 2019.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync First order motion model for image animation.Advances in neural informa- tion processing systems, 32, 2019

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:43.721816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:33.025935Z digest=sha256:2ac1e068495e94b8a91bb15c124c9cae356737a1e6bd3a2072fad1dac4d9a56a

Observation 634d6d80-3e21-434b-87c3-1a12f1065992 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.146640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.146640Z digest=sha256:e022a1f27fc27793c1d509b885035c2900b2efdb1d6a83c7fca23ca7d0397272

Observation f6285648-cc33-4b02-8ec7-0d7bc345be58 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Deep unsupervised learning using nonequilibrium thermodynamics

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.258476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.258476Z digest=sha256:5e9661030afe483377fb6030d5c1643b83394c968d078ea957808f1afd520e25

Observation 8a61772a-7f49-4d2c-b904-cb72350e8f88 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Score-Based Generative Modeling through Stochastic Differential Equations

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.344753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.344753Z digest=sha256:f49475242b9504e3482f4982c50feb3c0615ee1a6e8ccae8908f4975aa1682b6

Observation 4d5a9aed-d76c-4689-9b42-cb70207d4d25 · outbound

This paper cites UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.440194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.440194Z digest=sha256:29ebae7309bb89feb984b1335d2c4c6eb4f3fdc85a8c4ff5bb94a3840a97f671

Observation e030ceec-abb6-462c-bf73-a8f6a3641bd0 · outbound

This paper cites Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4): 1–9, 2024.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4): 1–9, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:43.329796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:33.563374Z digest=sha256:7666fe05def73585f56f8c1df9dc91cc45006242ddb949489e0bb92accabe98b

Observation c74ae494-b109-4b0d-ada8-38ea9c7732ed · outbound

This paper cites Mofa: Model-based deep convolutional face autoencoder for unsupervised monocular reconstruction.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Mofa: Model-based deep convolutional face autoencoder for unsupervised monocular reconstruction

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:43.026476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:33.671556Z digest=sha256:370d53a9a7d1e4f27fdf978f64cef965b2c23f31ae052b8fa457ef40e328c49f

Observation 4c3ffe1a-c906-4fa6-b52a-69d3472ea9fd · outbound

This paper cites Instance Normalization: The Missing Ingredient for Fast Stylization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Instance Normalization: The Missing Ingredient for Fast Stylization

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:33.789332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:33.789332Z digest=sha256:1400c330f342839f6009120b4bcc4e621fb8b38192e6f67ddaac266e2f6e17f7

Observation e95bcd4b-54e0-4c99-8704-59a1955c660f · outbound

This paper cites Seeing what you said: Talking face generation guided by a lip reading expert.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Seeing what you said: Talking face generation guided by a lip reading expert

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:42.762236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:33.915452Z digest=sha256:34248dc72a2cfe6c069c57919053a7aebcaf56e0acf26aa332c5234993ca7530

Observation cca0aa83-f00d-42b8-a2ec-60bf1bb2b0d0 · outbound

This paper cites JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:39:36.934732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.035222Z digest=sha256:55c183a404d5fd9cef0d69c33da17600da734070268fa05a709564cd776c49aa

Observation b0910bb9-7e8e-4faf-8ca2-8bc1a6c9c04c · outbound

This paper cites One-shot free-view neural talking- head synthesis for video conferencing.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync One-shot free-view neural talking- head synthesis for video conferencing

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:42.524760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.095342Z digest=sha256:5e1e656276a03c6ed36ed8b037d68180d39b0b6b77cb0223d5a98a2555ce8428

Observation 800e569a-8dc8-44f6-a940-de396405586b · outbound

This paper cites Im- itating arbitrary talking style for realistic audio-driven talking face synthesis.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Im- itating arbitrary talking style for realistic audio-driven talking face synthesis

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:42.177100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.191233Z digest=sha256:d667fca578583bbddfc962e1e6afb66970a8b0c400230235c1cb76445b20afb3

Observation e1cd2faa-24b2-4e10-b08b-b60d51c27c4d · outbound

This paper cites Group normalization.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Group normalization

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.321999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.321999Z digest=sha256:ac85597a39763e72dea354ea3227c362b8cbf7ce05e6165979777dbfbc55a806

Observation 3a0fb7a7-b975-4c4c-82b1-a21a80bac6a8 · outbound

This paper cites Vfhq: A high-quality dataset and benchmark for video face super-resolution.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Vfhq: A high-quality dataset and benchmark for video face super-resolution

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.956369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.448317Z digest=sha256:ffdbbc94fb7dfa4fc827621e39c176923d90bd18f406b9d90ca3d20d8800bdfa

Observation 5c7f96f4-0f53-472a-9c71-5ff53db2c77b · outbound

This paper cites High-fidelity generalized emotional talking face generation with multi-modal emotion space learning.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync High-fidelity generalized emotional talking face generation with multi-modal emotion space learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.706446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.547704Z digest=sha256:7345ab3a8fe8293d5f38dca0ee065afe4b44af1eaa489ba658978d577f616d31

Observation 776c542c-0ebf-4ce1-8283-319b259dc0f0 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.642522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.642522Z digest=sha256:92f18aae3413b6364f390ef4c1f0087661665f828c9f38ce8a3d9850be893bc1

Observation 664cfe33-f7d2-48b7-8b67-823303c094aa · outbound

This paper cites Vasa-1: Lifelike audio-driven talking faces generated in real time.Advances in Neural Information Processing Systems, 37: 660–684, 2025.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Vasa-1: Lifelike audio-driven talking faces generated in real time.Advances in Neural Information Processing Systems, 37: 660–684, 2025

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.478297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.736629Z digest=sha256:b7cc454a2a49ffd73aa29917fdca6a123fa335452b9e9a6134c8e52672e3dadd

Observation a621fd21-2307-4fea-94fe-f4f451b6878f · outbound

This paper cites Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.794532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.794532Z digest=sha256:bc6e7882021d4fd90ee50e4b4437afcef6b2868a034292ae40b7e11736957998

Observation d52bddd0-e699-4fc2-a245-258d573e19b4 · outbound

This paper cites Dynamic Neural Textures: Generating Talking-Face Videos with Continuously Controllable Expressions.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Dynamic Neural Textures: Generating Talking-Face Videos with Continuously Controllable Expressions

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:39:36.594730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:34.839399Z digest=sha256:dcfcdadc9be0a8bdccde31ad89ef30d709fdfe6f05a24edf2d8d32e344a08289

Observation 63ba9994-1cfe-4f93-b806-b72b859aa270 · outbound

This paper cites Audio-driven Talking Face Video Generation with Learning-based Personalized Head Pose.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Audio-driven Talking Face Video Generation with Learning-based Personalized Head Pose

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:34.940981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:34.940981Z digest=sha256:b2184dddd32385ac191f16fccaa2ab40145d74a01998e05aa88e9df5d7e0c67b

Observation 4157b59f-bfa4-4588-9252-5a8b327f1495 · outbound

This paper cites Face animation with an attribute-guided diffusion model.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Face animation with an attribute-guided diffusion model

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.262334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.033457Z digest=sha256:428e4262e27c6a98b5c6393d1231f972c6c8232d182d993c09b87789e43b9ec4

Observation 9bac0f53-df88-4098-a062-d3d80a038363 · outbound

This paper cites Facial: Synthesizing dynamic talking face with implicit attribute learning.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Facial: Synthesizing dynamic talking face with implicit attribute learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:41.002242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.108427Z digest=sha256:8e5a3bbad4a696dd5bbf45c7d36ae4e886eeea244079b5daf14c7c6259d57302

Observation 0422318c-2edd-4c9b-a99b-19dd9e1bdd53 · outbound

This paper cites Refa: Real-time egocentric facial animations for virtual reality.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Refa: Real-time egocentric facial animations for virtual reality

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:40.674745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.202295Z digest=sha256:67ef58753ec6cd6645f623643907f3c0e83371be25a66df277e7f7183fb137a0

Observation 879dae3c-ae67-4888-b272-598b973b81a8 · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:40.308738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.285388Z digest=sha256:951ea9844634bc70fafb5451704eb6b406fe2ef792e440187b65b4d2cf371dd9

Observation 9811c709-e290-485d-8a75-120a73b6c8ff · outbound

This paper cites MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:35.346930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:35.346930Z digest=sha256:133ed1b7faa5b3f654c5442f02e488c65fdd5998ac8b77bd2cce566738cba521

Observation 8b9d950b-6414-4087-a251-91625d824470 · outbound

This paper cites Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Flow-guided one-shot talking face generation with a high-resolution audio-visual dataset

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:40.054748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.424448Z digest=sha256:0b02f070236dc03aefccef9aa94cf817757ae8354c5e91c0d17d6b10ace4558b

Observation c020f487-236e-43b8-a64f-a0949742b08c · outbound

This paper cites Dinet: Deformation inpainting network for realistic face visually dubbing on high res- olution video.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Dinet: Deformation inpainting network for realistic face visually dubbing on high res- olution video

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.824124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.481510Z digest=sha256:8c9b9ee1d57a40b63f82fca1708fb33f97ec090fe28742ce18bb59e875e77932

Observation b6829569-aedf-4c87-bd57-1b7fd6d8b006 · outbound

This paper cites Avid: Any-length video inpainting with diffusion model.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Avid: Any-length video inpainting with diffusion model

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.579572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.584749Z digest=sha256:782f20eaa3b759ce3a0cce3896e3c045803551a0278783c3de10ab3f550d2d15

Observation b4b5f0c6-1f7b-416b-a740-3c14836a54bc · outbound

This paper cites General facial representation learn- ing in a visual-linguistic manner.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync General facial representation learn- ing in a visual-linguistic manner

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.366507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.672562Z digest=sha256:83dd9c947c06de48366d484b4f30a6b1f892905966789c32df4328e2c0e968fd

Observation db3a4d5b-bec0-4ead-afda-b9904fdd7efc · outbound

This paper cites Identity-preserving talking face generation with landmark and appearance priors.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Identity-preserving talking face generation with landmark and appearance priors

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:39.125796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.695632Z digest=sha256:1401abbacc979b47ed9fb0dedbde95816e51275c547a7c7686ceb1c346b93d24

Observation 30928c61-73dd-4168-b5f1-80bb44a3e58a · outbound

This paper cites On the continuity of rotation representations in neural networks.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync On the continuity of rotation representations in neural networks

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.914745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.804750Z digest=sha256:ae7a3b7bce3d0af28a005b44fa501d7d71ef7c1fdde65fe52d39398097d79bb6

Observation 16350022-bbee-41f7-8a26-41640f31ec27 · outbound

This paper cites Celebv-hq: A large-scale video facial attributes dataset.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Celebv-hq: A large-scale video facial attributes dataset

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T13:39:35.884744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:39:35.884744Z digest=sha256:8e3e42a9f4ab92635110aecd364e38b6c6be1d3e5e0bf61502b9da23c0447418

Observation 6ddcae72-4a54-4602-addd-1572a4f83b33 · outbound

This paper cites Face alignment across large poses: A 3d solution.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Face alignment across large poses: A 3d solution

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.679177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:35.995606Z digest=sha256:82917eb8413b8bf00873ee93494fefb0c3e0ca19b1879417621ce513a0dad053

Observation fa695015-4b98-413e-adea-91038e6223b8 · outbound

This paper cites We also include identity consistency lossL id =∥α 1 −α 2∥2 2, where 1 and 2 indicate identity param- eters extracted from the same video but from different frames.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync We also include identity consistency lossL id =∥α 1 −α 2∥2 2, where 1 and 2 indicate identity param- eters extracted from the same video but from different frames

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.329313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:36.109562Z digest=sha256:ebb04e4565d3f51d600351886d1a4557de9f27315ad226a588621d195976ed27

Observation 3b55c148-1b18-4996-b193-375f349ce86a · outbound

This paper cites Recall from Sec.

JOLT3D: Joint Learning of Talking Heads and 3DMM Parameters with Application to Lip-Sync Recall from Sec

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:39:38.046610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:39:36.267401Z digest=sha256:b21a1a67a649a3d75ac246d20f23bd4f1045abe0fe266c2101c874ad7f40ab02

Pith citing papers

No inbound Pith citation observations are available.