Pith. sign in

Paper Citation Record · LEDGER

Exploring Timeline Control for Facial Motion Generation

As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2505.20861.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20861 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:54.279621Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact6
  • verified fuzzy38
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e8f3d26e-c437-48ba-8b30-7823348a0832 · outbound

This paper cites Facetalk: Audio-driven motion diffusion for neural parametric head models.

Exploring Timeline Control for Facial Motion Generation Facetalk: Audio-driven motion diffusion for neural parametric head models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.840247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.163584Z digest=sha256:a69c5f3cd64f6f90b02c51dda655ead9d74efbb82422e6abb85b22ce0d3d2329

Observation fe2f996b-c31d-4885-abb5-64bffbfc8bab · outbound

This paper cites Teach: Temporal action composition for 3d hu- mans.

Exploring Timeline Control for Facial Motion Generation Teach: Temporal action composition for 3d hu- mans

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.573660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.274962Z digest=sha256:b2ccbe0795b97a809b0d870a798f9731f75315192e1bcc2766b29df386b951a8

Observation 6f71fbf8-c1c5-4618-8f0d-ab45897180f3 · outbound

This paper cites Seamless human motion composition with blended posi- tional encodings.

Exploring Timeline Control for Facial Motion Generation Seamless human motion composition with blended posi- tional encodings

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.434110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.409933Z digest=sha256:e415e49ca0619ca45367e42957b0e04d99438c65ca22faed173d283d3707e95c

Observation 6ae2a930-2173-4b1b-8a5d-f18abc55febe · outbound

This paper cites Video generation models as world simulators.

Exploring Timeline Control for Facial Motion Generation Video generation models as world simulators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.260523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.473574Z digest=sha256:0dd2e5f7d1a9e0e725619c4597701ce860513bc060e93263575f688760729718

Observation 27264f34-ed73-43a0-8ca8-39f359726ba2 · outbound

This paper cites Animated conversation: rule- based generation of facial expression, gesture & spoken in- tonation for multiple conversational agents.

Exploring Timeline Control for Facial Motion Generation Animated conversation: rule- based generation of facial expression, gesture & spoken in- tonation for multiple conversational agents

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.074037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.530163Z digest=sha256:5d3666a9cdcf6ee97715aa5c0dc2f512723cf5c76e588e9261089a971c0549bc

Observation 17f683c4-da09-48fb-9451-d598bf33cdd2 · outbound

This paper cites EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions.

Exploring Timeline Control for Facial Motion Generation EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:48.623861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:48.623861Z digest=sha256:9e418e998a1d86544b520b8078cfd8a28c3a1621af7ecdd96055460526dc9aff

Observation 07b1eb48-97c4-4717-bd2a-aebb57b7b7d4 · outbound

This paper cites Capture, learning, and syn- thesis of 3d speaking styles.

Exploring Timeline Control for Facial Motion Generation Capture, learning, and syn- thesis of 3d speaking styles

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.883957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.727652Z digest=sha256:1660951407ff38dde469d9f0933238633754217056b4d0f8aa43e48b711e8fbc

Observation e0ae3a21-b1c7-4a09-b263-56c7e858a2b1 · outbound

This paper cites Emotional Speech-Driven Animation with Content-Emotion Disentanglement.

Exploring Timeline Control for Facial Motion Generation Emotional Speech-Driven Animation with Content-Emotion Disentanglement

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:48.844014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:48.844014Z digest=sha256:7e4683443a3a470dd1c82914c081c36c2830aa37f3b0f8f77c810e390bd704a0

Observation b1b53d5c-bd2d-48bb-aaf1-bad0ce9fdec8 · outbound

This paper cites SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting.

Exploring Timeline Control for Facial Motion Generation SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.964439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:48.933064Z digest=sha256:b0fa33bd0751c43e1f40132b7a3977d197ab7c45898f9c727db8d57a295bb0af

Observation 822efe5f-76e2-4f27-8abc-8f0a3ce42fde · outbound

This paper cites Facial action coding system (facs).A Human Face, Salt Lake City, 2002.

Exploring Timeline Control for Facial Motion Generation Facial action coding system (facs).A Human Face, Salt Lake City, 2002

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.735132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.028666Z digest=sha256:412fb506eb10144f4f69cbbc36f074c8d352743d9e34a73813eb61b27915f5af

Observation a7ddf34d-a855-4df6-8cb1-0381a09944cd · outbound

This paper cites UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model.

Exploring Timeline Control for Facial Motion Generation UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.678185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.139392Z digest=sha256:14bb613358a45fbbbb7a8c609523d6d3b80671bd0fb6a0b6f98f840e3beddd8b

Observation fe52b9b4-c4ad-4789-a00d-a04be8fa1760 · outbound

This paper cites Faceformer: Speech-driven 3d facial anima- tion with transformers.

Exploring Timeline Control for Facial Motion Generation Faceformer: Speech-driven 3d facial anima- tion with transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.675102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.256063Z digest=sha256:8417f81b33d1ef6545c5ed5db6ec13f96cc393daec92598d4905b21ac90edf58

Observation d4f137cf-6685-47d6-b02b-643efebeddff · outbound

This paper cites Affective Faces for Goal-Driven Dyadic Communication.

Exploring Timeline Control for Facial Motion Generation Affective Faces for Goal-Driven Dyadic Communication

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.337095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.337095Z digest=sha256:15b2c33ccc86b9b45832a2b107dd2bbace50fca69791deda110c2fabcdfdb01d

Observation aad20c1e-9712-48f9-bd8c-dbfc7eebeb56 · outbound

This paper cites ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer.

Exploring Timeline Control for Facial Motion Generation ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.426666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.426666Z digest=sha256:6583d45aeff9c106818dac2a869b0d6d02c3e877fe5a7e15ae3ca83d6e516f8c

Observation 22f7283d-2b19-4f7d-8bef-763e52069d5e · outbound

This paper cites Ac- tion2motion: Conditioned generation of 3d human motions.

Exploring Timeline Control for Facial Motion Generation Ac- tion2motion: Conditioned generation of 3d human motions

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.328214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.488190Z digest=sha256:3cc8f2be4286437480b090932a621a6c99579a585f002d559378b96d3ce7b62f

Observation 8836d187-ae97-423d-9065-3015b5ce0353 · outbound

This paper cites Generating diverse and natural 3d human motions from text.

Exploring Timeline Control for Facial Motion Generation Generating diverse and natural 3d human motions from text

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.054404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.614660Z digest=sha256:669d553ea3ae0b947a49f98c50f398b4eb4f897cd1e3ea56d3687b084d1bc0b5

Observation be5a18b7-4a50-4142-840c-442dc86ac749 · outbound

This paper cites Micro-expression spotting with multi-scale local transformer in long videos.Pattern Recognition Letters, 168:146–152,.

Exploring Timeline Control for Facial Motion Generation Micro-expression spotting with multi-scale local transformer in long videos.Pattern Recognition Letters, 168:146–152,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:02.655173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.702574Z digest=sha256:d1e656c9575f8ceb7572331f1adf69a08cc7cad92ca272b60ebd5f161d11bcae

Observation 2b08c393-1a02-4c93-a9f9-d48d1b22493c · outbound

This paper cites Toeplitz inverse covariance-based clustering of multivariate time series data.

Exploring Timeline Control for Facial Motion Generation Toeplitz inverse covariance-based clustering of multivariate time series data

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:02.255990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:49.787589Z digest=sha256:665ffa13b5689d7f27b9221513aec3914e09a7c1578583c2c008e6209e7f8ba9

Observation 22a6ae26-72c3-4ca4-9321-01241a6c5239 · outbound

This paper cites GAIA: Zero-shot Talking Avatar Generation.

Exploring Timeline Control for Facial Motion Generation GAIA: Zero-shot Talking Avatar Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.888105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.888105Z digest=sha256:4110f0435ca5ce2e3ccf0883dd5b03a924a92c2dc86a5cbd9b609931bd140786

Observation ee7f8684-adb2-42ab-aa47-4dbbccd8a1cd · outbound

This paper cites Classifier-free diffusion guidance.

Exploring Timeline Control for Facial Motion Generation Classifier-free diffusion guidance

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.960005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.960005Z digest=sha256:5ee5ea1113ecd137d7e4548c57e0d336169d664c1f35c4f82b918ee391e0f789

Observation b6c6ff47-0767-434e-9a14-75e61fffe57f · outbound

This paper cites Denoising dif- fusion probabilistic models.

Exploring Timeline Control for Facial Motion Generation Denoising dif- fusion probabilistic models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.993307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.036450Z digest=sha256:5a0a00552d06e5f57e91a1f124a35812cb206de1d10a52121fcc5dfc19e32764

Observation 4f32f28f-b3e8-4e23-82e4-0c8a45e95da2 · outbound

This paper cites Multi-aspect mining of complex sensor sequences.

Exploring Timeline Control for Facial Motion Generation Multi-aspect mining of complex sensor sequences

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.856176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.138655Z digest=sha256:14f94f71595b15e447ed16dc389a22bef8d15995f363b9e18649615a127c76df

Observation daea6143-b097-4dcd-9655-0baf2b2e4c00 · outbound

This paper cites Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency.

Exploring Timeline Control for Facial Motion Generation Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.225081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.225081Z digest=sha256:6f4b10ebe507660e55f635b726ed977e2b0977e35513d19c1a67aeb1d0061e7b

Observation e8a73c3e-bb03-4a1f-adb0-f4d6a886b69e · outbound

This paper cites Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017.

Exploring Timeline Control for Facial Motion Generation Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.699979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.291313Z digest=sha256:e9a578dea4dd61998cbb2060fb4d6ce5719bc42475b94891b0e82760ac51eb1e

Observation 490f9ba2-21a9-4c6a-8d8c-fce972478b48 · outbound

This paper cites Kmtalk: Speech-driven 3d facial animation with key motion embedding.

Exploring Timeline Control for Facial Motion Generation Kmtalk: Speech-driven 3d facial animation with key motion embedding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.508258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.366335Z digest=sha256:91739fb7bf28d5b374594d73eb5f1b60570e01478dcf4b58019beae023b11c9c

Observation 81dbbc9c-8254-40e5-a8a8-75a950bf2ed4 · outbound

This paper cites Talkinggaussian: Structure-persistent 3d talking head synthesis via gaussian splatting.

Exploring Timeline Control for Facial Motion Generation Talkinggaussian: Structure-persistent 3d talking head synthesis via gaussian splatting

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.301208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.439421Z digest=sha256:392ca30ae8947f9d6417ea6dfd9e620496f653c90980a34f3687b23ca21c4598

Observation b53e64dc-1548-4177-8884-3f649be8e618 · outbound

This paper cites PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation.

Exploring Timeline Control for Facial Motion Generation PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.422351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.530894Z digest=sha256:d78b35c2c975717769eba6004496f7198b9171c5e4525a8db252979dab8b9e67

Observation 6c502d7a-fa97-41c4-8b13-25c3d36acaa5 · outbound

This paper cites TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles.

Exploring Timeline Control for Facial Motion Generation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.626113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.626113Z digest=sha256:7b402a161c4bb092bed0260805cb8b7f831f2119fe42e977bc4dc0a0f5226518

Observation 2e233ae7-dcc4-47b7-8109-b47e81b207eb · outbound

This paper cites Amass: Archive of motion capture as surface shapes.

Exploring Timeline Control for Facial Motion Generation Amass: Archive of motion capture as surface shapes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.755671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.755671Z digest=sha256:5c1e50e426474e1b2d2422d686b1b701d1fb27ad9715a80c23540261c7e3cd38

Observation 66ea1bf1-d86a-4fab-b44e-a970f349f47f · outbound

This paper cites Autoplait: Automatic mining of co-evolving time se- quences.

Exploring Timeline Control for Facial Motion Generation Autoplait: Automatic mining of co-evolving time se- quences

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.018947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.855167Z digest=sha256:139a2a601c7f779f0f365e53e84e70e6e4da63621948614125fbba58a59ee9a8

Observation 3d041e45-53be-4903-9741-7e1c872a4d74 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

Exploring Timeline Control for Facial Motion Generation From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.836240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:50.954333Z digest=sha256:8e1a2275dc7124e262c60bb7cb4954cb32a14634148ced14438b058b5c40e942

Observation fba965a2-de61-4c99-bb49-a20b5e3efd38 · outbound

This paper cites ScanTalk: 3D Talking Heads from Unregistered Scans.

Exploring Timeline Control for Facial Motion Generation ScanTalk: 3D Talking Heads from Unregistered Scans

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.159075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:51.060403Z digest=sha256:9bbf0110355ac85253b52cb3fbffa4e24abc4bb3879a202ab1895fdd6aa7784e

Observation e67917d9-6478-48db-b074-3fc7c96540ea · outbound

This paper cites Generating facial expressions for speech.Cognitive science, 20(1):1–46, 1996.

Exploring Timeline Control for Facial Motion Generation Generating facial expressions for speech.Cognitive science, 20(1):1–46, 1996

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.584576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:51.131465Z digest=sha256:af4b9891af0f06ec99c0fe319aae68fb1b6ce7347beccb3ac80bbd3a514aa210

Observation d02df5be-365c-4f9e-8465-673fedcf9f4a · outbound

This paper cites Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion.

Exploring Timeline Control for Facial Motion Generation Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.332064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:51.228539Z digest=sha256:422bb085fe4f51393b7907444ba82e6ed6322f0886c6781358a6a8b1514e08d6

Observation dea93a64-0405-45d3-8630-0b68523be7b5 · outbound

This paper cites Multi-track timeline control for text-driven 3d human motion generation.

Exploring Timeline Control for Facial Motion Generation Multi-track timeline control for text-driven 3d human motion generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.174781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:51.332307Z digest=sha256:f010abfe5b2e3d08c0d7116643fbccff5d3443c91b172db27e43e5ad7ae36e26

Observation 2e22b04f-63c2-49c5-be05-b3c7611b7687 · outbound

This paper cites The kit motion-language dataset.Big data, 4(4):236–252,.

Exploring Timeline Control for Facial Motion Generation The kit motion-language dataset.Big data, 4(4):236–252,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:51.419858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:51.419858Z digest=sha256:3b6adcfb2df38469c21620e9375bb9c4dbb0cd32818be55a0ab21fdc5851a6a6

Observation 2558f091-5f20-4e36-a206-ab1e78618df1 · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

Exploring Timeline Control for Facial Motion Generation A lip sync expert is all you need for speech to lip generation in the wild

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.952512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:51.551263Z digest=sha256:e4f56f847bf4f6adc751cf6271a49f553bdb1eba293baf17b6f589d6bd928130

Observation 894cb30f-0a81-43d1-95e1-5d2ef6b154ec · outbound

This paper cites Cas(me) 2 : A database for sponta- neous macro-expression and micro-expression spotting and recognition.IEEE Transactions on Affective Computing, 9 (4):424–436, 2018.

Exploring Timeline Control for Facial Motion Generation Cas(me) 2 : A database for sponta- neous macro-expression and micro-expression spotting and recognition.IEEE Transactions on Affective Computing, 9 (4):424–436, 2018

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.740373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:51.652249Z digest=sha256:6a88061e40b4a34571f144716e7057d1aee0eda53ae7c7328cd86c9bc8f2a2cb

Observation 5934baf1-ad13-4c1b-b06c-947ff71c23f1 · outbound

This paper cites Human Motion Diffusion as a Generative Prior.

Exploring Timeline Control for Facial Motion Generation Human Motion Diffusion as a Generative Prior

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:51.788235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:51.788235Z digest=sha256:b043494736b0b26fda859c16a76aba5b286256080e06897dc89c5b1572784281

Observation f04e6d3f-db93-4a87-a9ea-b8a7ffbe50fd · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Exploring Timeline Control for Facial Motion Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:51.922931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:51.922931Z digest=sha256:c4918ff23931301f68d0d0c3b2fc8b593566e682b0e027d74456d867a2fd144c

Observation d6088f83-614d-48a6-a733-135abd67b98c · outbound

This paper cites Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4):1–9, 2024.

Exploring Timeline Control for Facial Motion Generation Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4):1–9, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.569621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:52.027296Z digest=sha256:0081ec0159705aee6103ed5d4e134ae48e5c30cfc5672159cf9298df3450c25c

Observation def676ab-7ab8-4b0a-bb2d-d85fc603ab6b · outbound

This paper cites Edtalk: Effi- cient disentanglement for emotional talking head synthesis.

Exploring Timeline Control for Facial Motion Generation Edtalk: Effi- cient disentanglement for emotional talking head synthesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.129530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.129530Z digest=sha256:93dbdd6b7bc27c2510fc7312be029a7a32eb7e0a8530d2a7e527a96a90a86cf9

Observation 0994e012-ab48-468a-aab5-4d606b581965 · outbound

This paper cites EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions.

Exploring Timeline Control for Facial Motion Generation EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.227418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.227418Z digest=sha256:a097ebdffce4849c6832091a656ebbd52732cfc50f12db9ad84cc7d787dac9a4

Observation 63d6ad76-c6d6-4f36-a81e-063e2c86f0d4 · outbound

This paper cites AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents.

Exploring Timeline Control for Facial Motion Generation AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.324149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.324149Z digest=sha256:ae49452518644866476ebbf67c02afb52e25b003e4e595f638fa6f2510989433

Observation ee323b2a-e72e-4563-ac71-c1c1da63d329 · outbound

This paper cites Mead: A large-scale audio-visual dataset for emotional talking-face generation.

Exploring Timeline Control for Facial Motion Generation Mead: A large-scale audio-visual dataset for emotional talking-face generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.382601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:52.431506Z digest=sha256:c1d391eb790048c5a557b26855de1a5dbbda7ad9bb0727d105fbf165d828ba0a

Observation 89f5637d-4240-4a0f-92d6-10bbdb6f5231 · outbound

This paper cites Faceverse: a fine-grained and detail- controllable 3d face morphable model from a hybrid dataset.

Exploring Timeline Control for Facial Motion Generation Faceverse: a fine-grained and detail- controllable 3d face morphable model from a hybrid dataset

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.144125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:52.513665Z digest=sha256:a66026ea47cbe8b1b79b16faa286cbb237e10006d91dcaa2466d925c485600e7

Observation e0b2f31c-4e2d-4d98-a6de-9374520f3c1d · outbound

This paper cites Styletalk++: A unified framework for controlling the speaking styles of talking heads.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

Exploring Timeline Control for Facial Motion Generation Styletalk++: A unified framework for controlling the speaking styles of talking heads.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.921199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:52.583573Z digest=sha256:6510dfc39129e8c9ff9bc4e34f1d47a72d6cf95b506095a4951d549a2738a252

Observation 0c779b70-8aea-4196-8848-70e36ee8f2fb · outbound

This paper cites Mes- net: A convolutional neural network for spotting multi-scale micro-expression intervals in long videos.IEEE Transac- tions on Image Processing, 30:3956–3969, 2021.

Exploring Timeline Control for Facial Motion Generation Mes- net: A convolutional neural network for spotting multi-scale micro-expression intervals in long videos.IEEE Transac- tions on Image Processing, 30:3956–3969, 2021

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.687769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:52.674076Z digest=sha256:c9f355296574a54eb0653b230ac112fc69435f7b13edbe257388040ab70ce1d0

Observation 5d68c1b8-b243-44ec-bda2-288fa97dce0e · outbound

This paper cites InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation.

Exploring Timeline Control for Facial Motion Generation InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.775927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.775927Z digest=sha256:c0205bfdbaa1a2d1ac25013c7bde76435189ae97692c1d4740bf319dbf13f948

Observation 1f0bcfd1-d675-4f87-ae55-6c336198b589 · outbound

This paper cites AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation.

Exploring Timeline Control for Facial Motion Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.852981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.852981Z digest=sha256:c38efd0a0c89b58631b33cf599eba32cf467dbf8ebfc6ae3271f60add5fe9cdf

Observation 5f9d5c69-790c-4151-a5fd-34304171d249 · outbound

This paper cites MMHead: Towards Fine-grained Multi-modal 3D Facial Animation.

Exploring Timeline Control for Facial Motion Generation MMHead: Towards Fine-grained Multi-modal 3D Facial Animation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.938222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.938222Z digest=sha256:3c74192164c2a53d478f7f9e5eb0964711f189689d0231640c74a9e1c0eff58d

Observation d2c14f9a-e377-4885-858a-e8fec47d3c01 · outbound

This paper cites Codetalker: Speech-driven 3d facial animation with discrete motion prior.

Exploring Timeline Control for Facial Motion Generation Codetalker: Speech-driven 3d facial animation with discrete motion prior

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.476783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.034434Z digest=sha256:6c41adb0f283abc7ca4a137b77013fc7865e3d08334866e9a6fd7e67c0342181

Observation 2bc0af3b-65de-4f02-bee0-74fa0848f548 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

Exploring Timeline Control for Facial Motion Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:53.135171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:53.135171Z digest=sha256:630dfc38e00a4204f0ce10a64d4ab27ade7cd26ebee6e9d2753b96b3824396b0

Observation 888ee6e9-f674-4db1-8804-01986df3a000 · outbound

This paper cites VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time.

Exploring Timeline Control for Facial Motion Generation VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:53.235375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:53.235375Z digest=sha256:8adbef0d8e4eb22b65ec0bd226653e99a8d6734ce1eb793ed6df0bec559a0e93

Observation c4d05344-ae25-49bd-b38e-f2ce30f0c207 · outbound

This paper cites Probabilistic speech- driven 3d facial motion synthesis: New benchmarks meth- ods and applications.

Exploring Timeline Control for Facial Motion Generation Probabilistic speech- driven 3d facial motion synthesis: New benchmarks meth- ods and applications

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.258184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.331132Z digest=sha256:31bc513eb315c7b845c29db1207da64449bb5b90791fc190e7d5998a69f8c030

Observation 59a38717-3606-49db-9fec-a5fb10493703 · outbound

This paper cites Samm long videos: A spontaneous facial micro-and macro- expressions dataset.

Exploring Timeline Control for Facial Motion Generation Samm long videos: A spontaneous facial micro-and macro- expressions dataset

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.039505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.398064Z digest=sha256:94e932a06cefa364f985b8930628580c3b2edfdca32c150395f3747487c2dba7

Observation a4ab37a4-bc3d-4553-bda2-888af71a5990 · outbound

This paper cites 3d-cnn for facial micro-and macro-expression spotting on long video sequences using temporal oriented reference frame.

Exploring Timeline Control for Facial Motion Generation 3d-cnn for facial micro-and macro-expression spotting on long video sequences using temporal oriented reference frame

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:57.843948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.512356Z digest=sha256:3a0d2c3ed0d1edbaad23cd233b1cba32fe69fb64915a470773f9a637bf575def

Observation 5769b326-6566-4671-a8bf-db05c7da5412 · outbound

This paper cites GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis.

Exploring Timeline Control for Facial Motion Generation GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:53.598544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:53.598544Z digest=sha256:ce677f355f6bc4b146ec3724d4aa53022214970dd1a547bbddad75dab5c01906

Observation d4088ada-b6a2-4119-818e-defb0818f912 · outbound

This paper cites Facial expression spotting based on optical flow features.

Exploring Timeline Control for Facial Motion Generation Facial expression spotting based on optical flow features

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:57.505213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.737429Z digest=sha256:975682252f04b73a8bf65c42344b943e4b1376896c7366effbc8a5f6570509fb

Observation 674c982a-6c39-4fdf-b4bf-a1cb43c28784 · outbound

This paper cites CelebV-Text: A large-scale facial text-video dataset.

Exploring Timeline Control for Facial Motion Generation CelebV-Text: A large-scale facial text-video dataset

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:57.077407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.833576Z digest=sha256:36ce7c206c2dee57739e1b326c008820b78fb616bc15f18e5869ce64dac7bd9a

Observation 8dd5130e-983b-4f44-9e1c-42c0c2714fbc · outbound

This paper cites Talking Head Generation with Probabilistic Audio-to-Visual Diffusion Priors.

Exploring Timeline Control for Facial Motion Generation Talking Head Generation with Probabilistic Audio-to-Visual Diffusion Priors

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:54.743484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:53.927766Z digest=sha256:e173a20239b0704eba456db2e6191ed7b16fa971aa620d8161e569baf61160ac

Observation 0037c0d5-adbf-4eea-9e87-e6602698941c · outbound

This paper cites Talking head generation with probabilistic audio-to-visual diffusion priors.

Exploring Timeline Control for Facial Motion Generation Talking head generation with probabilistic audio-to-visual diffusion priors

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:56.774445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:54.016200Z digest=sha256:d7511712e96c8e070655c9579442b61813df37b889ef1b52bdc52077571ed103

Observation 85deecbd-cb63-473f-9204-68e68961d28d · outbound

This paper cites PersonaTalk: Bring Attention to Your Persona in Visual Dubbing.

Exploring Timeline Control for Facial Motion Generation PersonaTalk: Bring Attention to Your Persona in Visual Dubbing

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:54.502560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:54.116546Z digest=sha256:22923f4cfdec8326724894bf29e88b06800d58d7bdc6c1ebf091ff99d4973fa5

Observation f548c61a-4c08-4251-888e-d79c7a6c3c74 · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation.

Exploring Timeline Control for Facial Motion Generation Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:56.504965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:54.196437Z digest=sha256:b8a0763ee7d55599cef191b7169899971f1c47bae5b78ff3944db7ad0f80feb4

Observation a075f5fe-8e44-415f-b1c2-12a5ab345e62 · outbound

This paper cites Media2face: Co-speech facial animation gen- eration with multi-modality guidance.

Exploring Timeline Control for Facial Motion Generation Media2face: Co-speech facial animation gen- eration with multi-modality guidance

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:56.242846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:49:54.279621Z digest=sha256:f2b2411544807f2108ae0d6f01d9cfa95b46ed497a63a09ba085c098cafb2cdd

Pith citing papers

No inbound Pith citation observations are available.