Pith. sign in

Paper Citation Record · LEDGER

Exploring Timeline Control for Facial Motion Generation

As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2505.20861.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20861 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:54.279621Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact6
  • verified fuzzy38
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e8f3d26e-c437-48ba-8b30-7823348a0832 · outbound

This paper cites Facetalk: Audio-driven motion diffusion for neural parametric head models.

Exploring Timeline Control for Facial Motion Generation Facetalk: Audio-driven motion diffusion for neural parametric head models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.840247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.163584Z digest=sha256:e4197b0feb61f2eb7f0e7346dd315165823069e072ef8421c129fe7fc15d89c6

Observation fe2f996b-c31d-4885-abb5-64bffbfc8bab · outbound

This paper cites Teach: Temporal action composition for 3d hu- mans.

Exploring Timeline Control for Facial Motion Generation Teach: Temporal action composition for 3d hu- mans

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.573660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.274962Z digest=sha256:d0e384a82128cd7ddff07185319b4ab53feebd71dfd2f19f802a6be32c120659

Observation 6f71fbf8-c1c5-4618-8f0d-ab45897180f3 · outbound

This paper cites Seamless human motion composition with blended posi- tional encodings.

Exploring Timeline Control for Facial Motion Generation Seamless human motion composition with blended posi- tional encodings

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.434110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.409933Z digest=sha256:4802983c25d45c47333ae684250687b550a84f07e8f4f2a74f8b97912f07be54

Observation 6ae2a930-2173-4b1b-8a5d-f18abc55febe · outbound

This paper cites Video generation models as world simulators.

Exploring Timeline Control for Facial Motion Generation Video generation models as world simulators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.260523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.473574Z digest=sha256:da82a555a73a4a1960a9fc3ff69d1e35795c9d8c22e5fb2b365ca58ca2c06961

Observation 27264f34-ed73-43a0-8ca8-39f359726ba2 · outbound

This paper cites Animated conversation: rule- based generation of facial expression, gesture & spoken in- tonation for multiple conversational agents.

Exploring Timeline Control for Facial Motion Generation Animated conversation: rule- based generation of facial expression, gesture & spoken in- tonation for multiple conversational agents

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:04.074037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.530163Z digest=sha256:eb5642ec0d845274129125f8d73f1333d4b2ac9c447613ef3fe37c071fa07500

Observation 17f683c4-da09-48fb-9451-d598bf33cdd2 · outbound

This paper cites EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions.

Exploring Timeline Control for Facial Motion Generation EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:48.623861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:48.623861Z digest=sha256:9e418e998a1d86544b520b8078cfd8a28c3a1621af7ecdd96055460526dc9aff

Observation 07b1eb48-97c4-4717-bd2a-aebb57b7b7d4 · outbound

This paper cites Capture, learning, and syn- thesis of 3d speaking styles.

Exploring Timeline Control for Facial Motion Generation Capture, learning, and syn- thesis of 3d speaking styles

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.883957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.727652Z digest=sha256:d624ef66585cfeea159c04d4d6810b4ae2d2acc6a2fb0c9a0d1b6a01d1643eff

Observation e0ae3a21-b1c7-4a09-b263-56c7e858a2b1 · outbound

This paper cites Emotional Speech-Driven Animation with Content-Emotion Disentanglement.

Exploring Timeline Control for Facial Motion Generation Emotional Speech-Driven Animation with Content-Emotion Disentanglement

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:48.844014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:48.844014Z digest=sha256:7e4683443a3a470dd1c82914c081c36c2830aa37f3b0f8f77c810e390bd704a0

Observation b1b53d5c-bd2d-48bb-aaf1-bad0ce9fdec8 · outbound

This paper cites SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting.

Exploring Timeline Control for Facial Motion Generation SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.964439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:48.933064Z digest=sha256:b90a96f810dc4740651e2c3da6a0b4e1598782b558160142c1a0602b57f00c1b

Observation 822efe5f-76e2-4f27-8abc-8f0a3ce42fde · outbound

This paper cites Facial action coding system (facs).A Human Face, Salt Lake City, 2002.

Exploring Timeline Control for Facial Motion Generation Facial action coding system (facs).A Human Face, Salt Lake City, 2002

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.735132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.028666Z digest=sha256:c5a5aeea1298782354190b189b22f62736cbd9885375ed583593abc1d8160ab2

Observation a7ddf34d-a855-4df6-8cb1-0381a09944cd · outbound

This paper cites UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model.

Exploring Timeline Control for Facial Motion Generation UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified Model

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.678185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.139392Z digest=sha256:f48f35fca80222ecfe8cb4f3f950c8c46c17a448d4ec2d382cef00ef738cdf89

Observation fe52b9b4-c4ad-4789-a00d-a04be8fa1760 · outbound

This paper cites Faceformer: Speech-driven 3d facial anima- tion with transformers.

Exploring Timeline Control for Facial Motion Generation Faceformer: Speech-driven 3d facial anima- tion with transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.675102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.256063Z digest=sha256:dd891ba65276800cfc89be3ac89a877aaa3304a847f626624b1923dbfa3f31f5

Observation d4f137cf-6685-47d6-b02b-643efebeddff · outbound

This paper cites Affective Faces for Goal-Driven Dyadic Communication.

Exploring Timeline Control for Facial Motion Generation Affective Faces for Goal-Driven Dyadic Communication

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.337095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.337095Z digest=sha256:15b2c33ccc86b9b45832a2b107dd2bbace50fca69791deda110c2fabcdfdb01d

Observation aad20c1e-9712-48f9-bd8c-dbfc7eebeb56 · outbound

This paper cites ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer.

Exploring Timeline Control for Facial Motion Generation ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.426666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.426666Z digest=sha256:6583d45aeff9c106818dac2a869b0d6d02c3e877fe5a7e15ae3ca83d6e516f8c

Observation 22f7283d-2b19-4f7d-8bef-763e52069d5e · outbound

This paper cites Ac- tion2motion: Conditioned generation of 3d human motions.

Exploring Timeline Control for Facial Motion Generation Ac- tion2motion: Conditioned generation of 3d human motions

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.328214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.488190Z digest=sha256:d19c18927ddb947f7bfa7c858e511cb062047cd15ac80c02845389fbbd94d6a1

Observation 8836d187-ae97-423d-9065-3015b5ce0353 · outbound

This paper cites Generating diverse and natural 3d human motions from text.

Exploring Timeline Control for Facial Motion Generation Generating diverse and natural 3d human motions from text

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:03.054404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.614660Z digest=sha256:376745c68dfd110e590b1e8597122cd67bcc2ee245da1a14007447093449912c

Observation be5a18b7-4a50-4142-840c-442dc86ac749 · outbound

This paper cites Micro-expression spotting with multi-scale local transformer in long videos.Pattern Recognition Letters, 168:146–152,.

Exploring Timeline Control for Facial Motion Generation Micro-expression spotting with multi-scale local transformer in long videos.Pattern Recognition Letters, 168:146–152,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:02.655173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.702574Z digest=sha256:8235a9f1a946026f69f29865d6b6dbecc1e515e2d56bce053e2156724f0b8d94

Observation 2b08c393-1a02-4c93-a9f9-d48d1b22493c · outbound

This paper cites Toeplitz inverse covariance-based clustering of multivariate time series data.

Exploring Timeline Control for Facial Motion Generation Toeplitz inverse covariance-based clustering of multivariate time series data

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:02.255990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:49.787589Z digest=sha256:28743e8677d1c0f6f680583fd5177029c2b9d75a63c9c200a5ce0d89c8a4ed4a

Observation 22a6ae26-72c3-4ca4-9321-01241a6c5239 · outbound

This paper cites GAIA: Zero-shot Talking Avatar Generation.

Exploring Timeline Control for Facial Motion Generation GAIA: Zero-shot Talking Avatar Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.888105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.888105Z digest=sha256:4110f0435ca5ce2e3ccf0883dd5b03a924a92c2dc86a5cbd9b609931bd140786

Observation ee7f8684-adb2-42ab-aa47-4dbbccd8a1cd · outbound

This paper cites Classifier-free diffusion guidance.

Exploring Timeline Control for Facial Motion Generation Classifier-free diffusion guidance

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:49.960005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:49.960005Z digest=sha256:5ee5ea1113ecd137d7e4548c57e0d336169d664c1f35c4f82b918ee391e0f789

Observation b6c6ff47-0767-434e-9a14-75e61fffe57f · outbound

This paper cites Denoising dif- fusion probabilistic models.

Exploring Timeline Control for Facial Motion Generation Denoising dif- fusion probabilistic models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.993307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.036450Z digest=sha256:d902f9e21ec29c5f7ac6098ec75b3f0d905a5eb96e49e1b258471a0070f7d921

Observation 4f32f28f-b3e8-4e23-82e4-0c8a45e95da2 · outbound

This paper cites Multi-aspect mining of complex sensor sequences.

Exploring Timeline Control for Facial Motion Generation Multi-aspect mining of complex sensor sequences

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.856176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.138655Z digest=sha256:42ba80b67a914757a3d692e959841711078f256641cc56a8b2f6464d03b8b438

Observation daea6143-b097-4dcd-9655-0baf2b2e4c00 · outbound

This paper cites Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency.

Exploring Timeline Control for Facial Motion Generation Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.225081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.225081Z digest=sha256:6f4b10ebe507660e55f635b726ed977e2b0977e35513d19c1a67aeb1d0061e7b

Observation e8a73c3e-bb03-4a1f-adb0-f4d6a886b69e · outbound

This paper cites Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017.

Exploring Timeline Control for Facial Motion Generation Audio-driven facial animation by joint end- to-end learning of pose and emotion.ACM Transactions on Graphics (TOG), 36(4):1–12, 2017

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.699979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.291313Z digest=sha256:165e63e5d85aa6aece3f112fe9d9dec09a99654aac9e1e174438dcf2cf0edbf5

Observation 490f9ba2-21a9-4c6a-8d8c-fce972478b48 · outbound

This paper cites Kmtalk: Speech-driven 3d facial animation with key motion embedding.

Exploring Timeline Control for Facial Motion Generation Kmtalk: Speech-driven 3d facial animation with key motion embedding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.508258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.366335Z digest=sha256:16eadab13360df19dabb829a337dd7391cbdf2486b29f1c5c25f515ef47d6439

Observation 81dbbc9c-8254-40e5-a8a8-75a950bf2ed4 · outbound

This paper cites Talkinggaussian: Structure-persistent 3d talking head synthesis via gaussian splatting.

Exploring Timeline Control for Facial Motion Generation Talkinggaussian: Structure-persistent 3d talking head synthesis via gaussian splatting

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.301208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.439421Z digest=sha256:5ae1b80c488a11187603d6809648add1f3d8a096b29b6925cb69da322eecb1fd

Observation b53e64dc-1548-4177-8884-3f649be8e618 · outbound

This paper cites PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation.

Exploring Timeline Control for Facial Motion Generation PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.422351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.530894Z digest=sha256:51ceaf749473ddc11c0d64eb22610e0978ed735858499dcfd52b80fa7b1d999c

Observation 6c502d7a-fa97-41c4-8b13-25c3d36acaa5 · outbound

This paper cites TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles.

Exploring Timeline Control for Facial Motion Generation TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.626113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.626113Z digest=sha256:7b402a161c4bb092bed0260805cb8b7f831f2119fe42e977bc4dc0a0f5226518

Observation 2e233ae7-dcc4-47b7-8109-b47e81b207eb · outbound

This paper cites Amass: Archive of motion capture as surface shapes.

Exploring Timeline Control for Facial Motion Generation Amass: Archive of motion capture as surface shapes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:50.755671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:50.755671Z digest=sha256:5c1e50e426474e1b2d2422d686b1b701d1fb27ad9715a80c23540261c7e3cd38

Observation 66ea1bf1-d86a-4fab-b44e-a970f349f47f · outbound

This paper cites Autoplait: Automatic mining of co-evolving time se- quences.

Exploring Timeline Control for Facial Motion Generation Autoplait: Automatic mining of co-evolving time se- quences

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:01.018947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.855167Z digest=sha256:d335f3609563a10f59255f6aceae4360eb984c3952dc3d39254f4a06254a8759

Observation 3d041e45-53be-4903-9741-7e1c872a4d74 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

Exploring Timeline Control for Facial Motion Generation From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.836240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:50.954333Z digest=sha256:491a63bcae92f5e918ed6c6d51394e365ae0bc6832af93f6b6ca171c831cd4db

Observation fba965a2-de61-4c99-bb49-a20b5e3efd38 · outbound

This paper cites ScanTalk: 3D Talking Heads from Unregistered Scans.

Exploring Timeline Control for Facial Motion Generation ScanTalk: 3D Talking Heads from Unregistered Scans

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:55.159075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:51.060403Z digest=sha256:d4f4d8a7b4f459e536121253b7e8e418077ba4fabbde7d6264f6d87b6cd6dbb0

Observation e67917d9-6478-48db-b074-3fc7c96540ea · outbound

This paper cites Generating facial expressions for speech.Cognitive science, 20(1):1–46, 1996.

Exploring Timeline Control for Facial Motion Generation Generating facial expressions for speech.Cognitive science, 20(1):1–46, 1996

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.584576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:51.131465Z digest=sha256:fcf62e8e0c90bf03ddd87eb374cf922e9e4627f7cec395dddfb84cbd3f244ef0

Observation d02df5be-365c-4f9e-8465-673fedcf9f4a · outbound

This paper cites Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion.

Exploring Timeline Control for Facial Motion Generation Emotalk: Speech-driven emotional disentanglement for 3d face anima- tion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.332064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:51.228539Z digest=sha256:98d8d79743f746429ea26fb92de732a5955661fdc707df04d3d4cbb2def4f0b2

Observation dea93a64-0405-45d3-8630-0b68523be7b5 · outbound

This paper cites Multi-track timeline control for text-driven 3d human motion generation.

Exploring Timeline Control for Facial Motion Generation Multi-track timeline control for text-driven 3d human motion generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:50:00.174781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:51.332307Z digest=sha256:02463c44919c305ce42d149fe093f015650b7533cec55526ce49e183777c6da7

Observation 2e22b04f-63c2-49c5-be05-b3c7611b7687 · outbound

This paper cites The kit motion-language dataset.Big data, 4(4):236–252,.

Exploring Timeline Control for Facial Motion Generation The kit motion-language dataset.Big data, 4(4):236–252,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:51.419858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:51.419858Z digest=sha256:3b6adcfb2df38469c21620e9375bb9c4dbb0cd32818be55a0ab21fdc5851a6a6

Observation 2558f091-5f20-4e36-a206-ab1e78618df1 · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

Exploring Timeline Control for Facial Motion Generation A lip sync expert is all you need for speech to lip generation in the wild

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.952512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:51.551263Z digest=sha256:ca3de64f13aa9fb663bd88025d4aba87d9565318324661e8820c5d07daed04c7

Observation 894cb30f-0a81-43d1-95e1-5d2ef6b154ec · outbound

This paper cites Cas(me) 2 : A database for sponta- neous macro-expression and micro-expression spotting and recognition.IEEE Transactions on Affective Computing, 9 (4):424–436, 2018.

Exploring Timeline Control for Facial Motion Generation Cas(me) 2 : A database for sponta- neous macro-expression and micro-expression spotting and recognition.IEEE Transactions on Affective Computing, 9 (4):424–436, 2018

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.740373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:51.652249Z digest=sha256:9d09d48a0d7a355544fd03988d81f6c4fc29ef7d09ebca3c863a643bada5b2cf

Observation 5934baf1-ad13-4c1b-b06c-947ff71c23f1 · outbound

This paper cites Human Motion Diffusion as a Generative Prior.

Exploring Timeline Control for Facial Motion Generation Human Motion Diffusion as a Generative Prior

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:51.788235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:51.788235Z digest=sha256:b043494736b0b26fda859c16a76aba5b286256080e06897dc89c5b1572784281

Observation f04e6d3f-db93-4a87-a9ea-b8a7ffbe50fd · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Exploring Timeline Control for Facial Motion Generation Score-Based Generative Modeling through Stochastic Differential Equations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:51.922931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:51.922931Z digest=sha256:c4918ff23931301f68d0d0c3b2fc8b593566e682b0e027d74456d867a2fd144c

Observation d6088f83-614d-48a6-a733-135abd67b98c · outbound

This paper cites Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4):1–9, 2024.

Exploring Timeline Control for Facial Motion Generation Diffposetalk: Speech-driven stylistic 3d facial animation and head pose generation via diffusion models.ACM Transactions on Graphics (TOG), 43(4):1–9, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.569621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:52.027296Z digest=sha256:b4b3b1959d1bdb43db0f561af472dd32847c0f65b07558ae341ceedfe2d4e18d

Observation def676ab-7ab8-4b0a-bb2d-d85fc603ab6b · outbound

This paper cites Edtalk: Effi- cient disentanglement for emotional talking head synthesis.

Exploring Timeline Control for Facial Motion Generation Edtalk: Effi- cient disentanglement for emotional talking head synthesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.129530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.129530Z digest=sha256:93dbdd6b7bc27c2510fc7312be029a7a32eb7e0a8530d2a7e527a96a90a86cf9

Observation 0994e012-ab48-468a-aab5-4d606b581965 · outbound

This paper cites EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions.

Exploring Timeline Control for Facial Motion Generation EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.227418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.227418Z digest=sha256:a097ebdffce4849c6832091a656ebbd52732cfc50f12db9ad84cc7d787dac9a4

Observation 63d6ad76-c6d6-4f36-a81e-063e2c86f0d4 · outbound

This paper cites AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents.

Exploring Timeline Control for Facial Motion Generation AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.324149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.324149Z digest=sha256:ae49452518644866476ebbf67c02afb52e25b003e4e595f638fa6f2510989433

Observation ee323b2a-e72e-4563-ac71-c1c1da63d329 · outbound

This paper cites Mead: A large-scale audio-visual dataset for emotional talking-face generation.

Exploring Timeline Control for Facial Motion Generation Mead: A large-scale audio-visual dataset for emotional talking-face generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.382601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:52.431506Z digest=sha256:f64bd74311832ff54d361e5ac9c160c210e4961b7ede862180584099944e8dcd

Observation 89f5637d-4240-4a0f-92d6-10bbdb6f5231 · outbound

This paper cites Faceverse: a fine-grained and detail- controllable 3d face morphable model from a hybrid dataset.

Exploring Timeline Control for Facial Motion Generation Faceverse: a fine-grained and detail- controllable 3d face morphable model from a hybrid dataset

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:59.144125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:52.513665Z digest=sha256:e5831ffddcfbef902928ed4578cf091ec8872759a931fb95979bbea5b97b9619

Observation e0b2f31c-4e2d-4d98-a6de-9374520f3c1d · outbound

This paper cites Styletalk++: A unified framework for controlling the speaking styles of talking heads.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

Exploring Timeline Control for Facial Motion Generation Styletalk++: A unified framework for controlling the speaking styles of talking heads.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.921199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:52.583573Z digest=sha256:f99bbdd889aa85f5b6c73d612d425f3890bf09105bb781886c11a3ffe255b297

Observation 0c779b70-8aea-4196-8848-70e36ee8f2fb · outbound

This paper cites Mes- net: A convolutional neural network for spotting multi-scale micro-expression intervals in long videos.IEEE Transac- tions on Image Processing, 30:3956–3969, 2021.

Exploring Timeline Control for Facial Motion Generation Mes- net: A convolutional neural network for spotting multi-scale micro-expression intervals in long videos.IEEE Transac- tions on Image Processing, 30:3956–3969, 2021

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.687769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:52.674076Z digest=sha256:e37a4efc8ea21c036b1fe0481b033908e755803f436a61c9bc6569ce7d401e25

Observation 5d68c1b8-b243-44ec-bda2-288fa97dce0e · outbound

This paper cites InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation.

Exploring Timeline Control for Facial Motion Generation InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.775927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.775927Z digest=sha256:c0205bfdbaa1a2d1ac25013c7bde76435189ae97692c1d4740bf319dbf13f948

Observation 1f0bcfd1-d675-4f87-ae55-6c336198b589 · outbound

This paper cites AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation.

Exploring Timeline Control for Facial Motion Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.852981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.852981Z digest=sha256:c38efd0a0c89b58631b33cf599eba32cf467dbf8ebfc6ae3271f60add5fe9cdf

Observation 5f9d5c69-790c-4151-a5fd-34304171d249 · outbound

This paper cites MMHead: Towards Fine-grained Multi-modal 3D Facial Animation.

Exploring Timeline Control for Facial Motion Generation MMHead: Towards Fine-grained Multi-modal 3D Facial Animation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.938222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.938222Z digest=sha256:3c74192164c2a53d478f7f9e5eb0964711f189689d0231640c74a9e1c0eff58d

Observation d2c14f9a-e377-4885-858a-e8fec47d3c01 · outbound

This paper cites Codetalker: Speech-driven 3d facial animation with discrete motion prior.

Exploring Timeline Control for Facial Motion Generation Codetalker: Speech-driven 3d facial animation with discrete motion prior

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.476783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.034434Z digest=sha256:5375dfb8612dffd44b3ba4efe60d4f21484e156498451e35997f9c82f08a1322

Observation 2bc0af3b-65de-4f02-bee0-74fa0848f548 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

Exploring Timeline Control for Facial Motion Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:53.135171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:53.135171Z digest=sha256:630dfc38e00a4204f0ce10a64d4ab27ade7cd26ebee6e9d2753b96b3824396b0

Observation 888ee6e9-f674-4db1-8804-01986df3a000 · outbound

This paper cites VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time.

Exploring Timeline Control for Facial Motion Generation VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:53.235375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:53.235375Z digest=sha256:8adbef0d8e4eb22b65ec0bd226653e99a8d6734ce1eb793ed6df0bec559a0e93

Observation c4d05344-ae25-49bd-b38e-f2ce30f0c207 · outbound

This paper cites Probabilistic speech- driven 3d facial motion synthesis: New benchmarks meth- ods and applications.

Exploring Timeline Control for Facial Motion Generation Probabilistic speech- driven 3d facial motion synthesis: New benchmarks meth- ods and applications

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.258184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.331132Z digest=sha256:b18e78490ec54c70f90dbc4031e453533616ddf6f2c4a65910daa325f1681111

Observation 59a38717-3606-49db-9fec-a5fb10493703 · outbound

This paper cites Samm long videos: A spontaneous facial micro-and macro- expressions dataset.

Exploring Timeline Control for Facial Motion Generation Samm long videos: A spontaneous facial micro-and macro- expressions dataset

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:58.039505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.398064Z digest=sha256:1cb8c773660e57d1eb7d13df6a1afc4f8042c17c8ae71a97651d59618755475a

Observation a4ab37a4-bc3d-4553-bda2-888af71a5990 · outbound

This paper cites 3d-cnn for facial micro-and macro-expression spotting on long video sequences using temporal oriented reference frame.

Exploring Timeline Control for Facial Motion Generation 3d-cnn for facial micro-and macro-expression spotting on long video sequences using temporal oriented reference frame

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:57.843948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.512356Z digest=sha256:611d4896bf7284c6642fa8aa1d78d6598e2009bfcf45fac8f25cbe6f78e7fb2e

Observation 5769b326-6566-4671-a8bf-db05c7da5412 · outbound

This paper cites GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis.

Exploring Timeline Control for Facial Motion Generation GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:53.598544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:53.598544Z digest=sha256:ce677f355f6bc4b146ec3724d4aa53022214970dd1a547bbddad75dab5c01906

Observation d4088ada-b6a2-4119-818e-defb0818f912 · outbound

This paper cites Facial expression spotting based on optical flow features.

Exploring Timeline Control for Facial Motion Generation Facial expression spotting based on optical flow features

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:57.505213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.737429Z digest=sha256:4835ffa3568e4a94b983b7f76a73260d0a0e6b74ea8f17367f0ea8b2a4b28892

Observation 674c982a-6c39-4fdf-b4bf-a1cb43c28784 · outbound

This paper cites CelebV-Text: A large-scale facial text-video dataset.

Exploring Timeline Control for Facial Motion Generation CelebV-Text: A large-scale facial text-video dataset

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:57.077407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.833576Z digest=sha256:3796a7c276f63da5947b5313f8b9ebb439c45eacf4b32deae4821a2394d48617

Observation 8dd5130e-983b-4f44-9e1c-42c0c2714fbc · outbound

This paper cites Talking Head Generation with Probabilistic Audio-to-Visual Diffusion Priors.

Exploring Timeline Control for Facial Motion Generation Talking Head Generation with Probabilistic Audio-to-Visual Diffusion Priors

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:54.743484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:53.927766Z digest=sha256:5139d490dbe2c5971cb37842a7638f5c6b14e705ae0b4333e6f5f77970ba204d

Observation 0037c0d5-adbf-4eea-9e87-e6602698941c · outbound

This paper cites Talking head generation with probabilistic audio-to-visual diffusion priors.

Exploring Timeline Control for Facial Motion Generation Talking head generation with probabilistic audio-to-visual diffusion priors

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:56.774445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:54.016200Z digest=sha256:fb7d515f577ca58b2397aa6f4df1f173df5b31197c70ef41f42bf393754bb9dd

Observation 85deecbd-cb63-473f-9204-68e68961d28d · outbound

This paper cites PersonaTalk: Bring Attention to Your Persona in Visual Dubbing.

Exploring Timeline Control for Facial Motion Generation PersonaTalk: Bring Attention to Your Persona in Visual Dubbing

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:54.502560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:54.116546Z digest=sha256:106ecf8c972cb4955f6f71b5a9f46e21e226aec590989545cf14773c9b5f4902

Observation f548c61a-4c08-4251-888e-d79c7a6c3c74 · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation.

Exploring Timeline Control for Facial Motion Generation Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:56.504965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:54.196437Z digest=sha256:ca541bcd8784200fb2fce376c7cc886496569be36bafa10d1ea386fb634cc110

Observation a075f5fe-8e44-415f-b1c2-12a5ab345e62 · outbound

This paper cites Media2face: Co-speech facial animation gen- eration with multi-modality guidance.

Exploring Timeline Control for Facial Motion Generation Media2face: Co-speech facial animation gen- eration with multi-modality guidance

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:56.242846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:54.279621Z digest=sha256:bd8a0c0369e4c2f9e49d7c5dffd3b7d67b21d9bd150b3ec80d7e0da103c4514c

Pith citing papers

No inbound Pith citation observations are available.