Pith. sign in

Paper Citation Record · LEDGER

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion

As of 11 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2604.02941.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.02941 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-13T13:40:21.868428Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 955f449a-a4ab-47fd-8c0c-725c6ac113e1 · outbound

This paper cites Deep speech 2: End- to-end speech recognition in english and mandarin.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Deep speech 2: End- to-end speech recognition in english and mandarin

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:790dbdd80012a1af323ff3c320dec58b891a7909eb1763fec663d0d64b32e7eb

Observation a5039727-dee9-4396-96f8-a3f265e42d16 · outbound

This paper cites Uncertainty-Aware Weakly Supervised Action De- tection from Untrimmed Videos.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Uncertainty-Aware Weakly Supervised Action De- tection from Untrimmed Videos

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:2670ea87d6041a61ebd2208926bd6cd26dac3e5207ad0424e0629efe43eb9186

Observation 2e01c389-bd6f-4579-aef5-0093b4edeb58 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion wav2vec 2.0: A framework for self-supervised learning of speech representations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:d8923f55dc99f5ce48bdd45d2e36f8001e67bc521d3ff6ef3208e7a0cb2e3f50

Observation 542add47-c248-483b-8dcd-0798dae528b8 · outbound

This paper cites SCM: Spatial Continuity Modeling for Weakly Supervised Object Localization.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion SCM: Spatial Continuity Modeling for Weakly Supervised Object Localization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:adf6ba5606ea13cff335cfc5e87eff979700f01bb9c6bef2711e46ebf0151fab

Observation a671660b-9c74-483c-9303-104917d93623 · outbound

This paper cites an unresolved cited work.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:f3fd01988a8dcaf5b313ddc3b7474955ca148b7f70138bf2eb3cd4b8d6fd8716

Observation 159b2e5a-75d1-4a28-afd1-7ad38f70741e · outbound

This paper cites Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:76a71822d6df7cea168a666510a4da5f0075eedc4de756a2e128d184a39a6742

Observation bd4c7419-7755-4b66-9514-5b9d2eeae97e · outbound

This paper cites A flexible model for training action local- ization with varying levels of supervision.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion A flexible model for training action local- ization with varying levels of supervision

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:626d6a205151138c56024cce03190fc3817c28fa4ca16e02201c8300e8e5740f

Observation 6e81054d-ea57-41b4-9b22-fb78f65c1543 · outbound

This paper cites Evaluating weakly supervised object localization methods right.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Evaluating weakly supervised object localization methods right

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:1d395315b34debe540df1cc7f19d05d0e67db3312a8ca976cb75b18eaec940af

Observation d8d60310-b36a-4b8b-b08c-d2c1e2035e82 · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Bert: Pre-training of deep bidirectional trans- formers for language understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:efd99dd9381df1587d348124ab3b58e9f013215e35db0b6eff1fd19d71a86a2c

Observation c0861cf4-fe86-4490-a1ac-e8aab2d95a36 · outbound

This paper cites Gonzalez, and Trevor Darrell.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Gonzalez, and Trevor Darrell

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:c6faf9aeb026a7c1e507123f0689879b1dd6226119d13c9b7a05379e89661c2d

Observation 5ad1b318-e4f1-4005-909a-8f131ebdbea5 · outbound

This paper cites Scaling laws of synthetic images for model training.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Scaling laws of synthetic images for model training

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:dc48d4eb884002f4bbf84f345e0c7b6050d6e82af921d6f2418b8b6d1787db0c

Observation ed9ec79a-95ea-43e6-b80a-5d47c6ec379b · outbound

This paper cites Attention branch network: Learning of attention mechanism for visual explanation.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Attention branch network: Learning of attention mechanism for visual explanation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:a22eb4d0aaba94f902bb23698c0e33f6c9de131668ab9fa8a4a3f1f930389891

Observation f70fb42f-15b7-4a8a-b19b-1554a91a45f5 · outbound

This paper cites TS-CAM: To- ken Semantic Coupled Attention Map for Weakly Supervised Object Localization.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion TS-CAM: To- ken Semantic Coupled Attention Map for Weakly Supervised Object Localization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:0ca21e8242130421790586504f21b35baca992b0de2636264babcf7ded37b3b1

Observation f8e0ffe8-6c8a-474c-a53e-80e0175679d2 · outbound

This paper cites Wichmann.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Wichmann

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:5e5ba3f016792423e8cfe9e1ce1cfa2c64385d468badb65de6679a85a61ad7a6

Observation 95e694c1-68d2-4659-b267-b19247193194 · outbound

This paper cites Unified Keypoint-Based Action Recognition Framework via Struc- tured Keypoint Pooling.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unified Keypoint-Based Action Recognition Framework via Struc- tured Keypoint Pooling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:725de612e888502746d72fe672141915514d837ddda1f21236b79391b6266afd

Observation 2fb9318d-44e4-49ba-aa79-fe378ad54d02 · outbound

This paper cites Deep Residual Learning for Image Recognition.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Deep Residual Learning for Image Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:3779b28b5e38392d583634a5df66ebcd65e6e4b89cbfc46797d4257020675c7a

Observation e8e752a5-ebbe-4f15-baee-6108714c4c63 · outbound

This paper cites On the Unreasonable Effectiveness of Last- Layer Retraining.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion On the Unreasonable Effectiveness of Last- Layer Retraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:f6a007e96ffb9ddb8f8cdd89abd2dcc606c39576c20030501f9a9a5e9e14c5ba

Observation 1d362c31-c3b3-43d4-bb75-78c30b6298af · outbound

This paper cites Mitigating Simplicity Bias in Neural Net- works: A Feature Sieve Modification, Regularization, and Self-Supervised Augmentation Approach.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Mitigating Simplicity Bias in Neural Net- works: A Feature Sieve Modification, Regularization, and Self-Supervised Augmentation Approach

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:f1a222d1b1ba2267a63efb5d3af417b4f74e67d1302753eadc31b0e357efedb2

Observation a27dcd5c-169a-477b-9759-07345b3552c5 · outbound

This paper cites Puz- zle mix: Exploiting saliency and local statistics for optimal mixup.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Puz- zle mix: Exploiting saliency and local statistics for optimal mixup

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:9148becaf8ff9371bfe7077ac2136489a9577884763884531366052103c0955a

Observation 694beee3-7df1-456e-b921-ba2879d6c81e · outbound

This paper cites Wilds: A benchmark of in-the-wild distribution shifts.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Wilds: A benchmark of in-the-wild distribution shifts

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:74fbbc18354984c883307b4b166316ef0812b5e9743ea04703771e7f29313eab

Observation c8a3ae1a-eb1d-4416-8438-dd39cfcbfacf · outbound

This paper cites an unresolved cited work.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:7b4fef249f236d954135ddc9333b19d9231b5ca82c2ece90fa1a28853a3c4308

Observation 315904c5-5989-44fb-bf93-11b5e65c4376 · outbound

This paper cites A threshold selection method from gray- level histograms.IEEE Transactions on Systems, Man, and Cybernetics, 9(1):62–66, 1979.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion A threshold selection method from gray- level histograms.IEEE Transactions on Systems, Man, and Cybernetics, 9(1):62–66, 1979

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:7f80f76e77219bc97dfca8b7d4d528907c21c62582a4206ffb48b0b0f8c78d21

Observation 28cfd5a3-9b78-45ca-8c36-460b28f9593e · outbound

This paper cites ResizeMix: Mixing Data with Preserved Object Information and True Labels.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion ResizeMix: Mixing Data with Preserved Object Information and True Labels

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:137469da5e33e4af0865c0495be02d48fb71401f3b6984404c8e931189edbe06

Observation 27f8395a-ba30-4e6e-8581-10e96286a9e7 · outbound

This paper cites You Only Look Once: Unified, Real-Time Object Detection.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion You Only Look Once: Unified, Real-Time Object Detection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:7b42b20e400fe9f769cda470d5e1221a9b6dc456d269ad3b14b9ec7219668d1d

Observation 28c7b39c-04b8-46c3-845b-cb479b3a146d · outbound

This paper cites Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:0355c35543d0aba8cb43fff25c7b79cacbceba58d7a0c304df3caf16b018a6bd

Observation 583d00d8-bf38-4483-ac7d-e99cf51fcf88 · outbound

This paper cites Why Should I Trust You?.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Why Should I Trust You?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:22514832b2fbfd9f031f5aed0538dba803cd0705afca98b8872a2c9473f94d24

Observation eddffb0f-c213-4a8e-bfd5-ffd8eb135976 · outbound

This paper cites Richter, Vibhav Vineet, Stefan Roth, and Vladlen Koltun.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Richter, Vibhav Vineet, Stefan Roth, and Vladlen Koltun

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:d420d089808f1fe45b758d8d4e6c98faa3883f4a410b6771c89b57c2d5db44e6

Observation 185b8d5d-bfbb-4f51-b00e-f75be0c89dae · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion High-resolution image syn- thesis with latent diffusion models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:7eabb3246b51aaf2394488af35383bd4371b6a095cffba036a0bec3a1ea184d0

Observation 25e6075b-5b74-49a0-8ff6-aee61d96adfa · outbound

This paper cites Hughes, and Finale Doshi- Velez.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Hughes, and Finale Doshi- Velez

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:85a40ffc373963f6077cdbc8dfa5b9475359a1d579b7b6e780c75d516875d4ff

Observation 808261c6-0132-49bd-9160-ead7de148c9c · outbound

This paper cites Berg, and Li Fei-Fei.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Berg, and Li Fei-Fei

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:162a1a7ef3b140ba85f4e1cf4a3e0c684c478150ea831cb064cae762d126ac2b

Observation 761c3ec0-8040-43b6-be89-45800a370b04 · outbound

This paper cites Distributionally Robust Neural Networks for Group Shifts: On the Importance of Regularization for Worst-Case Generalization.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Distributionally Robust Neural Networks for Group Shifts: On the Importance of Regularization for Worst-Case Generalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:30de43d90bdc895c6ed05405923be1f017a93c6bedacf0825103f55c6823b76f

Observation 37bd9596-2a70-4c96-b466-2e575d708ee2 · outbound

This paper cites Fake it till you make it: Learning trans- ferable representations from synthetic imagenet clones.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Fake it till you make it: Learning trans- ferable representations from synthetic imagenet clones

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:2872408d6f341c713727c9eee0017f46db32844d6f0f4fd39464fdc9737a13a6

Observation 019481a8-db7a-44d1-80c6-94858853ca4e · outbound

This paper cites Grad-CAM: visual explanations from deep networks via gradient-based localization.IJCV, 2020.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Grad-CAM: visual explanations from deep networks via gradient-based localization.IJCV, 2020

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:bae098bbf93f35a3816b33233c1c6f7c83caf7a5047fc4e8befbb81b7e5c9db6

Observation 4ce575b6-a68d-470e-abaf-1fa315b44328 · outbound

This paper cites Counterfactual Co-occurring Learning for Bias Mitigation in Weakly-supervised Object Localiza- tion.IEEE Transactions on Multimedia, 2026.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Counterfactual Co-occurring Learning for Bias Mitigation in Weakly-supervised Object Localiza- tion.IEEE Transactions on Multimedia, 2026

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:611428a08593ddc96af41bc1b9f3146a12a9024e91839570a2cc5e920acbda52

Observation e5823782-5fa3-45d6-b9dc-36cf1c35eefd · outbound

This paper cites Very Deep Con- volutional Networks for Large-Scale Image Recognition.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Very Deep Con- volutional Networks for Large-Scale Image Recognition

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:a0cc0630f570e53d6dc59f4637f1a1374d84876d27f0d74bb2684fbed141f769

Observation c744f365-8a81-47cb-83ef-8a5ce6438929 · outbound

This paper cites Singh and Y .J.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Singh and Y .J

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:3fa1f18e7db9246b8d0b1f97c279fcb86748604668f2d7761703aec562ba77ca

Observation b3d8ebf4-2c01-4701-a86f-c746f982931a · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:e38738abe490fa84c0187a74d8228a9599273eec30e7839afc1d214d90cd74f2

Observation 65d08dee-e502-43bb-b162-ef51f492a404 · outbound

This paper cites Deep High-Resolution Representation Learning for Human Pose Estimation.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Deep High-Resolution Representation Learning for Human Pose Estimation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:2f508a714e4ce2edc4460961628445477ba84d1965cad17d24f3ac52d485d375

Observation 5913bedc-8b28-4797-97a4-3fdd92123730 · outbound

This paper cites Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:c29a843245f8790af17da9f33751cad037ab027b31c3af455a664129b45ac72c

Observation 7914bf73-9fd2-47b9-ade0-f32af419a22e · outbound

This paper cites Training Data-Efficient Image Transformers & Distillation through Attention.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Training Data-Efficient Image Transformers & Distillation through Attention

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:4be7c452e9805b1e2b2fbbb24709230d6ade4410aa6625ca7047015d6f4739d7

Observation f6e18546-9b10-4162-a30c-fdf902958c04 · outbound

This paper cites Black, Ivan Laptev, and Cordelia Schmid.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Black, Ivan Laptev, and Cordelia Schmid

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:485d1d6fba50711916160cf82aabf7a5947796ecfd17a064cf9ae61a576cde92

Observation bd870079-2859-47bd-987a-c1829f99d15a · outbound

This paper cites The Caltech-UCSD Birds-200- 2011 Dataset.Caltech Technical Report, 2011.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion The Caltech-UCSD Birds-200- 2011 Dataset.Caltech Technical Report, 2011

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:6bd3d6d393aeecb701fbd646bfd14f290261598974055b06f3536c939c6c65cf

Observation d795d151-f84f-4972-805d-a2c220b8c27d · outbound

This paper cites Spatial-Aware Token for Weakly Supervised Object Localization.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Spatial-Aware Token for Weakly Supervised Object Localization

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:4aba4592ef259c7a5aa4ddad3381c354091581fc9426ea75b326a92b28eaa81d

Observation 294f640f-8ba0-44f2-8803-85addc3044a2 · outbound

This paper cites Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:7f30723653251b9dd48cf101480c298b6ccd84f7e4f7999c359d0b5721110d7b

Observation e7ef1230-51f2-4107-be61-0286465dcfca · outbound

This paper cites CutMix: Regular- ization Strategy to Train Strong Classifiers With Localizable Features.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion CutMix: Regular- ization Strategy to Train Strong Classifiers With Localizable Features

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:28969395d8e398f632dc020e62747953491f2bad66c1195dc663dd49f4ebabc5

Observation 1034c450-46e5-4a69-9e75-5ee9cdef4603 · outbound

This paper cites Dauphin, and David Lopez-Paz.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Dauphin, and David Lopez-Paz

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:817ac4cdb1918778535339f3ceb5338e2bfa8e9c22f146491069e24ea4f30d3d

Observation 9e24ef8d-11b1-4384-88bf-7e9dda6d887a · outbound

This paper cites an unresolved cited work.

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-13T13:40:21.868428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:40:21.868428Z digest=sha256:94297e3fbcabdfbedc41504c4d0ae7e47235506acadaf2e2431293f338b9d270

Pith citing papers

No inbound Pith citation observations are available.