Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:40:21.868428Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2604.02941.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:40:21.868428Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 955f449a-a4ab-47fd-8c0c-725c6ac113e1 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Deep speech 2: End- to-end speech recognition in english and mandarin
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5039727-dee9-4396-96f8-a3f265e42d16 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Uncertainty-Aware Weakly Supervised Action De- tection from Untrimmed Videos
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e01c389-bd6f-4579-aef5-0093b4edeb58 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion wav2vec 2.0: A framework for self-supervised learning of speech representations
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 542add47-c248-483b-8dcd-0798dae528b8 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion SCM: Spatial Continuity Modeling for Weakly Supervised Object Localization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a671660b-9c74-483c-9303-104917d93623 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 159b2e5a-75d1-4a28-afd1-7ad38f70741e · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd4c7419-7755-4b66-9514-5b9d2eeae97e · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion A flexible model for training action local- ization with varying levels of supervision
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e81054d-ea57-41b4-9b22-fb78f65c1543 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Evaluating weakly supervised object localization methods right
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8d60310-b36a-4b8b-b08c-d2c1e2035e82 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Bert: Pre-training of deep bidirectional trans- formers for language understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0861cf4-fe86-4490-a1ac-e8aab2d95a36 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Gonzalez, and Trevor Darrell
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ad1b318-e4f1-4005-909a-8f131ebdbea5 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Scaling laws of synthetic images for model training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed9ec79a-95ea-43e6-b80a-5d47c6ec379b · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Attention branch network: Learning of attention mechanism for visual explanation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f70fb42f-15b7-4a8a-b19b-1554a91a45f5 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion TS-CAM: To- ken Semantic Coupled Attention Map for Weakly Supervised Object Localization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8e0ffe8-6c8a-474c-a53e-80e0175679d2 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Wichmann
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95e694c1-68d2-4659-b267-b19247193194 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unified Keypoint-Based Action Recognition Framework via Struc- tured Keypoint Pooling
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fb9318d-44e4-49ba-aa79-fe378ad54d02 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Deep Residual Learning for Image Recognition
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8e752a5-ebbe-4f15-baee-6108714c4c63 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion On the Unreasonable Effectiveness of Last- Layer Retraining
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d362c31-c3b3-43d4-bb75-78c30b6298af · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Mitigating Simplicity Bias in Neural Net- works: A Feature Sieve Modification, Regularization, and Self-Supervised Augmentation Approach
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a27dcd5c-169a-477b-9759-07345b3552c5 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Puz- zle mix: Exploiting saliency and local statistics for optimal mixup
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 694beee3-7df1-456e-b921-ba2879d6c81e · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Wilds: A benchmark of in-the-wild distribution shifts
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8a3ae1a-eb1d-4416-8438-dd39cfcbfacf · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 315904c5-5989-44fb-bf93-11b5e65c4376 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion A threshold selection method from gray- level histograms.IEEE Transactions on Systems, Man, and Cybernetics, 9(1):62–66, 1979
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28cfd5a3-9b78-45ca-8c36-460b28f9593e · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion ResizeMix: Mixing Data with Preserved Object Information and True Labels
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27f8395a-ba30-4e6e-8581-10e96286a9e7 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion You Only Look Once: Unified, Real-Time Object Detection
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28c7b39c-04b8-46c3-845b-cb479b3a146d · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 583d00d8-bf38-4483-ac7d-e99cf51fcf88 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Why Should I Trust You?
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eddffb0f-c213-4a8e-bfd5-ffd8eb135976 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Richter, Vibhav Vineet, Stefan Roth, and Vladlen Koltun
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 185b8d5d-bfbb-4f51-b00e-f75be0c89dae · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion High-resolution image syn- thesis with latent diffusion models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25e6075b-5b74-49a0-8ff6-aee61d96adfa · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Hughes, and Finale Doshi- Velez
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 808261c6-0132-49bd-9160-ead7de148c9c · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Berg, and Li Fei-Fei
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 761c3ec0-8040-43b6-be89-45800a370b04 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Distributionally Robust Neural Networks for Group Shifts: On the Importance of Regularization for Worst-Case Generalization
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37bd9596-2a70-4c96-b466-2e575d708ee2 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Fake it till you make it: Learning trans- ferable representations from synthetic imagenet clones
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 019481a8-db7a-44d1-80c6-94858853ca4e · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Grad-CAM: visual explanations from deep networks via gradient-based localization.IJCV, 2020
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ce575b6-a68d-470e-abaf-1fa315b44328 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Counterfactual Co-occurring Learning for Bias Mitigation in Weakly-supervised Object Localiza- tion.IEEE Transactions on Multimedia, 2026
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5823782-5fa3-45d6-b9dc-36cf1c35eefd · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Very Deep Con- volutional Networks for Large-Scale Image Recognition
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c744f365-8a81-47cb-83ef-8a5ce6438929 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Singh and Y .J
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3d8ebf4-2c01-4701-a86f-c746f982931a · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65d08dee-e502-43bb-b162-ef51f492a404 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Deep High-Resolution Representation Learning for Human Pose Estimation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5913bedc-8b28-4797-97a4-3fdd92123730 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7914bf73-9fd2-47b9-ade0-f32af419a22e · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Training Data-Efficient Image Transformers & Distillation through Attention
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e18546-9b10-4162-a30c-fdf902958c04 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Black, Ivan Laptev, and Cordelia Schmid
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd870079-2859-47bd-987a-c1829f99d15a · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion The Caltech-UCSD Birds-200- 2011 Dataset.Caltech Technical Report, 2011
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d795d151-f84f-4972-805d-a2c220b8c27d · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Spatial-Aware Token for Weakly Supervised Object Localization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 294f640f-8ba0-44f2-8803-85addc3044a2 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7ef1230-51f2-4107-be61-0286465dcfca · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion CutMix: Regular- ization Strategy to Train Strong Classifiers With Localizable Features
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1034c450-46e5-4a69-9e75-5ee9cdef4603 · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Dauphin, and David Lopez-Paz
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e24ef8d-11b1-4384-88bf-7e9dda6d887a · outbound
MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.