Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:57:41.937067Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2501.03332.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:57:41.937067Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cefb8f7e-1b1c-4885-895d-066ea60b3fd7 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Multimodal personality recognition using cross-attention transformer and behaviour encoding
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9a3d884a-bb0a-4e77-989b-edcb82da6259 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Multimodal vision transformers with forced attention for behavior analysis
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1eaf3175-b942-4fb1-8eb4-6ab022416907 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Text
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c420e10-9362-479c-a3a8-236326b87c4a · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Vivit: A video vi- sion transformer
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e79fad3-0e5a-4983-a89b-43454524ad2d · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Bodily be- haviors in social interaction: Novel annotations and state-of- the-art evaluation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation be8de12e-1c86-499f-b53b-51028f9afdec · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets AdaptFormer: Adapting Vision Transformers for Scalable Visual Recognition
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29b47da0-5ead-46ef-98c8-e46eb20e0c02 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Rescaling Egocentric Vision
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d7d4960-2027-4a56-9d79-a5fc39f7bf2c · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets A transformer-based joint-encoding for emotion recognition and sentiment analysis
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 517d6aa6-979a-4842-a0e5-e3e409206b80 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets PPT: Pre-trained Prompt Tuning for Few-shot Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57d12086-2d8c-486e-96e5-af7792988656 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Towards a Unified View of Parameter-Efficient Transfer Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17187cc0-9144-4617-b319-154d44b933fc · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Parameter-efficient trans- fer learning for NLP
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d50bb71c-9b26-4cca-9ca9-3d8dff2a5375 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Lora: Low-rank adaptation of large language models, 2021
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5f3ad6b7-82a6-4838-98de-3a7c28a17858 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets LoRA: Low-Rank Adaptation of Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47638112-c3bb-4db1-b42e-0a4f6c7fee1f · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Mumu: Cooperative mul- titask learning-based guided multimodal fusion
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1e937211-8683-4eed-a255-fd36306eb630 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Vi- sual prompt tuning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 062b865c-3421-4328-b567-87405a285e13 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Compacter: Efficient low-rank hypercomplex adapter layers
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 37329977-4498-4e23-aac8-b181ca94c9e5 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Parameter-efficient multi-task fine-tuning for transformers via shared hypernetworks
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 039df51b-eee5-410e-b860-2f411990162c · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Transformers in vision: A survey
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecc15948-d0c0-46eb-b0a4-d2f7e61e1c0c · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Gated Mechanism for Attention Based Multimodal Sentiment Analysis
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f7680172-9a00-48f8-b83c-08944022ab9a · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets The power of scale for parameter-efficient prompt tuning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a2452fab-058f-49ef-a192-af15732c1147 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Prefix-Tuning: Optimizing Continuous Prompts for Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ec927f3-ed76-41bd-aca2-7d3101e3d2f0 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Video swin transformer
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c4eee4e8-5c4c-4682-9226-73692234e2f5 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Parameter-efficient Multi-task Fine-tuning for Transformers via Shared Hypernetworks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f125b378-f6e3-48ad-9f6f-80238fe6635f · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets UniPELT: A Unified Framework for Parameter-Efficient Language Model Tuning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 791ab3e0-3809-4a6c-a617-1c3875df456e · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Tiny adapters for vision transformers, 2023
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3ab3513c-5b1e-444a-84ab-a79ca64d5114 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Context- aware personality inference in dyadic scenarios: Introducing the udiva dataset
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f5f113a2-2ebc-4f1c-9de3-62d70bf10728 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets ST-Adapter: Parameter-Efficient Image-to-Video Transfer Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70d7c14b-1d06-43ed-a5ee-6204b3fdae4c · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Dual-path adaptation from image to video transformers
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c53a2f50-64df-4e3c-9c6f-09538f2edcfb · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Learning transferable visual models from natural language supervision
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7dcef711-74ce-4ef1-a1d0-c7859b704e57 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Learning multiple visual domains with residual adapters.Ad- vances in neural information processing systems , 30, 2017
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b3cddac2-21f3-4624-99d5-391385aecde6 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Efficient parametrization of multi-domain deep neural net- works
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 05eb2b85-d972-4d7e-bba8-745c2bd5e3f7 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Multilingual Detection of Check-Worthy Claims using World Languages and Adapter Fusion
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 332d9f91-1186-490b-a9da-593c90a511ec · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 811a7111-1a81-440f-b2b4-57d8cfb7adc8 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation df61c7e2-bb0c-4dfa-929a-4697adc19e80 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Training neu- ral networks with fixed sparse masks
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1051fd7a-9059-410d-b500-fc5f29f641da · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre-training
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e61394d-7238-44fd-9e1c-9f45e904f98e · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9217669-4064-436c-9976-7f7f8fe6acda · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets M&M Mix: A Multimodal Multiview Transformer Ensemble
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31ca9c65-cca4-4d16-a0fd-a23491cd5239 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Multiview transformers for video recognition
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 80143f1f-e6e2-45f7-9206-914917274a3d · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Bitfit: Simple parameter-efficient fine-tuning for transformer-based masked language-models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fab0796f-1759-49e8-a8e4-9bd83f972ab1 · outbound
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
No inbound Pith citation observations are available.