Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:10:40.870961Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2507.03531.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:10:40.870961Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3f64a285-5486-4481-a848-41d37450c95b · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Learning transferable visual models from natural language supervision
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08f5321e-940a-492b-930e-a8d2926d4753 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Cowen, Stefanos Zafeiriou, Irene Kotsia, Eric Granger, Marco Pedersoli, Simon L
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a08ef3e8-8b7d-4929-bdad-2564cdb6b6d9 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Advancements in affective and behavior analysis: The 8th abaw workshop and competition
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1e90cbc-c112-4c54-b824-e6368022d9fd · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding 7th ABAW Competition: Multi-Task Learning and Compound Expression Recognition
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35fda226-7fbe-475b-bd42-1b768c83a350 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding The 6th affective behavior analysis in-the-wild (abaw) competition
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7445a316-cce0-437c-9b9a-b346f8832216 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Distribution matching for multi-task learning of classification tasks: A large-scale study on faces & beyond
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f58196a3-b9e7-48ef-aabb-b4550e736c0a · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Abaw: Valence-arousal estimation, expression recognition, action unit detection & emotional reaction intensity estimation challenges
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03e8bd93-488b-4c2c-aa4e-c3e8d09a1303 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Multi-label compound expression recognition: C-expr database & network
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 072d26f9-c235-417b-97cc-5b06ba8674f6 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Abaw: Valence-arousal estimation, expression recognition, action unit detection & emotional reaction intensity estimation challenges
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f63ccf43-2d0e-4eeb-829b-c0f7dcb30879 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Abaw: Valence-arousal estimation, expression recognition, action unit detection & multi-task learning challenges
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a085e703-44d0-4474-9ae1-31069f6f3644 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Analysing affective behavior in the second abaw2 competition
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87117d46-f234-483e-a142-2a2469971925 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Analysing affective behavior in the first abaw 2020 competition
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1689b15b-0d71-49b2-a3cb-f3e06dc87cfa · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Distribution Matching for Heterogeneous Multi-Task Learning: a Large-scale Face Study
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b24b0fc-aa14-4630-9a09-26b0a906c5be · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Affect Analysis in-the-wild: Valence-Arousal, Expressions, Action Units and a Unified Framework
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6a9dfae-7c77-4cc3-9499-58c56f06e7a9 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0045ef01-93c9-4bd8-9c95-7beb30d01b9c · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Face Behavior a la carte: Expressions, Affect and Action Units in a Single Network
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d5046d1-8763-454e-828a-e274d0327289 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2df671bc-8334-47fb-89ab-627009e25e1e · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Dvd: A comprehensive dataset for advancing violence detection in real-world scenarios
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 736a5e90-dd88-4472-88b7-674ad358d992 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Deep residual learning for image recognition
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7691a0c4-a1e8-4862-9972-962336ec3e0e · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Efficientnetv2: Smaller models and faster training
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fd62586-fcd0-458d-a8b2-606daee5c610 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Masked autoencoders are scalable vision learners
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0f5d483-74e6-4ac4-8d94-069b27293e7c · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Cnn architectures for large-scale audio classification
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 635b295a-8207-4728-a989-d738270442ec · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding wav2vec 2.0: A framework for self- supervised learning of speech representations
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bbd5d9cb-16bf-45c9-aef4-bd7838eb0c62 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Bert: Pre-training of deep bidirectional transformers for language understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6bb8e1d-d169-4914-941a-52d1d0f315db · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Attention is all you need
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38f6ce61-6586-45ce-9f5f-36e4e8be9a1c · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Learning phrase representations using rnn encoder-decoder for statistical machine translation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fae51b66-8d97-4a4e-a4c1-ff63160464e0 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding An Empirical Evaluation of Generic Convolutional and Recurrent Networks for Sequence Modeling
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ca3f941-0204-4fcd-8856-70ab839f6d5b · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Contrastive Training of Complex-Valued Autoencoders for Object Discovery
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2935fdb-3a17-4d9c-8fe2-1c96ee5ed57b · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8410bae-100f-4ba2-b17a-20ffe6b0573e · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Focal loss for dense object detection
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbdc3a54-adea-4007-a5d8-8526b33355a3 · outbound
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding Decoupled weight decay regularization
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.