Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:09:18.680337Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2502.06012.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T17:09:18.680337Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c9307a56-36dc-43b4-add9-feb4284afeb0 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Is Someone Speaking? Exploring Long-term Temporal Features for Audio-visual Active Speaker Detection,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 47ddadcb-d087-4eb1-86d6-06434784b766 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Combining Residual Networks with LSTMs for Lipreading,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4b4a39e7-a127-4a86-8281-b656e5e37680 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Active Speakers in Context,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1136810-d3a3-402f-819d-de2d8f289c8e · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings ASD-Transformer: Efficient Active Speaker Detection Using Self And Multimodal Transformers,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68a840b7-088d-4d2b-8965-d29d9c87d898 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Hello! My name is... Buffy
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09db0d23-3d42-46be-82b3-4652a159bc36 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings MAAS: Multi-modal Assignation for Active Speaker Detection,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d87e82d8-2b51-441b-bd18-8c5834c0465f · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Target Active Speaker Detection with Audio-visual Cues,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f99e09c-4b40-4109-b829-75bf046371a9 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Ava Active Speaker: An Audio-Visual Dataset for Active Speaker De- tection,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac4f4dfd-872a-45dc-891f-bf0e2010181a · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Improving Audiovisual Active Speaker Detection in Egocentric Recordings with the Data-Efficient Image Transformer,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 21ef0601-6f5b-49da-a5f2-e9ad0d680dfc · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Ego4D: Around the World in 3,000 Hours of Egocentric Video,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36daf742-d0ff-45dd-9b53-f0257df901ad · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings End-to-End Active Speaker Detection, author=Juan Leon Alcazar and Moritz Cordes and Chen Zhao and Bernard Ghanem,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 04974023-5009-4843-8fbd-f24c2d8d0f4f · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings How to Design a Three-Stage Architecture for Audio-Visual Active Speaker Detection in the Wild,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb15f594-52b5-4eb8-a1a7-cfaff91a8646 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings A Light Weight Model for Active Speaker Detection,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 672bc6f7-3ec5-4460-969d-ccaff54edcc4 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings LoCoNet: Long-Short Context Network for Active Speaker Detection,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ba33c77-cfdf-4790-afb3-92d924c42c6a · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Learning Long-Term Spatial-Temporal Graphs for Active Speaker Detection,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7985d78f-7386-4bc4-bfe9-64a21de0d797 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9c9d624-29e9-44ae-9f97-53500061f0d5 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Personal VAD 2.0: Optimizing Personal Voice Activity Detection for On-Device Speech Recognition
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7eda9495-fd1b-4522-bbbf-c9191089d911 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Efficient Personal V oice Activity Detection with Wake Word Reference Speech,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9fe9b316-a7de-4906-a1bf-4b9883193f88 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings V oxCeleb: A Large-Scale Speaker Identification Dataset,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8792db4-8364-4ab7-9d03-35e7e6946e63 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Ring Loss: Convex Feature Normalization for Face Recognition,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c34fe6a-207f-4beb-a411-a14f2c7ce986 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings ArcFace: Additive Angular Margin Loss for Deep Face Recognition,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 786bcc82-f5bf-4bc3-acaa-a9ff2a5cfed2 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings SphereFace: Deep Hypersphere Embedding for Face Recognition,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e72199d4-0acd-4d42-85b9-72cdffc84cec · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings CosFace: Large Margin Cosine Loss for Deep Face Recognition,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 599cbf32-1274-4f73-a393-10805583039d · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Partial FC: Training 10 Million Identities on a Single Ma- chine,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72c51122-fd47-4229-8770-65d107bced6b · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Attention is All you Need,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 497d1872-2ede-461d-b85a-85a0398a7510 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Robust Object Recogni- tion Through Symbiotic Deep Learning In Mobile Robots,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a683ce62-4341-4bca-9986-5edeae446089 · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings The PASCAL Visual Object Classes Challenge 2012 (VOC2012) Results,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23a06e36-1d2c-49dd-80aa-aab1d90798dc · outbound
Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings Technical Report for Ego4D Long Term Action Anticipation Challenge 2023
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.