Pith. sign in

Paper Citation Record · LEDGER

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings

As of 9 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2506.18055.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.18055 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:28:04.024662Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved4
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3d519f9d-1d03-4f53-ae21-48a237127be6 · outbound

This paper cites Active Speakers in Context,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Active Speakers in Context,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:11.966029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.161274Z digest=sha256:ad1730c8350f5633fd9bd0747657a4356954003fff817f14c3877d23760fdd38

Observation 595f5d18-6525-4a29-b1ee-497625d96228 · outbound

This paper cites Ava Active Speaker: An Audio-Visual Dataset for Active Speaker De- tection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Ava Active Speaker: An Audio-Visual Dataset for Active Speaker De- tection,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:11.526572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.211659Z digest=sha256:3d1203042ed04fb77fd16f01b2447cf8c8660f718ea2d46a488c1151db5555d4

Observation 675e0db4-f8b4-45d3-8e6a-fde1299fe0ce · outbound

This paper cites ASD-Transformer: Efficient Active Speaker Detection Using Self And Multimodal Transformers,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings ASD-Transformer: Efficient Active Speaker Detection Using Self And Multimodal Transformers,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:11.244934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.334739Z digest=sha256:9f49432b5224dac922c98a35112fd8fbe4b8d09c4c3badfa81085065e1e4ddb7

Observation 01600124-09ba-4662-8c9c-f10ea044203a · outbound

This paper cites Hello! My name is... Buffy.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Hello! My name is... Buffy

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:10.993651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.395980Z digest=sha256:657c8a195d5a275aa3079981b4e07379092a960ec348ea79d6c7e17b9b2d1721

Observation 1d152d5d-2af3-4d3b-b54c-3a4bd1ae02d4 · outbound

This paper cites Is Someone Speaking? Exploring Long-term Temporal Features for Audio-visual Active Speaker Detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Is Someone Speaking? Exploring Long-term Temporal Features for Audio-visual Active Speaker Detection,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:10.769025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.502763Z digest=sha256:6b2835f2d4d5efa297716b530a9519451b621d24b2c4b9df903b7856edd2c77f

Observation c9c5a4f2-537b-4f30-82e8-5af3a23ccf49 · outbound

This paper cites A Light Weight Model for Active Speaker Detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings A Light Weight Model for Active Speaker Detection,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:10.599063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.578729Z digest=sha256:a2c90eeb4792dc6bb05c29d1f86b45ceed2d496268d5d7daad21a340d6a17d98

Observation 0a1d72ee-4f06-4b28-84fb-db88ad03bbe5 · outbound

This paper cites How to Design a Three-Stage Architecture for Audio-Visual Active Speaker Detection in the Wild,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings How to Design a Three-Stage Architecture for Audio-Visual Active Speaker Detection in the Wild,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:10.352982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.661110Z digest=sha256:1b8bb13ae1e087861e4615c6084c71f817c4066560edea61028b97eef472841f

Observation 560f2392-93ce-4dd9-8852-7e5e5e6e0f87 · outbound

This paper cites Rethinking audio- visual synchronization for active speaker detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Rethinking audio- visual synchronization for active speaker detection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:10.001902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.743867Z digest=sha256:fd0c02fe21ed90176256c582d5909da2404fa9d6b35cb960eef690f8e15ced77

Observation a6840b8f-f578-4718-8d99-81eb95934fba · outbound

This paper cites Look who’s talking: speaker detection using video and audio correlation,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Look who’s talking: speaker detection using video and audio correlation,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:09.792847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.830166Z digest=sha256:e13abbaec3550249a47c5f821a5b9608ef440125baf9cedc572aac80776b27d9

Observation c24fe963-85d5-4720-890b-e6b43059b1c9 · outbound

This paper cites End-to-End Active Speaker Detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings End-to-End Active Speaker Detection,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:09.581114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.907466Z digest=sha256:48fdc7d0d5449ed409d809996753c22de346d5f89cbb10eb86e9b03f0d3f7c08

Observation fdd7f602-0b28-4fbd-becd-a953dd51785a · outbound

This paper cites Learning Long-Term Spatial-Temporal Graphs for Active Speaker Detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Learning Long-Term Spatial-Temporal Graphs for Active Speaker Detection,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:09.351206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:00.958946Z digest=sha256:8de6e66ffa9dfb19614156ea9a384687f9aa5141532ae6717b1106303ddb89f3

Observation 485f6188-833d-4be2-ac95-f5721ce9ff00 · outbound

This paper cites MAAS: Multi-modal Assignation for Active Speaker Detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings MAAS: Multi-modal Assignation for Active Speaker Detection,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:09.128276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.097316Z digest=sha256:edeab3a7e33cb6f80113e236bb5d38788010d76519647be17b887910afd7231b

Observation 39561c4c-958e-44df-873b-cdad5357f4d1 · outbound

This paper cites Improving Audiovisual Active Speaker Detection in Egocentric Recordings with the Data-Efficient Image Transformer,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Improving Audiovisual Active Speaker Detection in Egocentric Recordings with the Data-Efficient Image Transformer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:08.895941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.251608Z digest=sha256:1858093cb89d8f45ff5f3953c8ad8c33819e34632eab500317adf8a8aa83882e

Observation 7c1b49ea-fa35-443d-b6f4-ad6f9019017b · outbound

This paper cites Target Active Speaker Detection with Audio-visual Cues,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Target Active Speaker Detection with Audio-visual Cues,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:08.743478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.401263Z digest=sha256:b87ff4ac1b526b41c4924e0d8f1b1a563dbd7d25db37b220f06dbe148548c5ce

Observation 6934f742-1c5e-4600-a4e2-622a37b09024 · outbound

This paper cites Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Speaker Embedding Informed Audiovisual Active Speaker Detection for Egocentric Recordings,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:08.516435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.510834Z digest=sha256:ab95b6d114585dac5f1915122d41c844caa71453aeb826e9f108b7ecfe22210e

Observation 6b46b576-648a-42dc-9e1b-bc703fae9e9f · outbound

This paper cites Ego4D: Around the World in 3,000 Hours of Egocentric Video,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Ego4D: Around the World in 3,000 Hours of Egocentric Video,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:08.272269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.613046Z digest=sha256:2c8f4241b52e64da84316107c4da3754a02639b424a407e41eebf3196714ad6d

Observation 3ff72ce4-a91e-44ca-8841-3d3cca488ac6 · outbound

This paper cites Technical Report for Ego4D Long Term Action Anticipation Challenge 2023.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Technical Report for Ego4D Long Term Action Anticipation Challenge 2023

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:01.668716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:01.668716Z digest=sha256:8473834c8cb1312be780f9c4fb7101ed08303cadcc5642317d0918cc8d982794

Observation 3c85c3c6-8f4e-470d-8fcb-67259ac708cb · outbound

This paper cites Seeking the shape of sound: An adaptive framework for learning voice-face association,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Seeking the shape of sound: An adaptive framework for learning voice-face association,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:08.058693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.716049Z digest=sha256:322701e4b1d31c0091d9e7915a0da24a06d2394bb26ced9cc287c4bae7da92c1

Observation 03f7becb-6ffe-425f-8532-08d88f4dccff · outbound

This paper cites Self-lifting: A novel framework for unsupervised voice-face association learning,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Self-lifting: A novel framework for unsupervised voice-face association learning,

Reference 19

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T23:28:04.554745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:01.837550Z digest=sha256:26e92a73647a8abcd05009e76e58bc4837ddd6ea9bf8c9c1300bd60f8f211673

Observation c818b741-5277-4e16-abd0-a8d1291c9357 · outbound

This paper cites Learnable pins: Cross- modal embeddings for person identity,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Learnable pins: Cross- modal embeddings for person identity,

Reference 20

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T23:28:07.868774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.082727Z digest=sha256:e373038c3c744db8bdefd59f27bb822e6ec49512762764259026bd2e1074c2d7

Observation 6877f71c-2b62-4c0b-ac45-67030f6e342c · outbound

This paper cites Single-branch network for multimodal training,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Single-branch network for multimodal training,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:07.663855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.239663Z digest=sha256:386d47afd1cc10037bcd83856970c4a90324f06836e66e6809ae7d8c4d5b2d9d

Observation 45f549bc-58ed-4e08-901d-8064243fca8f · outbound

This paper cites Disentangled representation learning for cross-modal biometric matching,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Disentangled representation learning for cross-modal biometric matching,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:07.467614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.377727Z digest=sha256:775a569fdf077133c15a0a91da1f0555a2ad0dcf1c35a89949c4df15e79c97a3

Observation a7cf0fa6-e9bc-4d58-81ee-4a56267393e8 · outbound

This paper cites An efficient momentum framework for face- voice association learning,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings An efficient momentum framework for face- voice association learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:07.247169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.488724Z digest=sha256:2e3c63cade2259b3e52cb78ce282a3102473de4520d33a33b44722dac5bffa54

Observation a5379e78-54f8-44f9-b109-a007618e45a8 · outbound

This paper cites Audio-visual activity guided cross-modal identity association for active speaker detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Audio-visual activity guided cross-modal identity association for active speaker detection,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:07.078961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.572698Z digest=sha256:a737dc459b812ce484b03e2377f8b01e7e5d878c744255a69c39639b251f8ce6

Observation 4b0cb8aa-1595-430f-9381-c9c9d544d02c · outbound

This paper cites Favoa: Face-voice association favours ambiguous speaker detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Favoa: Face-voice association favours ambiguous speaker detection,

Reference 25

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T23:28:06.912665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.704753Z digest=sha256:d457c3cda87359922cb33c0c2678904423ed013010525003636a14ae86c7c279

Observation 94448a55-3457-4788-9073-46b186ac08be · outbound

This paper cites What in the world do we hear?: An ecological approach to auditory event perception,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings What in the world do we hear?: An ecological approach to auditory event perception,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:06.692666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.819345Z digest=sha256:b83b5fd5055bdeabb56da6251464c2072d313eab08464fe9f0d7291012481dd3

Observation 678a8f40-55c7-4a93-949c-da8d2397e8fe · outbound

This paper cites The development of infant learning about specific face–voice relations,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings The development of infant learning about specific face–voice relations,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:06.518843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:02.931958Z digest=sha256:b08265facbdadec9d68a72656cdc40e0750cce6f219df21c1d0de1aaeb324a37

Observation bfe23d5c-abdf-4b35-9598-368caf1b438d · outbound

This paper cites Powerset multi-class cross entropy loss for neural speaker diarization,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Powerset multi-class cross entropy loss for neural speaker diarization,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.064751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:03.064751Z digest=sha256:9e6e35fc1bf993557933fb844165752b99d71e64afa1a24e6b2afb0b6e4c9a8c

Observation 49998520-39b3-4c4d-980d-52b78574f43a · outbound

This paper cites Going deeper with convolutions,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Going deeper with convolutions,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:06.213326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:03.181816Z digest=sha256:01bce128d97b242f5cb981cab9763a3f1cb14307ab906286b6c3446890275922

Observation 84a012c6-bd7f-4fe3-a822-1043f0ddb7c7 · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.287092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:03.287092Z digest=sha256:d6a69c235b7739a88dd0acabbf11f4f04129130c1d47e7afe439e2b7952dbf7d

Observation 372c41aa-29c2-48f0-ba07-06083b3c28ce · outbound

This paper cites V oxCeleb: A Large-Scale Speaker Identification Dataset,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings V oxCeleb: A Large-Scale Speaker Identification Dataset,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.413480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:03.413480Z digest=sha256:c6d175d0581849a95d7e99b668cf1af5e1110e26d6bc1913efb86110ca1d8c15

Observation bd9368f2-295d-43ba-94b1-9dc87ed33b03 · outbound

This paper cites Silero V AD: pre-trained enterprise-grade V oice Activity Detector (V AD), Number Detector and Language Classifier,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Silero V AD: pre-trained enterprise-grade V oice Activity Detector (V AD), Number Detector and Language Classifier,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:06.010040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:03.514567Z digest=sha256:27cfc61949b717cb530fd0684e6e124fd566db54d99995f145cf34f437426882

Observation 1c82e53b-dd44-43cb-8d11-e1060c2a9caf · outbound

This paper cites Speaker change detection using support vector machines,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Speaker change detection using support vector machines,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:05.791533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:03.617037Z digest=sha256:6f3665d5e57bc4646fed9d05650b9016bab3593451c563b8cf3a9945472a530b

Observation 01524bdb-3042-4b5f-98d8-9487ec8a63ab · outbound

This paper cites Robust Object Recogni- tion Through Symbiotic Deep Learning In Mobile Robots,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings Robust Object Recogni- tion Through Symbiotic Deep Learning In Mobile Robots,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:05.513645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:03.722890Z digest=sha256:3304c3cdfbb6e31134db75d04fe5906c1b074d698288bc25c3f85aac7dd39657

Observation d9dcd20c-c5f4-481e-bd08-0a81574c4e69 · outbound

This paper cites The PASCAL Visual Object Classes Challenge 2012 (VOC2012) Results,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings The PASCAL Visual Object Classes Challenge 2012 (VOC2012) Results,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:05.334738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:03.826961Z digest=sha256:3924f90abf61037643d71eb744b8efa2bb6cbd8eed25be9c38865d91fc0c4dc4

Observation dc7409f3-f4be-4f02-b059-3ca6b14fc596 · outbound

This paper cites LoCoNet: Long-Short Context Network for Active Speaker Detection,.

Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings LoCoNet: Long-Short Context Network for Active Speaker Detection,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:28:04.953889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:28:04.024662Z digest=sha256:55a696258abd690919ae7fc6f033361fc14fbde0cce21e461478fbb6549a2aba

Pith citing papers

No inbound Pith citation observations are available.