Pith. sign in

Paper Citation Record · LEDGER

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence

As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2608.05816.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05816 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:04:01.886110Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6c274851-dcd5-4301-ac50-615b7b1006bc · outbound

This paper cites In: Proceedings of the IEEE international conference on computer vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE international conference on computer vision

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.603025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.554204Z digest=sha256:49507ab00de404dca3ad55b7e0ed5f4a96c4f1e6c2197d7f8601c31b998d1da2

Observation 1860f9dc-9b7b-4999-8dcb-9d2fc6394f17 · outbound

This paper cites In: Proceedings of the Euro- pean conference on computer vision (ECCV).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the Euro- pean conference on computer vision (ECCV)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.588893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.559635Z digest=sha256:379bc261010c901bbd92effc3fada6c5c4482c6d132c1e22e77778b1551ddc8b

Observation 37445cdf-c368-4921-907f-b9a63da27640 · outbound

This paper cites Advances in neural information processing systems29(2016).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Advances in neural information processing systems29(2016)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.572902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.564650Z digest=sha256:9cfe5c9f1f82a60bbf33c16f8ce5efc91a503b62cc4974af77022e7294c5fb39

Observation fe0cb4b8-2b9c-4535-bc18-b60886a618c4 · outbound

This paper cites Frontiers in Neuroscience11, 159 (2017).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Frontiers in Neuroscience11, 159 (2017)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.558217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.569568Z digest=sha256:76f96be61df3aaa8ebb86cd73a0a79a3c16d8afae3b67afcf728887a469041c6

Observation 2226b01b-410e-4810-8a64-7fb3295ee0d7 · outbound

This paper cites Attention, Perception, & Psychophysics77(5), 1465– 1487 (2015).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Attention, Perception, & Psychophysics77(5), 1465– 1487 (2015)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.543677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.575305Z digest=sha256:bd47787f79d66326440abb8c0ed3a16ec04eaae1c2e894d628e1ead9254d7871

Observation 70bbde47-6c15-44e2-accc-24068ee7d583 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.527389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.580317Z digest=sha256:0f27774dd76659aa9ea41de0ca48c13d8f75e1374111d82022b45749e63b0f32

Observation 04861d27-f7da-46b9-9d0c-60e141688c93 · outbound

This paper cites In: ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.510649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.586258Z digest=sha256:e5d7680d1b520be7444bb208ab9184fa7b81bdbf587a9e636f3a6aee8f0b0349

Observation aa60035c-c75c-4be2-815f-8443961b437e · outbound

This paper cites In: Conference on Computer Vision and Pattern Recognition (CVPR) (2020).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Conference on Computer Vision and Pattern Recognition (CVPR) (2020)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.492991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.591367Z digest=sha256:ec0b663d82cdd63b496c00e2a4ed67b03484a1f9373d54e4d6062602f6496d0d

Observation 24df23de-f6f3-4649-96cf-a1d483f84a1b · outbound

This paper cites In: Proceedings of the IEEE/CVF International Confer- ence on Computer Vision (ICCV) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF International Confer- ence on Computer Vision (ICCV) (2025)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.476703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.596497Z digest=sha256:dc3f8afe37192aee686da6577832ebfa6acf8635ac71f091a351fcc030406dd4

Observation b39f2f56-35e7-4d00-a4a4-535a8ff02df8 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence (AAAI).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the AAAI Conference on Artificial Intelligence (AAAI)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.460940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.601464Z digest=sha256:31e85e93d8c33a513234e34c3a8da26f8caf73d7c85b6b579716e0a6aad57af5

Observation 3a424f31-5dac-467d-b69f-93376678c2f1 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2025)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.444767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.613119Z digest=sha256:a6ca63f66db38b518730b7ed3224eb59ed5b046382120e8e33fc0c2a792ce594

Observation 8ea478e6-0ac9-4039-9617-0d09ef77e084 · outbound

This paper cites In: Advances in Neural Information Processing Systems (NeurIPS).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Advances in Neural Information Processing Systems (NeurIPS)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.429207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.628526Z digest=sha256:ce0e9c77ef418b02f3182bbd05c90d2de8bd344200e3f0f3c93ecdf9b5b913f0

Observation 210ccabd-7b88-4e58-a277-8c47d2959df1 · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.645386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.645386Z digest=sha256:eb5f8f3ceca69c3662f8a8ad5feb9f349becc1bc05f7cd665cd9e2b0ecb677a4

Observation d322b711-de24-4baa-b2a8-46f8d5ca0d46 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.401361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.656480Z digest=sha256:18f720381b09a3ee055db3fcc0f927dd0c9e405880ac896017b48c3120dc5dad

Observation a441b1f1-0c6e-4e0c-9643-8db7ab5d6bcf · outbound

This paper cites Advances in Neural Information Processing Systems33, 10077–10087 (2020).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Advances in Neural Information Processing Systems33, 10077–10087 (2020)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.386110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.667931Z digest=sha256:69cd89224624e7e310558747748bb2bd0450c28a6a7d5787ed97307d495592ab

Observation fa0c01c6-d212-4035-81ba-28fb7fbe33e1 · outbound

This paper cites In: British Machine Vision Conference (BMVC) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: British Machine Vision Conference (BMVC) (2025)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.369268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.679994Z digest=sha256:aa6870fec2249e44ab9847ac2c9c00723c7b7739d988bcbb3b06403081f1eeeb

Observation cc5e74c0-e9c9-454e-9394-3862c0ec41f4 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition (2022) Whence the Voice? 33.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition (2022) Whence the Voice? 33

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.354037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.692047Z digest=sha256:b3ea701c1c7a752be35c8c6321ee059a19c6333b2fbffb7657708eda5cd9fecf

Observation a00b0d01-9f57-469b-939c-66eb56d2f339 · outbound

This paper cites In: British Machine Vision Conference (BMVC) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: British Machine Vision Conference (BMVC) (2025)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.337579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.704370Z digest=sha256:109d6475459991513f959d4b62b92ddbc632de025226985ec32b463956df0f08

Observation 01159ed1-3a04-4182-abd5-d43280ebe315 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.320224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.717929Z digest=sha256:ea5b3b3af5483b82ae0de4bfcccc3b22ecb6a4b5a6bb1e43b2bdb3042a9ad38c

Observation 0a9cdadb-e65e-470b-9873-9f088ea81ed5 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.303101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.731400Z digest=sha256:9c021398f53f3536837d568557cf49a4bd7f03782f7e2a269c9b0b1ceb056a78

Observation c799eb3e-e1a8-4529-98df-28647321f6d7 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Adam: A Method for Stochastic Optimization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.744967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.744967Z digest=sha256:644d0db9054f148f7d9f2e6a107fd953a26e2b4be619a4fa392e2c2bfb4ecb51

Observation 5eb0effc-344b-4a37-bbe8-7d646a28c510 · outbound

This paper cites In: Proceedings of the IEEE/CVF international conference on computer vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF international conference on computer vision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.759221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.759221Z digest=sha256:d9c979fc67ec6688d0c909669e20447a9bae9bff1b1c22150fc53485670157b8

Observation 9a3a2b95-8243-4ef0-b3e9-89a7ae82b892 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.275731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.771411Z digest=sha256:0b2849e0c9901761f3137414f730f0e9bc027927b86ae27c95973778d2cc1fd6

Observation afe5a8c0-2469-42dc-86ac-082e3f35f50f · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.260518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.776984Z digest=sha256:e2914af92c3e8559b12cc2c41e9b50450f40413259109bdb974aedd9203ad5fc

Observation 6289fb5c-0299-4add-99be-b18439494471 · outbound

This paper cites In: European Confer- ence on Computer Vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: European Confer- ence on Computer Vision

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.243499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.781720Z digest=sha256:e3ff7b44472321accc3782e13382c2984422daf6b1e8570e7d7a9267b095fa34

Observation 058b684d-7c8c-48f1-acb4-3e8b47e2325f · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:04:02.227849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.786501Z digest=sha256:38d46dd0bc43d7cc553e8cd429722d4d950dae53feb423888ed7be644c909b16

Observation 7d9570fc-408d-4b6a-b6aa-a5a106d818b8 · outbound

This paper cites Multi-scale Multi-instance Visual Sound Localization and Segmentation.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Multi-scale Multi-instance Visual Sound Localization and Segmentation

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T23:04:01.982686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.791902Z digest=sha256:385ed1b26395a57ce7192a26650759f48232a79fbc3cdf4b7089f9ef045ce136

Observation 0b86ad80-4dc9-45c6-b617-66949b7bdde8 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Representation Learning with Contrastive Predictive Coding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.797039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.797039Z digest=sha256:ee6fcdac22ca432220616c26d9f8e70c9e2662ad94e95c79d01407dc6872c8ea

Observation f8669019-ec0f-4286-b572-2f2f0744da26 · outbound

This paper cites In: Proceedings of the European conference on computer vision (ECCV).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the European conference on computer vision (ECCV)

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.213675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.803709Z digest=sha256:7676a6e522a1a867ec44ff2056a9ff1f10fa39ddb7fb7f40c3f2db0c81fd7d06

Observation 872970d9-7b66-42c4-8e95-ee3866be1439 · outbound

This paper cites In: European conference on computer vi- sion.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: European conference on computer vi- sion

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.198824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.810853Z digest=sha256:956d3eac7657b83e87367d76bf83880bf26de4496315e72cbb42fe9d159ee9da

Observation ab100932-b508-46dd-be9d-d176605f76d5 · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:04:02.184569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.815816Z digest=sha256:361425e47cbf212534611efe11b0b3ee64b06517e807bfb2a3b50d3e4786a2b0

Observation 2965b4f1-3b10-43fa-87b4-e4afc5d408c5 · outbound

This paper cites In: European Conference on Computer Vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: European Conference on Computer Vision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.169557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.823082Z digest=sha256:e7851cf549021e0e18be36693afe412ae07cc82bb8defb5a15eefeee0341afb6

Observation d95d7cd5-6837-4960-9822-0fff0adb090a · outbound

This paper cites In: Proceedings of the IEEE conference on computer vision and pattern recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE conference on computer vision and pattern recognition

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.154225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.828317Z digest=sha256:3702693aeb46986f09dbfab38c65143c66523237932c0f2852891861f7e9f914

Observation 56b0f1df-d741-4d6d-bcb7-4044eb785e8f · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence42(8), 1984–1997 (2020).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence IEEE Transactions on Pattern Analysis and Machine Intelligence42(8), 1984–1997 (2020)

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.137544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.834636Z digest=sha256:70700ae680a5f33f4307b6026df4dbe65124ccc8cb4408024dec7d6052579e5e

Observation 035a9a13-d765-4d3f-976b-538cc424f664 · outbound

This paper cites Hu et al.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Hu et al

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.120779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.847282Z digest=sha256:2316992209e36ce981f09e6002c62b8dad1e3c69c7e582f55a3525a55439cacc

Observation 9e0ca096-c3ef-4693-9934-eccc92b981c8 · outbound

This paper cites Journal of Speech, Language, and Hearing Research60(10), 2989–3000 (2017).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Journal of Speech, Language, and Hearing Research60(10), 2989–3000 (2017)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.102936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.854989Z digest=sha256:24c5a7517cf0d5baaf96df6ccef7cb803b9e3a8d1339727320c9529c1d2420a1

Observation 4395275b-730b-4d7f-90e9-fa21db9ffef9 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.086380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.860761Z digest=sha256:730cf6e303362ecf14846a1a9375a430e346aa2cefaf3f89641ae181afd117e0

Observation 0e8bbfe5-f322-4844-9222-d57d9560d1b3 · outbound

This paper cites In: International Conference on Machine Learning (2023).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: International Conference on Machine Learning (2023)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.068961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.866450Z digest=sha256:a24068b90827d42667e5cfa23cb4ff0dbc28a30597bc80c156f1d85f7135583b

Observation 1af77587-3acd-4773-83bf-f867cd91f2cc · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:04:02.050943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.873644Z digest=sha256:1591ace0d372e96f66ae00258e79114c4a62d26c371f725d4884bb6a91f65ec9

Observation 027d935b-c4ad-4bfa-976f-f0b81b9b0548 · outbound

This paper cites Audio-Visual Segmentation with Semantics.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Audio-Visual Segmentation with Semantics

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.879461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.879461Z digest=sha256:49754f33189a2c7df8f9ed46b18d32e955e6e011fae6c1bb052a5946b639a875

Observation 70c20b9a-212b-49ef-89d3-3efa484953b0 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV) (2022).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the European Conference on Computer Vision (ECCV) (2022)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.027050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T23:04:01.886110Z digest=sha256:5f16be4efbf5bdb25c16c11990879f714fc2aca7ba39be54565c626b47c4960a

Pith citing papers

No inbound Pith citation observations are available.