Pith. sign in

Paper Citation Record · LEDGER

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE

As of 9 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2608.03964.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03964 v2

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:38:54.409813Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 112c890a-9b33-4f67-b96c-f88dd6add5c8 · outbound

This paper cites ACM Transactions on Graphics , volume =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE ACM Transactions on Graphics , volume =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.211546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.211546Z digest=sha256:e4adec6b8503f8da5a67d8b84948af2f149c882eff935e9f26e6034ebafcb094

Observation 1b927815-17bc-4886-93d9-8328dba9f8f2 · outbound

This paper cites 2025 , publisher=.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2025 , publisher=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.216918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.216918Z digest=sha256:45c5bacb68c0d3ee6c3e83c0525a7bbca2c15452ecc0586c94123fa2a66ebda0

Observation 7e894187-9100-46a6-8547-10060c2cb37e · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.220345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.220345Z digest=sha256:6d5ade1516b7c778ebd4f28fe341b5d4d952a4726aa53b0302109193982b0527

Observation a800df13-33f0-4c3c-a4f2-850072516a9c · outbound

This paper cites Interspeech , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Interspeech , pages =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.223368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.223368Z digest=sha256:0c27506b5d438feb92efdda226c4a001b2243e56c0cb6264dedfcbca192010bf

Observation 61dd3564-683a-450b-845d-d1287cf94a61 · outbound

This paper cites 2022 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2022 , doi =

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.226354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.226354Z digest=sha256:2fffe547950fdca393a7d6ec0202116f73aa9bdfea39fee0e1a5770ca50d4844

Observation 60a2f955-fd1c-414c-84b4-e3f86d41a8db · outbound

This paper cites 2019 , publisher=.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2019 , publisher=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.230320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.230320Z digest=sha256:b592b672d22d1eb5a9dbae2cc671ee67b74a04697c9228855f4e4f0c8320a488

Observation b1da064b-7657-4104-8e98-88b95502d472 · outbound

This paper cites IEEE/ACM Transactions on Audio, Speech, and Language Processing , volume =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE IEEE/ACM Transactions on Audio, Speech, and Language Processing , volume =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.234316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.234316Z digest=sha256:8aa766f67073cfeb371fb772eb254794dfd61f1fd2338e9203bdcc40f26edd13

Observation 79a3f8dc-cd62-4314-9f00-1e1d96af9493 · outbound

This paper cites Interspeech , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Interspeech , pages =

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.238054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.238054Z digest=sha256:fed0ed87db86941d58a439375fb3a41bfd06782a66221f4f2b09f591fe3dd826

Observation 20b3a242-fc20-4c28-9ed8-f26fd7675079 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.241584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.241584Z digest=sha256:e5095ab1f8e67ac32e1180aafe1f0896b6455649183ee111797e65ebc3e0f4b4

Observation 2194eccc-70b0-434a-b99e-f97e018babd5 · outbound

This paper cites 2018 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2018 , doi =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.245641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.245641Z digest=sha256:d55d976dff3a93ebaa57ae53a03b33f94f53d4ed24911f05ff44551c73525228

Observation d1d82b5c-84ad-4417-9f7b-419d06fa569a · outbound

This paper cites , booktitle =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE , booktitle =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.249341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.249341Z digest=sha256:83b8711479110a8846f5496e5a6f365f4957a4f82057f2dc84edfcfeeb581d64

Observation 539382fb-ce8d-46d8-a2ae-417fbefb4304 · outbound

This paper cites 2017 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2017 , doi =

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.252885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.252885Z digest=sha256:f7df091ff04b54f6240c43378163a2673f51561917d46210d14bc04f4b42ba1b

Observation 7a3a86a3-2fcf-41b5-9193-2af8c5f0e682 · outbound

This paper cites 2023 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2023 , doi =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.256292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.256292Z digest=sha256:089d708b573cc5ba88167fe8f7fa32405601e6396ef2feade482834e7894eba0

Observation ad1c46d9-1b70-4424-8411-d6c996dd1d43 · outbound

This paper cites Scenario-Aware Audio-Visual.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Scenario-Aware Audio-Visual

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.259841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.259841Z digest=sha256:0983789f2193567fd3831bd4aced3dac6de6e5487795779b896c3424cd7f49aa

Observation 1d007458-36a0-48fd-9d24-6c4607fd5b62 · outbound

This paper cites , title =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE , title =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.263337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.263337Z digest=sha256:f02cd62a86b640c6f71b240dd670ccf353b234b04a0066dce40abc23d5faa29d

Observation e88cb523-481f-4f3f-be00-5b88767ba1c1 · outbound

This paper cites 2018 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2018 , doi =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.266927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.266927Z digest=sha256:ac136b0b1e096bbb97872da672e06e7ea30b445f4e3bda7611062098f25ed1ac

Observation 3f31e0d4-41af-40d5-93f5-3a3e0df90af0 · outbound

This paper cites 2020 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2020 , doi =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.270343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.270343Z digest=sha256:1f68fed655d51be3c330c0a3bf1b374320efb9926ef74778174cc89f836c7a96

Observation ef4f2bf9-77aa-4a9e-8de0-12b4b396d984 · outbound

This paper cites 2020 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2020 , doi =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.274039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.274039Z digest=sha256:0704dd3a384dafa2e8b4dcf3ce28320300242856a5cd8dcc326cbbcba9800b6a

Observation 4266f7a6-6c7c-4505-9e9b-285d22f33640 · outbound

This paper cites IEEE International Conference on Computer Vision , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE IEEE International Conference on Computer Vision , pages =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.277716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.277716Z digest=sha256:132a835cc8897a89da29a0f4fba2d918cf824744f77a8dc94d406206994f0222

Observation 6f8e3df5-751c-4209-8af3-da698735f802 · outbound

This paper cites 2015 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2015 , doi =

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.281291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.281291Z digest=sha256:0c43de3ceb5bf84d5baa96a51eb14b3399e2779486eb5de2851587306d8321df

Observation 807b4805-0e65-461e-aa90-f9a65874a224 · outbound

This paper cites and Zisserman, Andrew , booktitle =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE and Zisserman, Andrew , booktitle =

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.285144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.285144Z digest=sha256:7a9ce62995c278ef418efe743259567db77f7b6fff37e0466cac28578a878159

Observation ab9bbf30-5cc8-4f92-b1bd-abca76a9406e · outbound

This paper cites IEEE Transactions on Information Theory , volume =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE IEEE Transactions on Information Theory , volume =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.289014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.289014Z digest=sha256:1c305b66c75b56ebbd8467253ad044321e66ff25aafea50f53228883dd8231fb

Observation 5984d529-99f9-419b-a51a-12911f513f01 · outbound

This paper cites ACM Multimedia , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE ACM Multimedia , pages =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.292831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.292831Z digest=sha256:e2577d56b5d0a3868c2813fe2e82ecf19ec7d68ca15605e307e98443ea9589cd

Observation 28122e67-d504-4689-9c89-c85ca1dddeed · outbound

This paper cites The Annals of Statistics , volume =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE The Annals of Statistics , volume =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.296240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.296240Z digest=sha256:e5424ed424406dbff947ce178ddb2be425133273e8e8fadac1c5a4b634db625b

Observation a02291c4-9efc-460e-99ac-86008ad23aab · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.308651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.308651Z digest=sha256:3e0bc705db943d898985e3d7e1777f8ca2bd4dbb244555d78ac3390dd6319e68

Observation 03294fde-dc69-438b-a96b-649434e648de · outbound

This paper cites 2026 , doi =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2026 , doi =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.311963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.311963Z digest=sha256:0ef2d245d8141f5d096006d44900057f3c564ad22e707b2dfe903f3067ae59e4

Observation 7b468ec1-be23-4dfc-8e6b-404de216cc77 · outbound

This paper cites 2025 , pages =.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE 2025 , pages =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.315406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.315406Z digest=sha256:f2dfa17b69f11c1c5764998cbd265f9c3b0d197ff53f027a786e92f0e6df33bd

Observation b333e601-2da3-4327-8a19-d5cfc83df7de · outbound

This paper cites ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , pages=.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , pages=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.318691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.318691Z digest=sha256:df0403a54a213dc9a5e3defa002bdc80bfb828c02caa3b85cc2d93270ac5caab

Observation 86254853-121b-4cfb-a3a2-9fd481c3351c · outbound

This paper cites LRS3-TED: a large-scale dataset for visual speech recognition.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE LRS3-TED: a large-scale dataset for visual speech recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.321955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.321955Z digest=sha256:538840fe79df1f9d64ab50a048fff3b0a001f0f91ec390be9600a38e6caed111

Observation 474d28ca-f1d0-43b0-993b-49d471749e2c · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.325800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.325800Z digest=sha256:586964df5ee29cc0e5d50f8ac30c6dc908e5ab3e71e2c2007f3554f85000a21e

Observation 0a328093-3103-4743-9785-21100acc16f8 · outbound

This paper cites Parkhi, and Andrew Zisserman.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Parkhi, and Andrew Zisserman

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.329420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.329420Z digest=sha256:9ec7783704f3fa2fc761bee151001f89b1728875dc91c24a6cc5d80246393254

Observation 65c05404-29c9-4730-bba4-30d59e2db9d7 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.333168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.333168Z digest=sha256:fb9ce33e1fed7c576dd8b7382673fe7ed76130c79c8c9cef48264c9351b2d47c

Observation 8f61b9d5-5d86-4ab4-afb3-33107d51200a · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.336600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.336600Z digest=sha256:60108e67fc85b324f71821772b853ea05c7c58c3bcb48f2ff22901eb84245778

Observation 313acf9d-0870-4420-abe3-4b99a0d74da6 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.339648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.339648Z digest=sha256:1fdb1996d59c4ac87d1932c38056f254e8f646255093ef76ed5918bf807ec2f1

Observation ea59e885-92f5-4560-a346-74d4ce5299b2 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.342520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.342520Z digest=sha256:c8c8c1e2cf1fa6f3740654552a4b9399503db2e07ffb3764aa9a310071d6cebb

Observation bb2f8a39-2f60-4774-bd1c-231f8f3b3861 · outbound

This paper cites Freeman, and Michael Rubinstein.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Freeman, and Michael Rubinstein

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.345337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.345337Z digest=sha256:77ace6c7b2ebe02b109e374faaa3ad8ccd856bd0ad15afad7b5bc99b4122fca0

Observation 96e1c602-b0f3-4613-b820-20faf5b74c85 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 39

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-08T00:38:55.208961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:38:54.348174Z digest=sha256:e4931b06570938a41976aa7978a54b1b976ebe36aa1ddab4013549c033aebe11

Observation 46ca842a-cc4c-49d1-9e69-0373a9275f21 · outbound

This paper cites Levenshtein.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Levenshtein

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.351032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.351032Z digest=sha256:c2e40342726e8bb3768a5ecad5cb7dd64215ebc285e8ad9336880af02431f1d2

Observation f662acfe-81d3-45c6-b467-6f7ec3a4da23 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.353708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.353708Z digest=sha256:2a0850d4326c523da479310f3bca16e1717449732915d04119f604908dad964e

Observation 044b5cc1-d2ee-45b1-9a6b-03240fad9dde · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.356936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.356936Z digest=sha256:fc3e05c570577e667e46d17d6796543eefb2dca5e5879464c5eb19856781f7b6

Observation 5ed23fc1-67b8-474b-ac58-dba8ecfaf0ca · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.359955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.359955Z digest=sha256:eed3be251153dea45cfa719d4d2a7f991fd782e69cfb19ff13bfb8232102bde5

Observation 042ccf02-0a6f-4e25-b1fe-615dfb96d307 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.363402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.363402Z digest=sha256:10a589541fa7ea5ba18acb3014a235557f7513dbc45f402bfb3f87abb27f3221

Observation 61a2b157-5986-4c77-9182-5c40862a7f61 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.366914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.366914Z digest=sha256:904190fcbc9d6a4a98d9e3f0bf8ddf28ad0da3a8ca3c88e0fc68c35f381eedc6

Observation 92025b76-f9ab-4d82-832e-057e8e627e9b · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 46

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T00:38:55.094647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:38:54.370357Z digest=sha256:cc3c3cc94f6ce87a6bac1a78c7fbbbf611e0bb35aefef73248631cb6a8148160

Observation 94bb6707-9a05-4651-94e0-7ac82a880576 · outbound

This paper cites Germain, Sameer Khurana, Chiori Hori, and Jonathan Le Roux.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Germain, Sameer Khurana, Chiori Hori, and Jonathan Le Roux

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.373896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.373896Z digest=sha256:310e7c0e17d8efb58050a114975b1c4a3ca4d4ddfb4a662bcd12b6d2a69d08e8

Observation 124d17c8-d2f7-4916-aa70-48a6b57d695c · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.377234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.377234Z digest=sha256:5ef386ffc03e6d97f1a24f22c12f849357dbbc834d36c3ffe7d8ab7d4288c7c6

Observation 0308c74e-745e-4bdb-80a9-fc395a666122 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.380506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.380506Z digest=sha256:b3cc66c61678db13922f4925113b61c91d912ee35509c4a7a1eba9b31f93bb3a

Observation 137980d6-499a-466b-8246-37875d31c9a0 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.383895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.383895Z digest=sha256:8a05319fd0c5c4fdcdf19f8e7c5f1273cbcd030c5891abaf44b35467c424b85e

Observation f48c3caa-2414-4616-9134-fb48b647a39f · outbound

This paper cites Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.387173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.387173Z digest=sha256:c1d14bb0a8bc6bc54f05cd208c8db7340a2f58f5ccad2e44c3e2f1efe8b22b7f

Observation 5862b241-aec2-4ec6-929c-eb3016af799b · outbound

This paper cites Qwen3-ASR Technical Report.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Qwen3-ASR Technical Report

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.390520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.390520Z digest=sha256:9aae2210e619269abfe3d3591861c14de1f0df1d61e537c4e9bff27ac7538159

Observation 90aa7cfb-9ff0-4588-8bc7-9aca6531904e · outbound

This paper cites M2S-AVSR: Modality-aware Multi-view Self-supervised Representation for Robust Audio-Visual Speech Recognition.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE M2S-AVSR: Modality-aware Multi-view Self-supervised Representation for Robust Audio-Visual Speech Recognition

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.394031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.394031Z digest=sha256:995c95ecc738d7644cda4d12dbb806e745998f15870e884de265fc97238151fb

Observation 2524002a-08e6-4e81-966b-538796c011ba · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.397062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.397062Z digest=sha256:d2936aac0fe710c19a71187a4930b56fe5984e8ca052e5cc5508443d27327ca3

Observation f153512a-cb7b-4cd9-a6db-55261bd5f2c3 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.400282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.400282Z digest=sha256:cf86e3b470f0fc0093da5537d76c152429eee7d2ae195bad921d98c32a6f588d

Observation b0693668-e234-4b58-b3d5-4e5d405ea64d · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 56

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T00:38:54.704054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:38:54.403395Z digest=sha256:b37149399314f116baa6c86be44f8973853b44c410e97f1c94d9490cc5049b66

Observation 69a62879-b9c3-403c-8956-f050ccbad675 · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.406547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.406547Z digest=sha256:dbe246899941794efe19932612df4753577f0255fc408b36d8a4820584f76961

Observation 3fa79b04-247c-4abe-bdf6-87be5cef118c · outbound

This paper cites an unresolved cited work.

Identity-Faithful Audio-Visual Target Speaker Extraction with REAL-2MIX and VOXBLINK2-AVSE Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T00:38:54.409813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:38:54.409813Z digest=sha256:2e7456355543cd25aa0000919dac39e68ba6111630336d37ea7e96aa75897fde

Pith citing papers

No inbound Pith citation observations are available.