Pith. sign in

Paper Citation Record · LEDGER

CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2303.00332.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.00332 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:32:33.180493Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 48d44a37-312b-4421-a845-c9f38f110ee0 · inbound

A Multi-task Learning Balanced Attention Convolutional Neural Network Model for Few-shot Underwater Acoustic Target Recognition cites this paper.

A Multi-task Learning Balanced Attention Convolutional Neural Network Model for Few-shot Underwater Acoustic Target Recognition CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:01:57.741067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:00:17.697217Z digest=sha256:e13ae042e82c32fa359b24e819c20eaa1389930f735664c64c73ab651c97f91e

Observation e320bea5-6e39-462c-a823-256f39134ddf · inbound

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin cites this paper.

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:32:33.180493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:32:33.180493Z digest=sha256:4efec7f12a36c174718c91196c0a0937a3fe39145353ae97f9c24fccd357e84a

Observation e3a807a5-cb05-448e-beb9-10498a10e830 · inbound

Towards Robust Overlapping Speech Detection: A Speaker-Aware Progressive Approach Using WavLM cites this paper.

Towards Robust Overlapping Speech Detection: A Speaker-Aware Progressive Approach Using WavLM CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:55:24.199871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:55:24.199871Z digest=sha256:a07f2f4e1c5753e33b985ebaaf5549d1c47f7e20fc6e56bbca6e3e7f1126955e

Observation 23f45442-877c-47ea-b01a-055032f9d6d4 · inbound

Seewo's Submission to MLC-SLM: Lessons learned from Speech Reasoning Language Models cites this paper.

Seewo's Submission to MLC-SLM: Lessons learned from Speech Reasoning Language Models CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:51.583965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:51.583965Z digest=sha256:dd180c2021e030b76b98fdcbe148cb660515563eadebf1585af47371531621d8

Observation 1171e4f6-1351-435f-ae4e-ac3f5ac56389 · inbound

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training cites this paper.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.032757Z digest=sha256:666d59a5cf306a28622d8071bd0096e99b8dfe3116ded38bf3cf25f1ad85fbca

Observation 30aa2119-5b56-43ca-b1f6-497c92a57ab1 · inbound

DRASP: A Dual-Resolution Attentive Statistics Pooling Framework for Automatic MOS Prediction cites this paper.

DRASP: A Dual-Resolution Attentive Statistics Pooling Framework for Automatic MOS Prediction CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T14:22:47.686311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:22:47.686311Z digest=sha256:1de6cfc26a206d69776d2aebef5b8f122a9c2333da516ac88a2c51d4cee0bd04

Observation 81c03d8f-5cba-4908-9af8-282172876454 · inbound

Effective Modeling of Critical Contextual Information for TDNN-based Speaker Verification cites this paper.

Effective Modeling of Critical Contextual Information for TDNN-based Speaker Verification CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T18:29:41.653077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:29:41.653077Z digest=sha256:3dd42251e81ab12682c6f46ef2211b6f3f71a18922b0d7853d61d5fea2d7b07a

Observation 32da00fe-042d-4ab6-b446-69f2dc58c249 · inbound

Logics-Parsing-Omni Technical Report cites this paper.

Logics-Parsing-Omni Technical Report CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T13:40:01.757425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T13:37:44.189839Z digest=sha256:005583aa900af5ca8e7e90524a2425964f36525ceea668376775bd7d9bf2665c

Observation 99df1911-2c91-4c2a-862f-c308e9aada21 · inbound

Controllable Singing Style Conversion with Boundary-Aware Information Bottleneck cites this paper.

Controllable Singing Style Conversion with Boundary-Aware Information Bottleneck CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:50:51.793968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:51:38.030059Z digest=sha256:069a17e19ee82298258f6c4a97cef41a5ec6b6ae24d07a5ac887a195a6ec4567

Observation 899481e7-e519-45d9-998e-318cbbf2727c · inbound

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing cites this paper.

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:50:56.052673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:03:09.942984Z digest=sha256:3d3769571a035e999b7a2bf487a9df5162c07d6f273fa47a5767efaaccc107e3

Observation b8ce56cf-20b4-4b4f-9127-c046beb3624c · inbound

SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue cites this paper.

SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:26:13.029513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T21:05:54.061395Z digest=sha256:974fa524c597463bdda4ed13750c185dbfa66f992a3ebe21e3285480df183712

Observation 1d286042-38d6-46a9-87d2-4fdafd0c2d28 · inbound

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation cites this paper.

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:57:19.589478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T21:14:08.243894Z digest=sha256:d2022a945d4f7000f40b889e3ede537934a4d5dc1c08ea9cf031ae796db216b2

Observation 915faf9c-9a63-43c0-b69b-899a9e0a3c3b · inbound

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis cites this paper.

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:10:07.240619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-25T20:43:29.118351Z digest=sha256:f80608ea398f617c07d0bae79faac45d30c05ebd04bc11d9bd1269b2848f39f3

Observation 1873aa41-be37-4671-9839-1675ce152f0b · inbound

Pmeta-TLA: Backdoor Attacks for Speech Classification Models via Meta-Learning with Timbre Leakage Attack cites this paper.

Pmeta-TLA: Backdoor Attacks for Speech Classification Models via Meta-Learning with Timbre Leakage Attack CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:38:04.489786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T11:34:50.959335Z digest=sha256:fb74be265f6e00d764e65bd7c3b6ac592e97ad8cf169f641ce70b1b9c75a03d2

Observation 9ab2deed-da03-494e-aeca-ea8ca106afef · inbound

DRL-CLBA: A Clean Label Backdoor Attack for Speech Classification via DDPG Reinforcement Learning cites this paper.

DRL-CLBA: A Clean Label Backdoor Attack for Speech Classification via DDPG Reinforcement Learning CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:18:22.475395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T14:14:15.384573Z digest=sha256:9dc3c009a82a3c8825cb225f16d4528614e1bc255de213aabb7da23387e332ad

Observation 876a18b8-983d-45ac-8849-d68f7812da0f · inbound

NouveauVoice: Generating Novel Pseudo Speakers for Voice Anonymization cites this paper.

NouveauVoice: Generating Novel Pseudo Speakers for Voice Anonymization CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T22:31:42.563915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:31:42.563915Z digest=sha256:e1465b0cdaca9516f5d319e12592ef5ce86c5825b4e96098265d40aaf5bb77ca

Observation 9842536f-f687-4b94-bc37-531637d17157 · inbound

X-Translator: A Real-Time Multilingual Speaker-Aware Speech-to-Speech Translation System cites this paper.

X-Translator: A Real-Time Multilingual Speaker-Aware Speech-to-Speech Translation System CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:43.822357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:43.822357Z digest=sha256:a198f7be150ba573e5f8b2f439b16d642dca4001714995b72963a30e9db27ec7

Observation 7087df11-3eec-4a77-8acc-9f06df93ee5e · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-04T16:29:28.771675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:29:28.771675Z digest=sha256:d438e4a9fe0f5d5cb8364a404bb9fa16311f3b1c51b7ab3331a54545ecdb9f5f

Observation dde9d30c-28e3-48db-bfa2-a2f7981db857 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:48.305543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:48.305543Z digest=sha256:22b52318bd172d06bc821707fe76215befab3d8967048065a194bb7c1a88e238