Pith. sign in

Paper Citation Record · LEDGER

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances

As of 10 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2506.00636.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00636 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:05:03.227528Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:05:03.141073Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:05:03.298932Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact5
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a1891f3-1f4c-4d7d-881a-20ccc08b0562 · outbound

This paper cites ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.301955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.141073Z digest=sha256:e3fdf00a39338c091cdb7aa3392f014f949d26813707999ac453405b33033e70

Observation e1e4d5c6-e01b-496e-95d3-ba1f7f2f1334 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.508277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.144779Z digest=sha256:0c9f854d52c43be8be20e3b6e1e4959e91eb055b97add4b9b4cda06b4144c395

Observation fc2fbd6b-5810-4365-bb28-9d08b6151617 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.500041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.147970Z digest=sha256:ac3c7de8d6467eaead584d3d8c6abc3213e9b44038122ee28be54bd71c866988

Observation bb5d3e38-29c8-4642-b525-42dc8efed1d8 · outbound

This paper cites The process is outlined in the following sections: data pre-processing, evaluation metrics, and speech recognition experimental results.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances The process is outlined in the following sections: data pre-processing, evaluation metrics, and speech recognition experimental results

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.492201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.151062Z digest=sha256:76d213430d920ffd945c82977d4f227f70c7dafa3228d9913b1dd0658699631f

Observation 282a035e-c133-47e8-82a2-9baf33bbfe90 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.483692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.155247Z digest=sha256:b9a086bed6566ea34169c07063a1039d293e4f1db5dbbd986b6068c6d7f6cfcc

Observation 584d35fc-c01d-4a1f-833a-0634ffdc37bc · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.475684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.158100Z digest=sha256:a319904fa4fab4d70e8132dc01819d7f4733e332c21992403a9e52d858d35f6d

Observation 2c828336-73ec-4068-b1dd-cbe31a96f438 · outbound

This paper cites The Sounds of Cyber Threats.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances The Sounds of Cyber Threats

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.290876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.161082Z digest=sha256:f202d5cbeb61132b625319b61a6a33e2c5ae1367df9822cb7fce6cb6dd8f6575

Observation e0a19c19-706e-4291-9215-f19d090345ce · outbound

This paper cites Audio-based toxic lan- guage classification using self-attentive convolutional neural net- work,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Audio-based toxic lan- guage classification using self-attentive convolutional neural net- work,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.466836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.164369Z digest=sha256:5d7c046b62c14f024056aaed1b3f96a3f060be14ad98fd875173298652dbbcf0

Observation 601813b3-5b84-4a08-8102-d213c4359ff4 · outbound

This paper cites Linguistic analysis of toxic behavior in an online video game,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Linguistic analysis of toxic behavior in an online video game,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.457438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.167350Z digest=sha256:4861ce34b9cc0f5aa269f06e6205ffcd175e8d49875a627ff27035a03b9377aa

Observation 9eaa9569-0c6b-4c26-8a4e-60fbc12876cf · outbound

This paper cites The influence of violent media on children and adolescents: A public-health approach,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances The influence of violent media on children and adolescents: A public-health approach,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.448239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.170465Z digest=sha256:4d5546492bc6dd32c5651a7fa5ac89f6b72e37585999454a08ed9383b85a4e17

Observation b48be41c-63eb-4a03-999b-62cfcb001f08 · outbound

This paper cites An exploratory analysis of the relation between offensive language and mental health,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances An exploratory analysis of the relation between offensive language and mental health,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.439072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.173194Z digest=sha256:904a6b628252b6983b291b0ebf43e49e5fcf757f63396c9fe216dd01d7f309c6

Observation a044bc26-d0b9-492f-a130-2ef2e682e2ed · outbound

This paper cites Choosing appropriate language to reduce the stigma around mental illness and substance use disorders,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Choosing appropriate language to reduce the stigma around mental illness and substance use disorders,

Reference 12

Resolution
verified exact
doi, observed 2026-08-07T12:05:03.259911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.176436Z digest=sha256:b7e26b25e035205546aebb2357e35dbee1586d90d091d7600e041614b1a3d8d1

Observation 9bf5e81a-97de-4271-933d-0d434d35fcaa · outbound

This paper cites Exploring the distinctive tweeting patterns of toxic twitter users,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Exploring the distinctive tweeting patterns of toxic twitter users,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.430288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.179409Z digest=sha256:02a24b92e4d783d4b4622b2dba71b0ddaca66dea0fcd45381db3f771a3c130c0

Observation 23013af6-4c1c-43d8-9353-b706fb5cd854 · outbound

This paper cites Detoxy: A large-scale multimodal dataset for toxicity classifica- tion in spoken utterances,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Detoxy: A large-scale multimodal dataset for toxicity classifica- tion in spoken utterances,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.421683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.182238Z digest=sha256:f8bb22f5c862706bab965c5f7f99c3d415b8a0261de30a23e2486141beca407d

Observation 2fd4218d-b510-4b72-818e-d984a27a501b · outbound

This paper cites MuTox: Universal MUltilingual audio-based TOXicity dataset and zero- shot detector,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances MuTox: Universal MUltilingual audio-based TOXicity dataset and zero- shot detector,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.413069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.184864Z digest=sha256:f56ac8703de51af37a4d60371bb6fd4b40410ed875692bfdb99fa2ea87e051b8

Observation 9553aa5d-1e7e-4bad-88fd-00d5589a65ba · outbound

This paper cites Lightweight Toxicity Detection in Spoken Language: A Transformer-based Approach for Edge Devices.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Lightweight Toxicity Detection in Spoken Language: A Transformer-based Approach for Edge Devices

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.279574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.187654Z digest=sha256:2327c27987b83bccc9a643d00402e505e50555fc4437885150d14137963a47d3

Observation 765ba09d-117d-4692-b12e-a11afd668e06 · outbound

This paper cites Enhancing multilingual voice toxicity detection with speech-text alignment,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Enhancing multilingual voice toxicity detection with speech-text alignment,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.403902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.190570Z digest=sha256:b22413d952e7a4902e70795e883cf3269da916b365b1d1ab98999ac375553fa3

Observation fc62ed87-eb80-4177-a043-365fc6870659 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-08-07T12:05:03.250183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.193220Z digest=sha256:12f31e031f9b21722dc74be74f484b3b35bb0b59f1b2e700f0cf0c7d4732f92b

Observation 6ca94240-3683-4e59-8bc6-6983bf466a1a · outbound

This paper cites A large-scale dataset for hate speech detection on vietnamese social media texts,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances A large-scale dataset for hate speech detection on vietnamese social media texts,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.395573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.195986Z digest=sha256:55a0ce2184ec2d4dbf988dde748316b1d3731fb69507a3c7b74c0c964bbc0cdd

Observation 9fe3260f-4bb9-443e-a030-6c128f2efc2e · outbound

This paper cites ViHOS: Hate speech spans detection for Vietnamese,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViHOS: Hate speech spans detection for Vietnamese,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.385999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.199708Z digest=sha256:e348d10f9c50bce8fe11c83ce1c5b5164682173fcadd460e80c3783dcaee2a5b

Observation fe09ce50-4340-4f3f-805d-84bcabb35fa1 · outbound

This paper cites ViHateT5: Enhancing hate speech detection in Vietnamese with a unified text-to-text transformer model,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViHateT5: Enhancing hate speech detection in Vietnamese with a unified text-to-text transformer model,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.377160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.202438Z digest=sha256:617c7854bc169dbbfcfa6073f824087c7329bce3b52fab7e42fb0efa53ee8ce0

Observation 76572e13-54fc-46e9-905c-8d283365327d · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Robust speech recognition via large-scale weak su- pervision,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.368473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.205503Z digest=sha256:ee94a3676c7153cb373ae0a4268cfe264657fdcc1e5494c7c5b85e90634799a9

Observation d1520d0c-5e29-440e-b38f-c494d8c52efe · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:03.208155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:03.208155Z digest=sha256:0932346d60bc5bb97fc39068675967aa6bb38a5a9c656307081ee59e3e61e4d7

Observation d0ae4b93-1e95-4d97-b11f-4d4e3186235c · outbound

This paper cites PhoWhisper: Auto- matic Speech Recognition for Vietnamese,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances PhoWhisper: Auto- matic Speech Recognition for Vietnamese,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.354799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.210818Z digest=sha256:612eaa3f4a4afef996b5e3327ac0ac33104e9abe2422057341730fa8d56e7a5b

Observation c379f1c8-4b2f-44f3-8799-0cfedcdfaad3 · outbound

This paper cites Unsupervised cross-lingual representation learning at scale,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unsupervised cross-lingual representation learning at scale,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.345270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.213464Z digest=sha256:9d3a75bf9a642fcb8e9f515534b0008dd2f7ed2bda56908e6f6fb61b1b480d52

Observation 88481b36-ea3f-4be8-919b-74e849601a30 · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.337261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.216252Z digest=sha256:44a068ca6ccb32960423be2aec674c9a73512cb6078902edf561c2e415c8a117

Observation 4328bd90-4d4f-4dfc-a110-81e6ce1e0c46 · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:03.219241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:03.219241Z digest=sha256:902d386d44eae8726b0d60abfb107f591814d1cd16dba857c5bf55da384b5eed

Observation 527822ff-6376-4ce3-a254-99cf9691e5a9 · outbound

This paper cites PhoBERT: Pre-trained language models for Vietnamese,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances PhoBERT: Pre-trained language models for Vietnamese,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.328697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.222062Z digest=sha256:e5322577ff01fe94919de8fc25aeda6a56673b4b93e35ef121b44cce122aeb31

Observation b2116d6d-accf-4f4b-be0e-20c920601f98 · outbound

This paper cites ViSoBERT: A pre-trained language model for Vietnamese social media text processing,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViSoBERT: A pre-trained language model for Vietnamese social media text processing,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.319492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.224863Z digest=sha256:01f72ad2457c9b10ce0d10ef1c28335a836f25c61371ed093151bba668cc81ff

Observation ad2e927f-9ba8-4197-a31c-522f3c8784d5 · outbound

This paper cites VLUE: A new benchmark and multi-task knowledge transfer learning for Vietnamese natural language understanding,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances VLUE: A new benchmark and multi-task knowledge transfer learning for Vietnamese natural language understanding,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.311260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.227528Z digest=sha256:9afcb85aae691a6f2bf7b93fbecff1e50d4f860ffe1595909dc2171dbed1f4f1

Pith citing papers

Observation 8a1891f3-1f4c-4d7d-881a-20ccc08b0562 · inbound

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances cites this paper.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.301955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:05:03.141073Z digest=sha256:e3fdf00a39338c091cdb7aa3392f014f949d26813707999ac453405b33033e70