Pith. sign in

Paper Citation Record · LEDGER

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances

As of 8 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2506.00636.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00636 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:05:03.227528Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:05:03.141073Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:05:03.298932Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact5
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a1891f3-1f4c-4d7d-881a-20ccc08b0562 · outbound

This paper cites ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.301955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.141073Z digest=sha256:ad8e49c1a9fdfcf6faa0b60762ff22f4908df001b964d5795fe39e0bbee2146f

Observation e1e4d5c6-e01b-496e-95d3-ba1f7f2f1334 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.508277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.144779Z digest=sha256:bafc2ccfe0460fccdfbf1bcf3741308945f00fb1af441f64823a9b9a1f9d04e8

Observation fc2fbd6b-5810-4365-bb28-9d08b6151617 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.500041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.147970Z digest=sha256:efbc3f739dde3dd4d3690a55bfd267858b461c596903e94a4a3fb6961658bb09

Observation bb5d3e38-29c8-4642-b525-42dc8efed1d8 · outbound

This paper cites The process is outlined in the following sections: data pre-processing, evaluation metrics, and speech recognition experimental results.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances The process is outlined in the following sections: data pre-processing, evaluation metrics, and speech recognition experimental results

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.492201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.151062Z digest=sha256:57f687b1fef1b3a848b353e19a47742b6aa2b304cdb977bb3f73df89879ec798

Observation 282a035e-c133-47e8-82a2-9baf33bbfe90 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.483692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.155247Z digest=sha256:d0f20a05b55c5fef2557ad66b7fa7b85ad0c43a388a3067f9c5c48ba946adc61

Observation 584d35fc-c01d-4a1f-833a-0634ffdc37bc · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:05:03.475684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.158100Z digest=sha256:c3e8af17bd8c0205b15deb0e67e30eb85c2d7dc51bc4b0003c30aeff4f0e3602

Observation 2c828336-73ec-4068-b1dd-cbe31a96f438 · outbound

This paper cites The Sounds of Cyber Threats.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances The Sounds of Cyber Threats

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.290876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.161082Z digest=sha256:fa6f38527eb7638a1d3f411eec16b40b7641cd1e7797c341becab8b72efdfdad

Observation e0a19c19-706e-4291-9215-f19d090345ce · outbound

This paper cites Audio-based toxic lan- guage classification using self-attentive convolutional neural net- work,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Audio-based toxic lan- guage classification using self-attentive convolutional neural net- work,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.466836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.164369Z digest=sha256:a3616067526c26eb28ada0a09c96532d97fdf3f956c482a1b46d74512b7fc27b

Observation 601813b3-5b84-4a08-8102-d213c4359ff4 · outbound

This paper cites Linguistic analysis of toxic behavior in an online video game,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Linguistic analysis of toxic behavior in an online video game,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.457438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.167350Z digest=sha256:7a363a98578dd0e81691a253fc3c868214751fb0dc96da2ad6975405cb315e9f

Observation 9eaa9569-0c6b-4c26-8a4e-60fbc12876cf · outbound

This paper cites The influence of violent media on children and adolescents: A public-health approach,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances The influence of violent media on children and adolescents: A public-health approach,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.448239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.170465Z digest=sha256:621dcf8a7968f9024e7f7144603e9076f5f0162fce441d7ca2620dcd4ddeb6c3

Observation b48be41c-63eb-4a03-999b-62cfcb001f08 · outbound

This paper cites An exploratory analysis of the relation between offensive language and mental health,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances An exploratory analysis of the relation between offensive language and mental health,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.439072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.173194Z digest=sha256:6b01dc2ecd640b920cae4b56ae84f666a7769a57fa4d1a1f09abcec8ecf08fc9

Observation a044bc26-d0b9-492f-a130-2ef2e682e2ed · outbound

This paper cites Choosing appropriate language to reduce the stigma around mental illness and substance use disorders,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Choosing appropriate language to reduce the stigma around mental illness and substance use disorders,

Reference 12

Resolution
verified exact
doi, observed 2026-08-07T12:05:03.259911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.176436Z digest=sha256:2e3a48cb083842ae3ea4190d4ed63137ff62e8246086abf689e8cf9632009470

Observation 9bf5e81a-97de-4271-933d-0d434d35fcaa · outbound

This paper cites Exploring the distinctive tweeting patterns of toxic twitter users,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Exploring the distinctive tweeting patterns of toxic twitter users,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.430288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.179409Z digest=sha256:7135d24aa6b374f5207e2c8f8ebf217497f4ef7072f3eea257a5e155db173d2f

Observation 23013af6-4c1c-43d8-9353-b706fb5cd854 · outbound

This paper cites Detoxy: A large-scale multimodal dataset for toxicity classifica- tion in spoken utterances,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Detoxy: A large-scale multimodal dataset for toxicity classifica- tion in spoken utterances,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.421683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.182238Z digest=sha256:7b77567454b043390459bb6a4d9f4a952dc3f92b47ab4d35d02ec7c4d662434f

Observation 2fd4218d-b510-4b72-818e-d984a27a501b · outbound

This paper cites MuTox: Universal MUltilingual audio-based TOXicity dataset and zero- shot detector,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances MuTox: Universal MUltilingual audio-based TOXicity dataset and zero- shot detector,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.413069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.184864Z digest=sha256:a7008cc686b735ce2fe2deb73d721e6632e550c1a23a0010dc0d23bb3140f59a

Observation 9553aa5d-1e7e-4bad-88fd-00d5589a65ba · outbound

This paper cites Lightweight Toxicity Detection in Spoken Language: A Transformer-based Approach for Edge Devices.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Lightweight Toxicity Detection in Spoken Language: A Transformer-based Approach for Edge Devices

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.279574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.187654Z digest=sha256:7786363be05c5d5c264c29347b22681ec6f10b30941340762307080390578a6a

Observation 765ba09d-117d-4692-b12e-a11afd668e06 · outbound

This paper cites Enhancing multilingual voice toxicity detection with speech-text alignment,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Enhancing multilingual voice toxicity detection with speech-text alignment,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.403902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.190570Z digest=sha256:d198d3748699b7a77cfc41f969a5dabbaa3714aef67f31146b5cff088aa39d9b

Observation fc62ed87-eb80-4177-a043-365fc6870659 · outbound

This paper cites an unresolved cited work.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-08-07T12:05:03.250183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.193220Z digest=sha256:1935409c3f57cfb011010c5952ca1034d8250c36023dbe4c1bcbbbdd11917dcb

Observation 6ca94240-3683-4e59-8bc6-6983bf466a1a · outbound

This paper cites A large-scale dataset for hate speech detection on vietnamese social media texts,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances A large-scale dataset for hate speech detection on vietnamese social media texts,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.395573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.195986Z digest=sha256:f8fc02ee4a783c116e87e0d11d547374a35be2fe9895db9a84040df8f4a72abe

Observation 9fe3260f-4bb9-443e-a030-6c128f2efc2e · outbound

This paper cites ViHOS: Hate speech spans detection for Vietnamese,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViHOS: Hate speech spans detection for Vietnamese,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.385999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.199708Z digest=sha256:6505debeecdb146c332cfba644ceb5c8aec78b1119b3a18abdd2239205d5371b

Observation fe09ce50-4340-4f3f-805d-84bcabb35fa1 · outbound

This paper cites ViHateT5: Enhancing hate speech detection in Vietnamese with a unified text-to-text transformer model,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViHateT5: Enhancing hate speech detection in Vietnamese with a unified text-to-text transformer model,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.377160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.202438Z digest=sha256:ddc5898a11063c105424fcf29b7f855ce67c321a32c245f74587f9229e99dedd

Observation 76572e13-54fc-46e9-905c-8d283365327d · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Robust speech recognition via large-scale weak su- pervision,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.368473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.205503Z digest=sha256:79cb0b68a750947a72b967b9b4b2d51fa7d7f39335ada75d1909a2bc8d138960

Observation d1520d0c-5e29-440e-b38f-c494d8c52efe · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances wav2vec 2.0: A framework for self-supervised learning of speech repre- sentations,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:03.208155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:03.208155Z digest=sha256:2a7dbe8568bd92a80b82535744f0b6d57e1d75b3d9947bf5f33d095e628e99ad

Observation d0ae4b93-1e95-4d97-b11f-4d4e3186235c · outbound

This paper cites PhoWhisper: Auto- matic Speech Recognition for Vietnamese,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances PhoWhisper: Auto- matic Speech Recognition for Vietnamese,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.354799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.210818Z digest=sha256:74cd9156c764d7e3922fb7e2cb98559e9ad4877d9a1d8073f33b094947d043b6

Observation c379f1c8-4b2f-44f3-8799-0cfedcdfaad3 · outbound

This paper cites Unsupervised cross-lingual representation learning at scale,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances Unsupervised cross-lingual representation learning at scale,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.345270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.213464Z digest=sha256:e6f741693306ecd2c00316a7d547592a7b5593bff024028d484badc5e125f42d

Observation 88481b36-ea3f-4be8-919b-74e849601a30 · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.337261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.216252Z digest=sha256:fa959be82906b73c7016cb506a0df7676536d3293819e5a8f949ca9a0bfaefb1

Observation 4328bd90-4d4f-4dfc-a110-81e6ce1e0c46 · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:03.219241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:03.219241Z digest=sha256:ba3887c7af8c60caac2ad58ceb6323d77c34cdae12e367789ec5cd58bae8fe7d

Observation 527822ff-6376-4ce3-a254-99cf9691e5a9 · outbound

This paper cites PhoBERT: Pre-trained language models for Vietnamese,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances PhoBERT: Pre-trained language models for Vietnamese,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.328697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.222062Z digest=sha256:0e149250905f9251e77f5a157b3dae28c3e154facce8d0ad61c32b7a6d2f3a49

Observation b2116d6d-accf-4f4b-be0e-20c920601f98 · outbound

This paper cites ViSoBERT: A pre-trained language model for Vietnamese social media text processing,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViSoBERT: A pre-trained language model for Vietnamese social media text processing,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.319492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.224863Z digest=sha256:c98f1a872ddad839b074b35ad36642ab6503231d43965f546516dcff4410edcb

Observation ad2e927f-9ba8-4197-a31c-522f3c8784d5 · outbound

This paper cites VLUE: A new benchmark and multi-task knowledge transfer learning for Vietnamese natural language understanding,.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances VLUE: A new benchmark and multi-task knowledge transfer learning for Vietnamese natural language understanding,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:05:03.311260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.227528Z digest=sha256:97459834958844622a64a4ee9861b87ef9e85e887f2005dc1f4e2a3ab3b1f2ff

Pith citing papers

Observation 8a1891f3-1f4c-4d7d-881a-20ccc08b0562 · inbound

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances cites this paper.

ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:05:03.301955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:05:03.141073Z digest=sha256:ad8e49c1a9fdfcf6faa0b60762ff22f4908df001b964d5795fe39e0bbee2146f