Pith. sign in

REVIEW 3 cited by

Graph Contrastive Learning via Cluster-refined Negative Sampling for Semi-supervised Text Classification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.18130 v1 pith:V3ENEV6L submitted 2024-10-18 cs.LG cs.CL

classification cs.LGcs.CL
keywords textnegativeclassificationgraphclustersclustertextcontrastivelearning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Graph contrastive learning (GCL) has been widely applied to text classification tasks due to its ability to generate self-supervised signals from unlabeled data, thus facilitating model training. However, existing GCL-based text classification methods often suffer from negative sampling bias, where similar nodes are incorrectly paired as negative pairs. This can lead to over-clustering, where instances of the same class are divided into different clusters. To address the over-clustering issue, we propose an innovative GCL-based method of graph contrastive learning via cluster-refined negative sampling for semi-supervised text classification, namely ClusterText. Firstly, we combine the pre-trained model Bert with graph neural networks to learn text representations. Secondly, we introduce a clustering refinement strategy, which clusters the learned text representations to obtain pseudo labels. For each text node, its negative sample set is drawn from different clusters. Additionally, we propose a self-correction mechanism to mitigate the loss of true negative samples caused by clustering inconsistency. By calculating the Euclidean distance between each text node and other nodes within the same cluster, distant nodes are still selected as negative samples. Our proposed ClusterText demonstrates good scalable computing, as it can effectively extract important information from from a large amount of data. Experimental results demonstrate the superiority of ClusterText in text classification tasks.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SDR-GNN: Spectral Domain Reconstruction Graph Neural Network for Incomplete Multimodal Learning in Conversational Emotion Recognition

    cs.CL 2024-11 reject novelty 5.0 of 10

    SDR-GNN is a graph neural network that reconstructs missing multimodal features and labels utterance emotions, with reported gains over prior methods that are inconsistent across datasets.

  2. GroupFace: Imbalanced Age Estimation Based on Multi-hop Attention Graph Convolutional Network and Group-aware Margin Optimization

    cs.CV 2024-12 reject novelty 4.0 of 10

    GroupFace combines a multi-hop attention graph network with a reinforcement-learning margin scheduler for imbalanced face age estimation, reporting modest benchmark gains but with internal inconsistencies in the rewar...

  3. Dynamic Graph Neural ODE Network for Multi-modal Emotion Recognition in Conversation

    cs.CL 2024-12 reject novelty 4.0 of 10

    DGODE combines adaptive mixhop aggregation with a graph ODE for multimodal emotion recognition in conversation, reporting SOTA numbers on IEMOCAP and MELD, but the supporting derivation and experimental reporting are ...

Pith tools