Pith. sign in

REVIEW 1 cited by

Low-Resource Fast Text Classification Based on Intra-Class and Inter-Class Distance Calculation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.09922 v1 pith:UDTDNIMV submitted 2024-12-13 cs.CL

classification cs.CL
keywords classificationmethodstextinformationlow-resourceperformanceprocessingclass
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In recent years, text classification methods based on neural networks and pre-trained models have gained increasing attention and demonstrated excellent performance. However, these methods still have some limitations in practical applications: (1) They typically focus only on the matching similarity between sentences. However, there exists implicit high-value information both within sentences of the same class and across different classes, which is very crucial for classification tasks. (2) Existing methods such as pre-trained language models and graph-based approaches often consume substantial memory for training and text-graph construction. (3) Although some low-resource methods can achieve good performance, they often suffer from excessively long processing times. To address these challenges, we propose a low-resource and fast text classification model called LFTC. Our approach begins by constructing a compressor list for each class to fully mine the regularity information within intra-class data. We then remove redundant information irrelevant to the target classification to reduce processing time. Finally, we compute the similarity distance between text pairs for classification. We evaluate LFTC on 9 publicly available benchmark datasets, and the results demonstrate significant improvements in performance and processing time, especially under limited computational and data resources, highlighting its superior advantages.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Text Distance from Nested and Hierarchical Repetitions: A Compression-Based Perspective

    cs.CL 2026-06 conditional novelty 6.0 of 10

    Ladderpath-derived distances (NCD_lp, L_Dice, L_Jaccard) with k-NN outperform gzip-NCD and BERT on out-of-distribution and few-shot text classification without training.

Pith tools