Pith. sign in

REVIEW 2 cited by

Text Classification: Neural Networks VS Machine Learning Models VS Pre-trained Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.21022 v1 pith:CNERX2K6 submitted 2024-12-30 cs.LG

classification cs.LG
keywords modelsnetworksneurallearningmachineclassificationpre-trainedstandard
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Text classification is a very common task nowadays and there are many efficient methods and algorithms that we can employ to accomplish it. Transformers have revolutionized the field of deep learning, particularly in Natural Language Processing (NLP) and have rapidly expanded to other domains such as computer vision, time-series analysis and more. The transformer model was firstly introduced in the context of machine translation and its architecture relies on self-attention mechanisms to capture complex relationships within data sequences. It is able to handle long-range dependencies more effectively than traditional neural networks (such as Recurrent Neural Networks and Multilayer Perceptrons). In this work, we present a comparison between different techniques to perform text classification. We take into consideration seven pre-trained models, three standard neural networks and three machine learning models. For standard neural networks and machine learning models we also compare two embedding techniques: TF-IDF and GloVe, with the latter consistently outperforming the former. Finally, we demonstrate the results from our experiments where pre-trained models such as BERT and DistilBERT always perform better than standard models/algorithms.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Extracting Overlapping Microservices from Monolithic Code via Deep Semantic Embeddings and Graph Neural Network-Based Soft Clustering

    cs.SE 2025-08 reject novelty 6.0 of 10

    Mo2oM assigns classes to overlapping microservices using UniXcoder embeddings and NOCD soft clustering, claiming large gains in modularity metrics over hard-clustering baselines on four monoliths.

  2. CaresAI at SMM4H-HeaRD 2026: Predicting TNM Staging

    cs.CL 2026-07 conditional novelty 2.0 of 10

    LightGBM with TF-IDF predicts T/N/M stage from TCGA pathology reports with strong internal AUROC but weaker generalization on a second held-out test set.

Pith tools