REVIEW 6 cited by
Adaptable and Reliable Text Classification using Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Text classification is fundamental in Natural Language Processing (NLP), and the advent of Large Language Models (LLMs) has revolutionized the field. This paper introduces an adaptable and reliable text classification paradigm, which leverages LLMs as the core component to address text classification tasks. Our system simplifies the traditional text classification workflows, reducing the need for extensive preprocessing and domain-specific expertise to deliver adaptable and reliable text classification results. We evaluated the performance of several LLMs, machine learning algorithms, and neural network-based architectures on four diverse datasets. Results demonstrate that certain LLMs surpass traditional methods in sentiment analysis, spam SMS detection, and multi-label classification. Furthermore, it is shown that the system's performance can be further enhanced through few-shot or fine-tuning strategies, making the fine-tuned model the top performer across all datasets. Source code and datasets are available in this GitHub repository: https://github.com/yeyimilk/llm-zero-shot-classifiers.
Forward citations
Cited by 6 Pith papers
-
A Multi-Dimensional Evaluation of Explainability in Media Bias Detection
In media bias detection, explanation plausibility and mechanistic faithfulness are distinct axes that vary independently across model architectures and finetuning strategies.
-
SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use
A new cross-lingual benchmark shows large language models comply with explicit requests to use swear words far more often in Indic languages than in English, revealing a safety alignment gap.
-
AdaPhish: AI-Powered Adaptive Defense and Education Resource Against Deceptive Emails
An LLM-based phish bowl that automatically anonymizes reported phishing emails and combines nearest-neighbor retrieval with GPT-4o classification to detect and track new phishing campaigns.
-
Potential and Perils of Large Language Models as Judges of Unstructured Textual Data
LLM judges show only fair-to-moderate agreement with human raters on thematic summary alignment and consistently over-rate alignment compared to humans.
-
Innovative Sentiment Analysis and Prediction of Stock Price Using FinBERT, GPT-4 and Logistic Regression: A Data-Driven Approach
On Nigerian financial news from 2010 to 2024, logistic regression with TF-IDF outperformed FinBERT and a predefined GPT-4 approach for predicting NGX index direction, with 81.83% accuracy.
-
Balancing Accuracy and Efficiency in Multi-Turn Intent Classification for LLM-Powered Dialog Systems in Production
Compressing intent labels for LLM fine-tuning and using self-consistency-filtered LLM pseudo-labeling improve multi-turn intent classification accuracy and enable small, low-latency production models.
Discussion (0). Continue with ORCID to comment.