REVIEW 2 cited by
ZeroShotDataAug: Generating and Augmenting Training Data with ChatGPT
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this paper, we investigate the use of data obtained from prompting a large generative language model, ChatGPT, to generate synthetic training data with the aim of augmenting data in low resource scenarios. We show that with appropriate task-specific ChatGPT prompts, we outperform the most popular existing approaches for such data augmentation. Furthermore, we investigate methodologies for evaluating the similarity of the augmented data generated from ChatGPT with the aim of validating and assessing the quality of the data generated.
Forward citations
Cited by 2 Pith papers
-
Does online sustainability communication shape public discourse? Insights from six years of tenant-housing provider interactions
Tenant replies to housing-association sustainability posts form six discourse types driven mainly by organisational scale, rents and satisfaction, not by post design.
-
CCISolver: End-to-End Detection and Repair of Method-Level Code-Comment Inconsistency
A two-stage detector-plus-LLM-fixer trained on a new, LLM-filtered dataset reports state-of-the-art code-comment inconsistency detection (F1 89.54%) and 18.84% relative GLEU gain in repair.
Discussion (0). Sign in to comment.