Pith. sign in

REVIEW 2 cited by

ZeroShotDataAug: Generating and Augmenting Training Data with ChatGPT

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.14334 v1 pith:MSXPHTO7 submitted 2023-04-27 cs.AI

classification cs.AI
keywords datachatgptaugmentinggeneratedinvestigatetrainingapproachesappropriate
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this paper, we investigate the use of data obtained from prompting a large generative language model, ChatGPT, to generate synthetic training data with the aim of augmenting data in low resource scenarios. We show that with appropriate task-specific ChatGPT prompts, we outperform the most popular existing approaches for such data augmentation. Furthermore, we investigate methodologies for evaluating the similarity of the augmented data generated from ChatGPT with the aim of validating and assessing the quality of the data generated.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Does online sustainability communication shape public discourse? Insights from six years of tenant-housing provider interactions

    cs.CY 2026-07 conditional novelty 6.0 of 10

    Tenant replies to housing-association sustainability posts form six discourse types driven mainly by organisational scale, rents and satisfaction, not by post design.

  2. CCISolver: End-to-End Detection and Repair of Method-Level Code-Comment Inconsistency

    cs.SE 2025-06 conditional novelty 6.0 of 10

    A two-stage detector-plus-LLM-fixer trained on a new, LLM-filtered dataset reports state-of-the-art code-comment inconsistency detection (F1 89.54%) and 18.84% relative GLEU gain in repair.

Pith tools