Pith. sign in

REVIEW 4 cited by

How much data is needed to train a medical image deep learning system to achieve necessary high accuracy?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1511.06348 v2 pith:Z72K3F2P submitted 2015-11-19 cs.LG cs.CVcs.NE

classification cs.LGcs.CVcs.NE
keywords classificationaccuracydataimageimagesmedicalsystemsystems
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The use of Convolutional Neural Networks (CNN) in natural image classification systems has produced very impressive results. Combined with the inherent nature of medical images that make them ideal for deep-learning, further application of such systems to medical image classification holds much promise. However, the usefulness and potential impact of such a system can be completely negated if it does not reach a target accuracy. In this paper, we present a study on determining the optimum size of the training data set necessary to achieve high classification accuracy with low variance in medical image classification systems. The CNN was applied to classify axial Computed Tomography (CT) images into six anatomical classes. We trained the CNN using six different sizes of training data set (5, 10, 20, 50, 100, and 200) and then tested the resulting system with a total of 6000 CT images. All images were acquired from the Massachusetts General Hospital (MGH) Picture Archiving and Communication System (PACS). Using this data, we employ the learning curve approach to predict classification accuracy at a given training sample size. Our research will present a general methodology for determining the training data set size necessary to achieve a certain target classification accuracy that can be easily applied to other problems within such systems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Bayesian Neural Scaling Law Extrapolation with Prior-Data Fitted Networks

    cs.LG 2025-05 conditional novelty 6.0 of 10

    A Prior-data Fitted Network with a scaling-law-specific prior gives better point and uncertainty predictions for neural scaling law extrapolation than MCMC, BNSL, and LC-PFN baselines.

  2. Scaling Pre-training to One Hundred Billion Data for Vision Language Models

    cs.CV 2025-02 conditional novelty 6.0 of 10

    Scaling VLM pretraining from 10B to 100B image-text pairs yields saturation on standard benchmarks but large gains on cultural diversity, low-resource language retrieval, and subgroup disparity.

  3. Recursive Inference Scaling: A Winning Path to Scalable Inference in Language and Multimodal Systems

    cs.AI 2025-02 conditional novelty 6.0 of 10

    Recursively applying the first half of a transformer before the second half (RINS) improves language modeling and vision-language accuracy under compute-matched comparisons.

  4. Analysing User Reviews to Identify User Concerns Around Permissions in AI Apps

    cs.LG 2026-07 conditional novelty 4.0 of 10

    Permission-related AI-app reviews can be classified with reported 82% accuracy using GPT-4-simulated reviews as training templates, and the concerns users raise cluster by sentiment rather than by permission type.

Pith tools