Pith. sign in

REVIEW 1 cited by

An Evaluation of Large Language Models in Bioinformatics Research

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.13714 v1 pith:BYVZDXK3 submitted 2024-02-21 q-bio.QM cs.AIcs.LG

classification q-bio.QMcs.AIcs.LG
keywords bioinformaticsllmstasksmodelsresearchlanguagelargepotential
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Large language models (LLMs) such as ChatGPT have gained considerable interest across diverse research communities. Their notable ability for text completion and generation has inaugurated a novel paradigm for language-interfaced problem solving. However, the potential and efficacy of these models in bioinformatics remain incompletely explored. In this work, we study the performance LLMs on a wide spectrum of crucial bioinformatics tasks. These tasks include the identification of potential coding regions, extraction of named entities for genes and proteins, detection of antimicrobial and anti-cancer peptides, molecular optimization, and resolution of educational bioinformatics problems. Our findings indicate that, given appropriate prompts, LLMs like GPT variants can successfully handle most of these tasks. In addition, we provide a thorough analysis of their limitations in the context of complicated bioinformatics tasks. In conclusion, we believe that this work can provide new perspectives and motivate future research in the field of LLMs applications, AI for Science and bioinformatics.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Wafer Map Defect Classification Using Autoencoder-Based Data Augmentation and Convolutional Neural Network

    cs.CV 2024-11 conditional novelty 4.0 of 10

    An autoencoder with Gaussian noise injection in its latent space generates synthetic wafer maps to balance classes, and a CNN trained on the augmented data reports 98.56% accuracy on WM-811K with AUC and AP of 1.0000.

Pith tools