REVIEW 1 cited by
What the F-measure doesn't measure: Features, Flaws, Fallacies and Fixes
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The F-measure or F-score is one of the most commonly used single number measures in Information Retrieval, Natural Language Processing and Machine Learning, but it is based on a mistake, and the flawed assumptions render it unsuitable for use in most contexts! Fortunately, there are better alternatives.
Forward citations
Cited by 1 Pith paper
-
Multi-Scale Deep Learning for Colon Histopathology: A Hybrid Graph-Transformer Approach
A hybrid CNN-transformer-graph network is reported to reach 96% accuracy on LC25000, but the paper lacks architectural detail, code, and a described data split.
Discussion (0). Continue with ORCID to comment.