Pith. sign in

The iMet Collection 2019 Challenge Dataset

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Existing computer vision technologies in artwork recognition focus mainly on instance retrieval or coarse-grained attribute classification. In this work, we present a novel dataset for fine-grained artwork attribute recognition. The images in the dataset are professional photographs of classic artworks from the Metropolitan Museum of Art, and annotations are curated and verified by world-class museum experts. In addition, we also present the iMet Collection 2019 Challenge as part of the FGVC6 workshop. Through the competition, we aim to spur the enthusiasm of the fine-grained visual recognition research community and advance the state-of-the-art in digital curation of museum collections.

citation-role summary

dataset 1

citation-polarity summary

fields

cs.CV 1

years

2024 1

verdicts

CONDITIONAL 1

roles

dataset 1

polarities

baseline 1

representative citing papers

Understanding Museum Exhibits using Vision-Language Reasoning

cs.CV · 2024-12-02 · conditional · novelty 6.0

A new 65M-image, 200M-QA dataset for museum exhibits lets fine-tuned vision-language models beat general-purpose VLMs on museum attribute questions, especially on questions requiring background knowledge.

citing papers explorer

Showing 1 of 1 citing paper.

  • Understanding Museum Exhibits using Vision-Language Reasoning cs.CV · 2024-12-02 · conditional · none · ref 88 · internal anchor

    A new 65M-image, 200M-QA dataset for museum exhibits lets fine-tuned vision-language models beat general-purpose VLMs on museum attribute questions, especially on questions requiring background knowledge.