Pith. sign in

REVIEW 1 cited by

What Makes Good Examples for Visual In-Context Learning?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2301.13670 v2 pith:WDWBR5JI submitted 2023-01-31 cs.CV

classification cs.CV
keywords in-contextexampleslearningpromptvisionmodelsperformanceretrieval
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large-scale models trained on broad data have recently become the mainstream architecture in computer vision due to their strong generalization performance. In this paper, the main focus is on an emergent ability in large vision models, known as in-context learning, which allows inference on unseen tasks by conditioning on in-context examples (a.k.a.~prompt) without updating the model parameters. This concept has been well-known in natural language processing but has only been studied very recently for large vision models. We for the first time provide a comprehensive investigation on the impact of in-context examples in computer vision, and find that the performance is highly sensitive to the choice of in-context examples. To overcome the problem, we propose a prompt retrieval framework to automate the selection of in-context examples. Specifically, we present (1) an unsupervised prompt retrieval method based on nearest example search using an off-the-shelf model, and (2) a supervised prompt retrieval method, which trains a neural network to choose examples that directly maximize in-context learning performance. The results demonstrate that our methods can bring non-trivial improvements to visual in-context learning in comparison to the commonly-used random selection.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Segment Any Class (SAC): Multi-Class Few-Shot Semantic Segmentation via Class Region Proposals

    cs.CV 2024-11 conditional novelty 6.0 of 10

    A training-free prompting method that adapts SAM to multi-class few-shot segmentation and reports higher mIoU than trained baselines on COCO-20i as the number of classes grows.

Pith tools