Pith. sign in

REVIEW 2 cited by

Diverse Demonstrations Improve In-context Compositional Generalization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2212.06800 v3 pith:QQFNGTWQ submitted 2022-12-13 cs.CL

classification cs.CL
keywords demonstrationsin-contextcompositionaldiversegeneralizationlearningsetupsimilar
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In-context learning has shown great success in i.i.d semantic parsing splits, where the training and test sets are drawn from the same distribution. In this setup, models are typically prompted with demonstrations that are similar to the input utterance. However, in the setup of compositional generalization, where models are tested on outputs with structures that are absent from the training set, selecting similar demonstrations is insufficient, as often no example will be similar enough to the input. In this work, we propose a method to select diverse demonstrations that aims to collectively cover all of the structures required in the output program, in order to encourage the model to generalize to new structures from these demonstrations. We empirically show that combining diverse demonstrations with in-context learning substantially improves performance across three compositional generalization semantic parsing datasets in the pure in-context learning setup and when combined with finetuning.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MarginSel : Max-Margin Demonstration Selection for LLMs

    cs.LG 2025-06 conditional novelty 6.0 of 10

    MarginSel picks demonstrations by matching the LLM's own ambiguous candidate-label sets between training and test instances, and reports consistent few-shot F1 gains over random and kNN baselines.

  2. Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis

    cs.CL 2025-07 conditional novelty 5.0 of 10

    For Flamingo-style models, increasing the number of in-context examples improves language coherence but degrades visual-text alignment, and similarity-based image retrieval inflates CIDEr scores by encouraging caption...

Pith tools