Pith. sign in

REVIEW 5 cited by

Iteratively Prompt Pre-trained Language Models for Chain of Thought

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2203.08383 v3 pith:WJPAOYCP submitted 2022-03-16 cs.CL

classification cs.CL
keywords iterativeknowledgemulti-stepplmspromptingchaincontext-awarecontexts
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

While Pre-trained Language Models (PLMs) internalize a great amount of world knowledge, they have been shown incapable of recalling these knowledge to solve tasks requiring complex & multi-step reasoning. Similar to how humans develop a "chain of thought" for these tasks, how can we equip PLMs with such abilities? In this work, we explore an iterative prompting framework, a new prompting paradigm which progressively elicits relevant knowledge from PLMs for multi-step inference. We identify key limitations of existing prompting methods, namely they are either restricted to queries with a single identifiable relation/predicate, or being agnostic to input contexts, which makes it difficult to capture variabilities across different inference steps. We propose an iterative context-aware prompter, which addresses these limitations by learning to dynamically synthesize prompts conditioned on the current step's contexts. Experiments on three datasets involving multi-step reasoning show the effectiveness of the iterative scheme and the context-aware prompter design.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Derailing Non-Answers via Logit Suppression at Output Subspace Boundaries in RLHF-Aligned Language Models

    cs.CL 2025-05 conditional novelty 6.0 of 10

    Suppressing the double-newline token immediately after <think> substantially increases substantive answers to sensitive prompts in DeepSeek-R1 distillations without training.

  2. eaSEL: Promoting Social-Emotional Learning and Parent-Child Interaction through AI-Mediated Content Consumption

    cs.HC 2025-01 conditional novelty 6.0 of 10

    A system that generates social-emotional learning activities from children's videos increased emotion-word use in 5-8 year olds' story retellings, and parents saw it as helping family conversations.

  3. Optimizing Question Semantic Space for Dynamic Retrieval-Augmented Multi-hop Question Answering

    cs.IR 2025-05 conditional novelty 5.0 of 10

    Q-DREAM improves multi-hop retrieval-augmented QA by decomposing questions, rewriting dependent subquestions, and retrieving with cluster-specific LoRA embeddings.

  4. Leveraging LLM Agents for Automated Optimization Modeling for SASP Problems: A Graph-RAG based Approach

    cs.AI 2025-01 reject novelty 4.0 of 10

    A multi-agent LLM system with graph-based retrieval scores higher than prompt-only baselines on ten SASP modeling tasks, though the evaluation may be biased because the library contains the answers.

  5. QuaLLM-Health: An Adaptation of an LLM-Based Framework for Quantitative Data Extraction from Online Health Discussions

    cs.CL 2024-11 reject novelty 4.0 of 10

    QuaLLM-Health claims GPT-4o-mini can extract clinical variables from GLP-1 Reddit discussions with macro F1 above 0.90, but the evaluation uses the same gold standard for prompt tuning and testing.

Pith tools