Pith. sign in

REVIEW 1 cited by

Large Language Models as Sous Chefs: Revising Recipes with GPT-3

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2306.13986 v1 pith:GSNQCFME submitted 2023-06-24 cs.CL

classification cs.CL
keywords recipeslanguagelargemodelspromptchefsgpt-3llms
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

With their remarkably improved text generation and prompting capabilities, large language models can adapt existing written information into forms that are easier to use and understand. In our work, we focus on recipes as an example of complex, diverse, and widely used instructions. We develop a prompt grounded in the original recipe and ingredients list that breaks recipes down into simpler steps. We apply this prompt to recipes from various world cuisines, and experiment with several large language models (LLMs), finding best results with GPT-3.5. We also contribute an Amazon Mechanical Turk task that is carefully designed to reduce fatigue while collecting human judgment of the quality of recipe revisions. We find that annotators usually prefer the revision over the original, demonstrating a promising application of LLMs in serving as digital sous chefs for recipes and beyond. We release our prompt, code, and MTurk template for public use.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. On Recipe Memorization and Creativity in Large Language Models: Is Your Model a Creative Cook, a Bad Cook, or Merely a Plagiator?

    cs.CL 2025-06 conditional novelty 5.0 of 10

    Mixtral's recipe ingredients are mostly traceable to online recipes, and an LLM-as-judge pipeline can reproduce human memorization annotations with up to 78 percent accuracy.

Pith tools