Pith. sign in

REVIEW 1 cited by

Fine-tuning Large Language Models with Sequential Instructions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.07794 v3 pith:OCOJW6WF submitted 2024-03-12 cs.CL

classification cs.CL
keywords instructionssequentialtaskscomplexfine-tuninginstructionmodelstuning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Despite the success of existing instruction-tuned models, we find that they usually struggle to respond to queries with multiple instructions. This impairs their performance in complex problems whose solution consists of multiple intermediate tasks. Thus, we contend that part of the fine-tuning data mixture should be sequential--containing a chain of interrelated tasks. We first approach sequential instruction tuning from a task-driven perspective, manually creating interpretable intermediate tasks for multilingual and visual question answering: namely "translate then predict" and "caption then answer". Next, we automate this process by turning instructions in existing datasets (e.g., Alpaca and FlanCoT) into diverse and complex sequential instructions, making our method general-purpose. Models that underwent our sequential instruction tuning show improved results in coding, maths, and open-ended generation. Moreover, we put forward a new benchmark named SeqEval to evaluate a model's ability to follow all the instructions in a sequence, which further corroborates the benefits of our fine-tuning method. We hope that our endeavours will open new research avenues on instruction tuning for complex tasks.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. When Does Language Transfer Help? Sequential Fine-Tuning for Cross-Lingual Euphemism Detection

    cs.CL 2025-08 conditional novelty 5.0 of 10

    Sequential fine-tuning improves target-language euphemism detection in some pairs and models, but the effect is pair-dependent and often no larger than the run-to-run noise.

Pith tools