Pith. sign in

REVIEW 1 cited by

EventRL: Enhancing Event Extraction with Outcome Supervision for Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.11430 v1 pith:V6VFVM3R submitted 2024-02-18 cs.CL

classification cs.CL
keywords eventeventrlextractionllmsmodelslanguagelargeoutcome
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this study, we present EventRL, a reinforcement learning approach developed to enhance event extraction for large language models (LLMs). EventRL utilizes outcome supervision with specific reward functions to tackle prevalent challenges in LLMs, such as instruction following and hallucination, manifested as the mismatch of event structure and the generation of undefined event types. We evaluate EventRL against existing methods like Few-Shot Prompting (FSP) (based on GPT4) and Supervised Fine-Tuning (SFT) across various LLMs, including GPT-4, LLaMa, and CodeLLaMa models. Our findings show that EventRL significantly outperforms these conventional approaches by improving the performance in identifying and structuring events, particularly in handling novel event types. The study emphasizes the critical role of reward function selection and demonstrates the benefits of incorporating code data for better event extraction. While increasing model size leads to higher accuracy, maintaining the ability to generalize is essential to avoid overfitting.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Reflective Agreement: Combining Self-Mixture of Agents with a Sequence Tagger for Robust Event Extraction

    cs.CL 2025-08 reject novelty 6.0 of 10

    ARIS combines self-mixture-of-agents LLM decoding with a RoBERTa sequence tagger, consensus detection, confidence filtering, and LLM reflection to improve event extraction F1, but its headline SOTA claim is not suppor...

Pith tools