Pith. sign in

Citation notice #8235 · 2026-08-03 06:19:12.027079+00:00

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback

Correction Crossref Open

cites Machine Learning , author =, which carries a correction notice dated 2022-12-28. One-hop deterministic notice: the citation edge exists in the Pith bibliography graph; no model judged whether the citation was load-bearing.

This is not a judgment on the citing paper.

Citing paper Event page Original DOI Notice DOI File a formal challenge All reference changes

01Evidence

Raw extraction · bibliography line · bibliography index 84

Model-free inverse reinforcement learning with multi-intention, unlabeled, and overlapping demonstrations , volume =. Machine Learning , author =. 2023 , pages =. doi:10.1007/s10994-022-06273-x , abstract =

Parser render (TeX stripped for reading; raw above is the evidence)

Model-free inverse reinforcement learning with multi-intention, unlabeled, and overlapping demonstrations, volume =. Machine Learning, author =. 2023, pages =. doi:10.1007/s10994-022-06273-x, abstract =

02Event

Type
Correction
Source
Crossref
Original DOI
10.1007/s10994-022-06273-x
Notice DOI
10.1007/s10994-022-06298-2
Date
2022-12-28
Title
Correction to: Model-free inverse reinforcement learning with multi-intention, unlabeled, and overlapping demonstrations
Reasons
['Correction']
Work
Machine Learning , author = (2023) Machine Learning

Schema constants (for re-runners): correction · crossref

03Dispute this notice

If this citation does not depend on the flagged claim, or the event is wrong, say so. Disputes are public. For a signed challenge against the paper itself, use the formal challenge form.