Pith. sign in

REVIEW 3 cited by

Learning to Write with Cooperative Discriminators

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1805.06087 v1 pith:RSUBHMWU submitted 2018-05-16 cs.CL

classification cs.CL
keywords modelstextgeneratedlanguagelearningoverallusedamounts
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Recurrent Neural Networks (RNNs) are powerful autoregressive sequence models, but when used to generate natural language their output tends to be overly generic, repetitive, and self-contradictory. We postulate that the objective function optimized by RNN language models, which amounts to the overall perplexity of a text, is not expressive enough to capture the notion of communicative goals described by linguistic principles such as Grice's Maxims. We propose learning a mixture of multiple discriminative models that can be used to complement the RNN generator and guide the decoding process. Human evaluation demonstrates that text generated by our system is preferred over that of baselines by a large margin and significantly enhances the overall coherence, style, and information content of the generated text.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Decision Transformer: Reinforcement Learning via Sequence Modeling

    cs.LG 2021-06 accept novelty 8.0 of 10

    Decision Transformer casts RL as autoregressive sequence modeling conditioned on desired returns, past states and actions, matching or exceeding offline RL baselines on Atari, Gym and Key-to-Door tasks.

  2. Multi-agents based User Values Mining for Recommendation

    cs.IR 2025-05 conditional novelty 6.0 of 10

    A multi-LLM debate framework extracts Schwartz value labels from user interaction histories, and adding these labels through contrastive learning improves recommendation accuracy on PENS and MovieLens-1M.

  3. Reflection-Window Decoding: Text Generation with Selective Refinement

    cs.CL 2025-02 conditional novelty 6.0 of 10

    Selectively refining uncertain windows during decoding improves text quality over greedy and beam search, with a theory formalizing why greedy decoding can miss the joint-probability-optimal response.

Pith tools