Title resolution pending

Brandon Dang, Martin J Riedl, Matthew Lease · 2018 · arXiv 1804.10999

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

Title metadata for this work has not finished resolving. The hub is built from the citation graph; the title resolver retries DOI and OpenAlex on its next pass.

citation-role summary

background 1

citation-polarity summary

background 1

representative citing papers

Beyond Content Exposure: Systemic Factors Driving Moderators' Mental Health Crisis in Africa

cs.HC · 2026-03-03 · unverdicted · novelty 6.0

African content moderators suffer high psychological distress from systemic labor conditions, with former moderators showing lasting impacts and corporate wellness programs proving ineffective.

Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

cs.CL · 2022-08-23 · accept · novelty 6.0

RLHF-aligned language models show increasing resistance to red teaming with scale up to 52B parameters, unlike prompted or rejection-sampled models, supported by a released dataset of 38,961 attacks.

citing papers explorer

Showing 2 of 2 citing papers.

Beyond Content Exposure: Systemic Factors Driving Moderators' Mental Health Crisis in Africa cs.HC · 2026-03-03 · unverdicted · none · ref 25
African content moderators suffer high psychological distress from systemic labor conditions, with former moderators showing lasting impacts and corporate wellness programs proving ineffective.
Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned cs.CL · 2022-08-23 · accept · none · ref 15
RLHF-aligned language models show increasing resistance to red teaming with scale up to 52B parameters, unlike prompted or rejection-sampled models, supported by a released dataset of 38,961 attacks.

Title resolution pending

citation-role summary

citation-polarity summary

fields

years

verdicts

roles

polarities

representative citing papers

citing papers explorer