Jack Clark
Identifiers
- name variant Jack Clark 0.60 · backfill
Papers (11)
- Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training cs.CR · 2024 · author #18
- Towards Measuring the Representation of Subjective Global Opinions in Language Models cs.CL · 2023 · author #17
- Discovering Language Model Behaviors with Model-Written Evaluations cs.CL · 2022 · author #55
- In-context Learning and Induction Heads cs.LG · 2022 · author #23
- Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned cs.CL · 2022 · author #36
- Language Models (Mostly) Know What They Know cs.CL · 2022 · author #31
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback cs.CL · 2022 · author #27
- A General Language Assistant as a Laboratory for Alignment cs.CL · 2021 · author #19
- Learning Transferable Visual Models From Natural Language Supervision cs.CV · 2021 · author #10
- Language Models are Few-Shot Learners cs.CL · 2020 · author #26
- Release Strategies and the Social Impacts of Language Models cs.CL · 2019 · author #3
Mentions
- 1908.09203 #3 · arxiv_oai · confidence 0.70 Jack Clark
- 2306.16388 #17 · arxiv_oai · confidence 0.70 Jack Clark
Frequent Coauthors
- Amanda Askell 11 shared papers
- Jared Kaplan 9 shared papers
- Deep Ganguli 8 shared papers
- Sam McCandlish 8 shared papers
- Danny Hernandez 7 shared papers
- Dario Amodei 7 shared papers
- Kamal Ndousse 7 shared papers
- Nicholas Joseph 7 shared papers
- Nova DasSarma 7 shared papers
- Tom Henighan 7 shared papers
- Yuntao Bai 7 shared papers
- Zac Hatfield-Dodds 7 shared papers
- Andy Jones 6 shared papers
- Anna Chen 6 shared papers
- Ben Mann 6 shared papers
- Catherine Olsson 6 shared papers
- Dawn Drain 6 shared papers
- Jackson Kernion 6 shared papers
- Liane Lovitt 6 shared papers
- Nelson Elhage 6 shared papers