Media Summary: Dylan Hadfield-Menell is an Assistant Professor at MIT's CSAIL, specializing in Artificial Intelligence and Decision-Making. At an Anthropic Research Salon event in San Francisco, four of our researchers—Alex Tamkin, Jan Leike, Amanda Askell and ... Can an AI do the right thing for the wrong reason? Tim Scarfe speaks with Apollo Research's Alexander Meinke, Axel Højmark ...
Flexible Agent Alignment With Goal - Detailed Analysis & Overview
Dylan Hadfield-Menell is an Assistant Professor at MIT's CSAIL, specializing in Artificial Intelligence and Decision-Making. At an Anthropic Research Salon event in San Francisco, four of our researchers—Alex Tamkin, Jan Leike, Amanda Askell and ... Can an AI do the right thing for the wrong reason? Tim Scarfe speaks with Apollo Research's Alexander Meinke, Axel Højmark ... This video explores how YOU, YES YOU, are a case of misalignment with respect to evolution's implicit optimization Leaders must set vision and strategy and determine what people must do differently to execute on that vision and strategy. Agentic engineering so far has been a solo story: one developer and a dozen
Inspired by Hoenigman, Bradley, and Lim's Dylan Hadfield-Menell - "Preference Learning in