The Boston Globe
Prof. Dylan Hadfield-Menell speaks to The Boston Globe’s Joshua Miller about his concerns with the current state of AI systems and how they are optimized. “The parts [of AI systems] that are imitating the ways that humans use language, those parts of the systems actually are quite correctable. They’re still pretty unpredictable, but they do seem quite flexible,” says Hadfield-Menell. “On the other hand, there’s reinforcement learning, task-focused behavior, that can often be quite sticky. They [AI systems] seem to really want to push towards task completion.” He adds: “If you have a powerful optimization-driven system pointed at some goal, it can often do a lot of unexpected things that can have a lot of impact.”