Emergent Deception in LLMs
On today’s show, we are joined by Thilo Hagendorff, a Research Group Leader of Ethics of Generative AI at the University of Stuttgart. He joins us to discuss his research, [Deception Abilities Emerged in Large Language Models](https://arxiv.org/abs/2307.16513). Thilo discussed how machine psychology is useful in machine learning tasks. He shared examples of cognitive tasks that LLMs have improved at solving. He shared his thoughts on whether there’s a ceiling to the tasks ML can solve.
Guest
Thilo Hagendorff: Dr. Thilo Hagendorff is an expert in AI ethics, machine behavior in language models, as well as the intersection of machine learning and psychology. He is working as an Independent Research Group Leader at the University of Stuttgart (Germany). He was a visiting scholar at Stanford University as well as UC San Diego. As a lecturer, he teaches at the Hasso Plattner Institute (Germany), among others. thilo-hagendorff.info