StrategyQA and Big Bench

Did Aristotle Use a Laptop?  That's a question from the StrategyQA benchmark which highlights the stretch goals for current artificial intelligence systems.  Answering a question like that requires several cognitive steps and reasoning.  Constructing a dataset of similarly challenging questions is a major undertaking.  On today's episode, Mor Geva returns to share details about the creation of StrategyQA and the larger Big Bench dataset it has been included in.

Guest

Mor Geva: Mor Geva is a researcher at the Allen Institute for AI (AI2), working in the field of Natural Language Processing. Her research focuses on developing systems that can reason over text in a robust and interpretable manner. She completed her Ph.D. in Computer Science and B.Sc. in Bioinformatics at Tel Aviv University. During her Ph.D., Mor interned at AI2, Google AI, and Microsoft Media AI. She was awarded the Dan David prize for graduate students in the field of AI, was nominated as one of the MIT Rising Stars in EECS, and is a laureate of the Séphora Berrebi scholarship in Computer Science.

StrategyQA and Big Bench