PIQA: Reasoning about Physical Commonsense in Natural Language-Reference-Cited by-同舟云学术

PIQA: Reasoning about Physical Commonsense in Natural Language

Published:2020-04-03 Issue:05 Volume:34 Page:7432-7439
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Bisk Yonatan,Zellers Rowan,Le bras Ronan,Gao Jianfeng,Choi Yejin

Abstract

To apply eyeshadow without a brush, should I use a cotton swab or a toothpick? Questions requiring this kind of physical commonsense pose a challenge to today's natural language understanding systems. While recent pretrained models (such as BERT) have made progress on question answering over more abstract domains – such as news articles and encyclopedia entries, where text is plentiful – in more physical domains, text is inherently limited due to reporting bias. Can AI systems learn to reliably answer physical commonsense questions without experiencing the physical world?In this paper, we introduce the task of physical commonsense reasoning and a corresponding benchmark dataset Physical Interaction: Question Answering or PIQA. Though humans find the dataset easy (95% accuracy), large pretrained models struggle (∼75%). We provide analysis about the dimensions of knowledge that existing models lack, which offers significant opportunities for future research.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 53 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Harnessing the Power of Large Language Models for Automated Code Generation and Verification;Robotics;2024-09-11

2. Human Language to Analog Layout Using GLayout Layout Automation Framework;Proceedings of the 2024 ACM/IEEE International Symposium on Machine Learning for CAD;2024-09-09

3. PrimeNet: A Framework for Commonsense Knowledge Representation and Reasoning Based on Conceptual Primitives;Cognitive Computation;2024-08-30

4. LangBirds: An Agent for Angry Birds using a Large Language Model;2024 IEEE Conference on Games (CoG);2024-08-05

5. Efficient Pretraining and Finetuning of Quantized LLMs with Low-Rank Structure;2024 IEEE 44th International Conference on Distributed Computing Systems (ICDCS);2024-07-23