Characterizing and Improving Resilience of Accelerators to Memory Errors in Autonomous Robots

Author:

Shah Deval1,Xue Zi Yu2,Pattabiraman Karthik1,Aamodt Tor M.1

Affiliation:

1. University of British Columbia, Canada

2. Massachusetts Institute of Technology, USA

Abstract

Motion planning is a computationally intensive and well-studied problem in autonomous robots. However, motion planning hardware accelerators (MPA) must be soft-error resilient for deployment in safety-critical applications, and blanket application of traditional mitigation techniques is ill-suited due to cost, power, and performance overheads. We propose Collision Exposure Factor (CEF), a novel metric to assess the failure vulnerability of circuits processing spatial relationships, including motion planning. CEF is based on the insight that the safety violation probability increases with the surface area of the physical space exposed by a bit-flip. We evaluate CEF on four MPAs. We demonstrate empirically that CEF is correlated with safety violation probability, and that CEF-aware selective error mitigation provides 12.3 ×, 9.6 ×, and 4.2 × lower dangerous Failures-In-Time rate on average for the same amount of protected memory compared to uniform, bit-position, and access-frequency-aware selection of critical data. Furthermore, we show how to employ CEF to enable fault characterization using 23, 000 × fewer fault injection (FI) experiments than exhaustive FI, and evaluate our FI approach on different robots and MPAs. We demonstrate that CEF-aware FI can provide insights on vulnerable bits in an MPA while taking the same amount of time as uniform statistical FI. Finally, we use the CEF to formulate guidelines for designing soft-error resilient MPAs.

Publisher

Association for Computing Machinery (ACM)

Subject

Artificial Intelligence,Control and Optimization,Computer Networks and Communications,Hardware and Architecture,Human-Computer Interaction

Reference148 articles.

1. 2020. Design Compiler. https://www.synopsys.com/implementation-and-signoff/rtl-synthesis-test/dc-ultra.html. 2020. Design Compiler. https://www.synopsys.com/implementation-and-signoff/rtl-synthesis-test/dc-ultra.html.

2. 2022. 2022 International Conference on Robotics and Automation, ICRA 2022 , Philadelphia, PA, USA , May 23-27, 2022 . IEEE. https://doi.org/10.1109/ICRA46639.2022 10.1109/ICRA46639.2022 2022. 2022 International Conference on Robotics and Automation, ICRA 2022, Philadelphia, PA, USA, May 23-27, 2022. IEEE. https://doi.org/10.1109/ICRA46639.2022

3. 2022. IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2022 , Kyoto, Japan , October 23-27, 2022 . IEEE. https://doi.org/10.1109/IROS47612.2022 10.1109/IROS47612.2022 2022. IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2022, Kyoto, Japan, October 23-27, 2022. IEEE. https://doi.org/10.1109/IROS47612.2022

4. Evan Ackerman. 2019. Jaco Is a Low-Power Robot Arm That Hooks to Your Wheelchair. https://spectrum.ieee.org/the-human-os/robotics/medical-robots/robot-arm-helps-disabled-11-year-old-girl-show-horse-in-competition. Evan Ackerman. 2019. Jaco Is a Low-Power Robot Arm That Hooks to Your Wheelchair. https://spectrum.ieee.org/the-human-os/robotics/medical-robots/robot-arm-helps-disabled-11-year-old-girl-show-horse-in-competition.

5. John Amanatides and Kin Choi . 1995 . Ray Tracing Triangular Meshes . Proceedings of Western Computer Graphics Symposium. John Amanatides and Kin Choi. 1995. Ray Tracing Triangular Meshes. Proceedings of Western Computer Graphics Symposium.

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3