首页> 外文会议>International Conference on Reconfigurable Computing and FPGAs >Thread shadowing: On the effectiveness of error detection at the hardware thread level
【24h】

Thread shadowing: On the effectiveness of error detection at the hardware thread level

机译:线程阴影:关于硬件线程级别的错误检测的有效性

获取原文

摘要

Dynamic thread duplication is a known redundancy technique for multi-cores. Recent research applied this concept to hybrid multi-cores for error detection and introduced thread shadowing that runs hardware threads in the reconfigurable cores and compares their outputs for deviation at configurable signature levels. Previously published work evaluated this concept in terms of performance, error detection latency and resource consumption. In this paper we report on the error detection capabilities of thread shadowing by presenting an extensive fault injection campaign. We employ the Xilinx Soft Error Mitigation Controller for fault injection and the Xilinx Essential Bit facility to limit the fault injections to relevant bits in the configuration bitstream. Our findings from fault injection experiments with a sorting benchmark are threefold: First, up to 98% of all errors are detected by the operating system of the hybrid multi-core supported by thread shadowing. Second, thread shadowing's signature levels provide a useful trade-off between detected errors and effort needed, with around 5% of all errors detected in calls to operating system functions and around 52% of errors detected in memory accesses of the hardware thread. Third, essential bit testing is effective and cuts down the amount of bits to be tested by a factor of 14.48 compared to the total amount of bits available in the configuration address space.
机译:动态线程复制是一种已知的多核冗余技术。最近的研究将此概念应用于混合多核以进行错误检测,并引入了线程重影,该影子在可重配置内核中运行硬件线程,并在可配置签名级别比较其输出的偏差。先前发表的工作在性能,错误检测延迟和资源消耗方面评估了此概念。在本文中,我们通过提出广泛的故障注入活动来报告线程阴影的错误检测功能。我们使用Xilinx软错误缓解控制器进行故障注入,并使用Xilinx基本位工具将故障注入限制在配置位流中的相关位。我们从具有排序基准的故障注入实验中得出的结论有三方面:第一,线程阴影支持的混合多核操作系统最多可检测到98%的错误。其次,线程阴影的签名级别在检测到的错误和所需的工作量之间提供了一种有用的折衷,其中大约5%的错误是在对操作系统函数的调用中检测到的,而大约52%的错误是在硬件线程的内存访问中检测到的。第三,基本位测试是有效的,并且与配置地址空间中可用的位总数相比,将要测试的位数量减少了14.48倍。

著录项

相似文献

  • 外文文献
  • 中文文献
  • 专利
获取原文

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号