Demo: Resource Allocation for Wafer-Scale Deep Learning Accelerator

2021 IEEE 41st International Conference on Distributed Computing Systems (ICDCS) Pub Date : 2021-07-01 DOI:10.1109/ICDCS51616.2021.00114

Huihong Peng, Longkun Guo, Long Sun, Xiaoyan Zhang

引用次数: 1

Abstract

Due to the rapid development of deep learning (DL) has brought, artificial intelligence (AI) chips were invented incorperating the traditional computing architecture with the simulated neural network structure for the sake of improving the energy efficiency. Recently, emerging deep learning AI chips imposed the challenge of allocating computing resources according to a deep neural networks (DNN), such that tasks using the DNN can be processed in a parallel and distributed manner. In this paper, we combine graph theory and combinatorial optimization technology to devise a fast floorplanning approach based on kernel graph structure, which is provided by Cerebras Systems Inc. for mapping the layers of DNN to the mesh of computing units called Wafer-Scale-Engine (WSE). Numerical experiments were carried out to evaluate our method using the public benchmarks and evaluation criteria, demonstrating its performance gain comparing to the state-of-art algorithms.

查看原文本刊更多论文

随着深度学习技术的飞速发展，为了提高能效，人们发明了将传统计算架构与模拟神经网络结构相结合的人工智能芯片。最近，新兴的深度学习人工智能芯片提出了根据深度神经网络(DNN)分配计算资源的挑战，使得使用DNN的任务可以以并行和分布式的方式处理。本文将图论与组合优化技术相结合，设计了一种基于核图结构的快速布局方法，该方法由Cerebras Systems公司提供，用于将深度神经网络层映射到称为晶圆规模引擎(WSE)的计算单元网格上。数值实验使用公共基准和评估标准来评估我们的方法，与最先进的算法相比，展示了它的性能增益。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

2021 IEEE 41st International Conference on Distributed Computing Systems (ICDCS)

自引率

0.00%

发文量