How to cultivate a green decision tree without loss of accuracy?

Proceedings of the ACM/IEEE International Symposium on Low Power Electronics and Design Pub Date : 2020-08-10 DOI:10.1145/3370748.3406566

Tseng-Yi Chen, Yuan-Hao Chang, Ming-Chang Yang, Huang-wei Chen

{"title":"How to cultivate a green decision tree without loss of accuracy?","authors":"Tseng-Yi Chen, Yuan-Hao Chang, Ming-Chang Yang, Huang-wei Chen","doi":"10.1145/3370748.3406566","DOIUrl":null,"url":null,"abstract":"Decision tree is the core algorithm of the random forest learning that has been widely applied to classification and regression problems in the machine learning field. For avoiding underfitting, a decision tree algorithm will stop growing its tree model when the model is a fully-grown tree. However, a fully-grown tree will result in an overfitting problem reducing the accuracy of a decision tree. In such a dilemma, some post-pruning strategies have been proposed to reduce the model complexity of the fully-grown decision tree. Nevertheless, such a process is very energy-inefficiency over an non-volatile-memory-based (NVM-based) system because NVM generally have high writing costs (i.e., energy consumption and I/O latency). Such unnecessary data will induce high writing energy consumption and long I/O latency on NVM-based architectures, especially for low-power-oriented embedded systems. In order to establish a green decision tree (i.e., a tree model with minimized construction energy consumption), this study rethinks a pruning algorithm, namely duo-phase pruning framework, which can significantly decrease the energy consumption on the NVM-based computing system without loss of accuracy.","PeriodicalId":116486,"journal":{"name":"Proceedings of the ACM/IEEE International Symposium on Low Power Electronics and Design","volume":"7 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2020-08-10","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"3","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Proceedings of the ACM/IEEE International Symposium on Low Power Electronics and Design","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1145/3370748.3406566","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 3

Abstract

Decision tree is the core algorithm of the random forest learning that has been widely applied to classification and regression problems in the machine learning field. For avoiding underfitting, a decision tree algorithm will stop growing its tree model when the model is a fully-grown tree. However, a fully-grown tree will result in an overfitting problem reducing the accuracy of a decision tree. In such a dilemma, some post-pruning strategies have been proposed to reduce the model complexity of the fully-grown decision tree. Nevertheless, such a process is very energy-inefficiency over an non-volatile-memory-based (NVM-based) system because NVM generally have high writing costs (i.e., energy consumption and I/O latency). Such unnecessary data will induce high writing energy consumption and long I/O latency on NVM-based architectures, especially for low-power-oriented embedded systems. In order to establish a green decision tree (i.e., a tree model with minimized construction energy consumption), this study rethinks a pruning algorithm, namely duo-phase pruning framework, which can significantly decrease the energy consumption on the NVM-based computing system without loss of accuracy.

查看原文本刊更多论文

如何培育绿色决策树而不损失其准确性?

决策树是随机森林学习的核心算法，已广泛应用于机器学习领域的分类和回归问题。为了避免欠拟合，决策树算法将在模型是完全成熟的树时停止生长其树模型。然而，成熟的树会导致过拟合问题，降低决策树的准确性。在这种困境下，人们提出了一些后修剪策略来降低完全生长决策树的模型复杂性。然而，与基于非易失性存储器(NVM)的系统相比，这样的过程非常节能，因为NVM通常具有高写入成本(即能耗和I/O延迟)。在基于nvm的架构上，这些不必要的数据将导致高写入能耗和长I/O延迟，特别是对于面向低功耗的嵌入式系统。为了建立绿色决策树(即建筑能耗最小的树模型)，本研究重新思考了一种剪枝算法，即两阶段剪枝框架，该算法可以在不损失精度的情况下显著降低基于nvm的计算系统的能耗。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Proceedings of the ACM/IEEE International Symposium on Low Power Electronics and Design

自引率

0.00%

发文量