Evaluation of Machine Learning Algorithms in Predicting the Next SQL Query from the Future

ACM Transactions on Database Systems (TODS) Pub Date : 2021-03-18 DOI:10.1145/3442338

Venkata Vamsikrishna Meduri, Kanchan Chowdhury, Mohamed Sarwat

{"title":"Evaluation of Machine Learning Algorithms in Predicting the Next SQL Query from the Future","authors":"Venkata Vamsikrishna Meduri, Kanchan Chowdhury, Mohamed Sarwat","doi":"10.1145/3442338","DOIUrl":null,"url":null,"abstract":"Prediction of the next SQL query from the user, given her sequence of queries until the current timestep, during an ongoing interaction session of the user with the database, can help in speculative query processing and increased interactivity. While existing machine learning-- (ML) based approaches use recommender systems to suggest relevant queries to a user, there has been no exhaustive study on applying temporal predictors to predict the next user issued query. In this work, we experimentally compare ML algorithms in predicting the immediate next future query in an interaction workload, given the current user query or the sequence of queries in a user session thus far. As a part of this, we propose the adaptation of two powerful temporal predictors: (a) Recurrent Neural Networks (RNNs) and (b) a Reinforcement Learning approach called Q-Learning that uses Markov Decision Processes. We represent each query as a comprehensive set of fragment embeddings that not only captures the SQL operators, attributes, and relations but also the arithmetic comparison operators and constants that occur in the query. Our experiments on two real-world datasets show the effectiveness of temporal predictors against the baseline recommender systems in predicting the structural fragments in a query w.r.t. both quality and time. Besides showing that RNNs can be used to synthesize novel queries, we find that exact Q-Learning outperforms RNNs despite predicting the next query entirely from the historical query logs.","PeriodicalId":6983,"journal":{"name":"ACM Transactions on Database Systems (TODS)","volume":"20 1","pages":"1 - 46"},"PeriodicalIF":0.0000,"publicationDate":"2021-03-18","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"3","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"ACM Transactions on Database Systems (TODS)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1145/3442338","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 3

Abstract

Prediction of the next SQL query from the user, given her sequence of queries until the current timestep, during an ongoing interaction session of the user with the database, can help in speculative query processing and increased interactivity. While existing machine learning-- (ML) based approaches use recommender systems to suggest relevant queries to a user, there has been no exhaustive study on applying temporal predictors to predict the next user issued query. In this work, we experimentally compare ML algorithms in predicting the immediate next future query in an interaction workload, given the current user query or the sequence of queries in a user session thus far. As a part of this, we propose the adaptation of two powerful temporal predictors: (a) Recurrent Neural Networks (RNNs) and (b) a Reinforcement Learning approach called Q-Learning that uses Markov Decision Processes. We represent each query as a comprehensive set of fragment embeddings that not only captures the SQL operators, attributes, and relations but also the arithmetic comparison operators and constants that occur in the query. Our experiments on two real-world datasets show the effectiveness of temporal predictors against the baseline recommender systems in predicting the structural fragments in a query w.r.t. both quality and time. Besides showing that RNNs can be used to synthesize novel queries, we find that exact Q-Learning outperforms RNNs despite predicting the next query entirely from the historical query logs.

查看原文本刊更多论文

机器学习算法在预测未来的下一个SQL查询中的评价

在用户与数据库的持续交互会话期间，根据用户在当前时间步之前的查询序列，预测用户的下一个SQL查询，可以帮助推测性查询处理和增强交互性。虽然现有的基于机器学习(ML)的方法使用推荐系统向用户建议相关查询，但尚未对应用时间预测器来预测下一个用户发出的查询进行详尽的研究。在这项工作中，我们通过实验比较了ML算法在交互工作负载中预测下一个未来查询的能力，给定当前用户查询或到目前为止用户会话中的查询序列。作为其中的一部分，我们提出了两种强大的时间预测因子的适应:(a)循环神经网络(rnn)和(b)使用马尔可夫决策过程的强化学习方法，称为q -学习。我们将每个查询表示为一组全面的片段嵌入，这些片段嵌入不仅捕获SQL操作符、属性和关系，还捕获查询中出现的算术比较操作符和常量。我们在两个真实数据集上的实验表明，相对于基线推荐系统，时间预测器在预测查询质量和时间上的结构片段方面是有效的。除了表明rnn可以用于合成新的查询之外，我们发现精确的Q-Learning优于rnn，尽管完全从历史查询日志中预测下一个查询。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

ACM Transactions on Database Systems (TODS)

自引率

0.00%

发文量