article

A deep reinforcement learning model with plan value network for join order selection

  • International Journal of Wireless and Mobile Computing
  • Inderscience Publishers
Research footprint

At a glance

Citations
0
References
0
Comments
0
Paper overview

Abstract

The existing optimisers and dynamic programming methods rely on the cardinality estimation and the cost model of the local database. The resulting join plan does not reflect the execution time, and the errors of cardinality estimation will also lead to the join plans with poor quality. We propose a new learning optimiser, called DVJ (Deep reinforcement learning with plan Value networks for Join order selection). Compared with the existing deep reinforcement learning method, DVJ has two improvements: (1) the plan value network is designed to improve the reward mechanism in the existing deep reinforcement learning method, and the join plan generated by DVJ can reflect the latency; (2) applying Deep Q-Network to reduce the optimisation time and increase the chance of finding the best connection plan. Extensive experiments are conducted and the experimental results demonstrate that DVJ outperforms traditional optimisers and the existing deep reinforcement learning methods.

Record transparency

Publication details

DOI
10.1504/ijwmc.2021.121627
OpenAlex
W4255347645
Document type
article
Language
EN
Source
International Journal of Wireless and Mobile Computing
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.