Skip to main navigation Skip to search Skip to main content

Database Native Model Selection: Harnessing Deep Neural Networks in Database Systems

  • Naili Xing
  • , Shaofeng Cai
  • , Gang Chen
  • , Zhaojing Luo
  • , Beng Chin Ooi
  • , Jian Pei
  • National University of Singapore
  • Zhejiang University
  • Duke University

Research output: Contribution to journalConference articlepeer-review

Abstract

The growing demand for advanced analytics beyond statistical aggregation calls for database systems that support effective model selection of deep neural networks (DNNs). However, existing model selection strategies are based on either training-based algorithms that deliver high-performing models at the expense of high computational cost, or training-free algorithms that enhance computational efficiency with reduced effectiveness. These strategies often disregard computational cost and response time Service-Level Objectives (SLOs), which are of concern to average or budgetconscious machine learning users. In addition, they lack a welldesigned integration of the model selection algorithms with DBMSs, which hinders efficient in-database model selection. This paper presents TRAILS, a resource-efficient and SLO-aware in-database model selection system. To leverage the strengths of both trainingfree and training-based model selection, we first characterize nine state-of-the-art training-free model evaluation metrics and propose a more effective one named JacFlow, and then, restructure the conventional model selection procedure into two phases: filtering and refinement. A novel coordinator is also introduced to strike a balance between the high efficiency of train-free algorithms and the high effectiveness of training-based algorithms, ensuring high-performing model selection while adhering to target SLOs. Moreover, we incorporate the proposed algorithm into PostgreSQL to develop TRAILS, thereby both enhancing resource efficiency and reducing model selection latency. This integration establishes a foundation for declarative model definition and selection within DBMSs. Empirical results demonstrate that our TRAILS reduces model selection time and computational expenses considerably by up to 24.38x and 29.32x respectively compared to existing model selection systems.

Original languageEnglish
Pages (from-to)1020-1033
Number of pages14
JournalProceedings of the VLDB Endowment
Volume17
Issue number5
DOIs
Publication statusPublished - 2024
Externally publishedYes
Event50th International Conference on Very Large Data Bases, VLDB 2024 - Guangzhou, China
Duration: 24 Aug 202429 Aug 2024

Fingerprint

Dive into the research topics of 'Database Native Model Selection: Harnessing Deep Neural Networks in Database Systems'. Together they form a unique fingerprint.

Cite this