. 2023 Apr 24;35(10):10295-10308.

doi: 10.1109/TKDE.2023.3269592. eCollection 2023 Oct 1.

Pushing ML Predictions Into DBMSs

Matteo Paganelli¹, Paolo Sottovia², Kwanghyun Park³, Matteo Interlandi⁴, Francesco Guerra¹

Affiliations

¹ University of Modena and Reggio Emilia 41121 Modena Italy.
² Huawei Research Munich 80992 München Germany.
³ Yonsei University Seoul 03722 South Korea.
⁴ Microsoft Research Redmond WA 98052 USA.

PMID: 37954972
PMCID: PMC10620958
DOI: 10.1109/TKDE.2023.3269592

Pushing ML Predictions Into DBMSs

Matteo Paganelli et al. IEEE Trans Knowl Data Eng. 2023.

. 2023 Apr 24;35(10):10295-10308.

doi: 10.1109/TKDE.2023.3269592. eCollection 2023 Oct 1.

Authors

Matteo Paganelli¹, Paolo Sottovia², Kwanghyun Park³, Matteo Interlandi⁴, Francesco Guerra¹

Affiliations

¹ University of Modena and Reggio Emilia 41121 Modena Italy.
² Huawei Research Munich 80992 München Germany.
³ Yonsei University Seoul 03722 South Korea.
⁴ Microsoft Research Redmond WA 98052 USA.

PMID: 37954972
PMCID: PMC10620958
DOI: 10.1109/TKDE.2023.3269592

Abstract

In the past decade, many approaches have been suggested to execute ML workloads on a DBMS. However, most of them have looked at in-DBMS ML from a training perspective, whereas ML inference has been largely overlooked. We think that this is an important gap to fill for two main reasons: (1) in the near future, every application will be infused with some sort of ML capability; (2) behind every web page, application, and enterprise there is a DBMS, whereby in-DBMS inference is an appealing solution both for efficiency (e.g., less data movement), performance (e.g., cross-optimizations between relational operators and ML) and governance. In this article, we study whether DBMSs are a good fit for prediction serving. We introduce a technique for translating trained ML pipelines containing both featurizers (e.g., one-hot encoding) and models (e.g., linear and tree-based models) into SQL queries, and we compare in-DBMS performance against popular ML frameworks such as Sklearn and ml.net. Our experiments show that, when pushed inside a DBMS, trained ML pipelines can have performance comparable to ML frameworks in several scenarios, while they perform quite poorly on text featurization and over (even simple) neural networks.

Keywords: MLOPs; SQL; machine learning.

PubMed Disclaimer

Figures

**Fig. 1.**
A typical ML workflow. Rectangles are used to identify data artifacts (e.g., input data, or trained models); ellipses determine computations (e.g., data preparation and serving).

**Fig. 2.**
$MASQ$ applied to an ML predictive pipeline.

**Fig. 3.**
Parsing of the pipeline of Example 1. The pipeline (top) is parsed on a container DAG (bottom). Each container stores a reference to the operator, its signature and extractor.

**Fig. 4.**
Scaling and linear model in SQL.

**Fig. 6.**
SQL workflow for the one-hot encoding sparse implementation.

**Fig. 7.**
Pipeline with OHE followed by a linear regression executed in $MASQ$ with TRO and partitioning.

**Fig. 8.**
How tree ensembles over triplet are translated in $MASQ$ .

**Fig. 9.**
Throughput for Sklearn, ml.net (on CSV, MySQL and SQL Server) and $MASQ$ (on MySQL and SQL Server).

**Fig. 10.**
Scalability of the different frameworks, over MySQL, as we change the batch size.

**Fig. 11.**
Latency numbers over a single record (MySQL).

**Fig. 12.**
Operator breakdown for P7 and P8.

**Fig. 13.**
Operator breakdown for P9 ( $Criteo$ ) and P10 ( $Flight Delay$ ). For P9 we also compare GBDT vs SDCA.

**Fig. 14.**
Latency breakdown for $MASQ$ , ml.net (ML.) and Sklearn (SK.), for pipelines P7, P8, P9 and P10. The time spent is divided into three buckets: load, computation, and write.

**Fig. 15.**
Performance comparison with indexing (MySQL).

**Fig. 16.**
Operator fusion (OHE + GBDT) for P9 and P10.

**Fig. 17.**
Comparison of different tree implementation methods (MySQL).

**Fig. 18.**
Comparison of single intermediate data and multi-way join strategy for OHE + linear models.

**Fig. 19.**
Comparison of tree ensembles performance with variable number of leaves on P8.

**Fig. 20.**
Left hand-side: Sentiment Analysis over textual features. Right hand-side: an MLP model applied on $Credit Card$ .

See this image and copyright information in PMC

References

1. Dean J., “The deep learning revolution and its implications for computer architecture and chip design,” Nov. 2019, arXiv:1911.05289.
1. Ahmed Z. et al., “Machine learning at microsoft with ML.NET,” in Proc. ACM SIGKDD Int. Conf. Knowl. Discov. Data Mining, 2019, pp. 2448–2458.
1. Abadi M. et al., “TensorFlow: A system for large-scale machine learning,” in Proc. 12th USENIX Symp. Operating Syst. Des. Implementation, 2016, pp. 265–283. [Online]. Available: https://www.usenix.org/system/files/conference/osdi16/osdi16-abadi.pdf
1. Paszke A., “PyTorch: An Imperative Style, High-Performance Deep Learning Library,” Dec. 2019, arXiv:1912.01703.
1. Sergeev A. and Del Balso M., “Horovod: Fast and easy distributed deep learning in TensorFlow,” Feb. 2018, arXiv:1802.05799.

LinkOut - more resources

Full Text Sources
- Europe PubMed Central
- PubMed Central
Miscellaneous
- NCI CPTAC Assay Portal

Save citation to file

Email citation

Add to Collections

Add to My Bibliography

Your saved search

Create a file for external citation management software

Your RSS Feed

Pushing ML Predictions Into DBMSs

Affiliations

Pushing ML Predictions Into DBMSs

Authors

Affiliations

Abstract

Figures

References

LinkOut - more resources

Full Text Sources

Miscellaneous