Computer Science > Databases

arXiv:2305.01237 (cs)

[Submitted on 2 May 2023]

Title:Updatable Learned Indexes Meet Disk-Resident DBMS -- From Evaluations to Design Choices

Authors:Hai Lan, Zhifeng Bao, J. Shane Culpepper, Renata Borovica-Gajic

View PDF

Abstract:Although many updatable learned indexes have been proposed in recent years, whether they can outperform traditional approaches on disk remains unknown. In this study, we revisit and implement four state-of-the-art updatable learned indexes on disk, and compare them against the B+-tree under a wide range of settings. Through our evaluation, we make some key observations: 1) Overall, the B+-tree performs well across a range of workload types and datasets. 2) A learned index could outperform B+-tree or other learned indexes on disk for a specific workload. For example, PGM achieves the best performance in write-only workloads while LIPP significantly outperforms others in lookup-only workloads. We further conduct a detailed performance analysis to reveal the strengths and weaknesses of these learned indexes on disk. Moreover, we summarize the observed common shortcomings in five categories and propose four design principles to guide future design of on-disk, updatable learned indexes: (1) reducing the index's tree height, (2) better data structures to lower operation overheads, (3) improving the efficiency of scan operations, and (4) more efficient storage layout.

Comments:	22 pages
Subjects:	Databases (cs.DB)
Cite as:	arXiv:2305.01237 [cs.DB]
	(or arXiv:2305.01237v1 [cs.DB] for this version)
	https://doi.org/10.48550/arXiv.2305.01237
Related DOI:	https://doi.org/10.1145/3589284

Submission history

From: Hai Lan [view email]
[v1] Tue, 2 May 2023 07:34:58 UTC (3,775 KB)

Computer Science > Databases

Title:Updatable Learned Indexes Meet Disk-Resident DBMS -- From Evaluations to Design Choices

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Databases

Title:Updatable Learned Indexes Meet Disk-Resident DBMS -- From Evaluations to Design Choices

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators