arrow
Return

LiquidCache: Efficient Pushdown Caching for Cloud-Native Data Analytics

delete2025-09-01
delete0
PRE
AI
X
Xiangpeng Hao *
A
Andrew Lamb
W
Wu, Yibo
A
Andrea C. Arpaci-Dusseau
R
Remzi H. Arpaci-Dusseau
DOI:10.14778/3773731.3773741delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We present LiquidCache, a novel pushdown-based disaggregated caching system that evaluates filters on cache servers before transmitting data to compute nodes. Our key observation is that data decoding, not filter evaluation, is the primary bottleneck in existing systems. To address this challenge, we transcode Parquet data a lightweight Liquid format and cache it for efficient filter evaluation. The Liquid format resides solely in the cache layer, requiring no changes to existing deployments and enabling easy adoption of new encodings without breaking compatibility. Through integration with Apache DataFusion and evaluation with ClickBench and TPC-H, we demonstrate that LiquidCache reduces cache CPU time by up to 10x without increasing memory footprint, and duces network traffic by two orders of magnitudes compared non-pushdown systems.
Keywords:
STORAGE
COMPRESSION

Journal

P
Proceedings of the VLDB Endowment
IF:
3.3
Papers:
556
Citations:
1.2W

Organization

U
university of wisconsin madison
Scholars:
3.8W
Papers: 2.9W
Citations: 53
University of Wisconsin System cover
University of Wisconsin System
Scholars:
6.7W
Papers: 5.8W
Citations: 382