Building an ExaByte-level Data Lake Using Apache Hudi at ByteDanceSeptember 1, 2021 by Ziyue Guan, translated to English by yihuadata lakehouseperformance
Schema evolution with DeltaStreamer using KafkaSourceAugust 16, 2021 by sbernauerhudi streamerschemaapache kafka
Cost-Efficient Open Source Big Data Platform at UberAugust 11, 2021 by Zheng Shao and Mohammad Islamperformancedata platformincremental processinguber
MLOps Wars: Versioned Feature Data with a LakehouseAugust 3, 2021 by David Bzhalava and Jim Dowlingmlopsaiincremental processingquerying
Baixin bank's real-time data lake evolution scheme based on Apache HudiJuly 26, 2021 by Developpaperdata lakehouseincremental processing
Part1: Query apache hudi dataset in an amazon S3 data lake with amazon athena : Read optimized queriesJuly 16, 2021 by Dhiraj Thakur, Sameer Goel and Imtiaz Sayedqueryingaws