
6/10/2019 · Maria Basmanova, Ying Su, Tim Meehan, Orri Erling
What this post added
This post details the Aria Presto initiative to improve PrestoDB efficiency, focusing on optimizing table scans for ORC format. Key technical contributions include subfield pruning for complex data types, adaptive filter ordering to reduce CPU cycles, and efficient row skipping to minimize data reading. The new architecture shifts filter evaluation to the Hive connector and introduces stream readers and record readers to manage filtering and column scanning dynamically. These optimizations aim for a 2-3x decrease in CPU time for Hive queries.