Data Warehousing and Analytics Platform
Aria Presto: Making table scan more efficient

Aria Presto: Making table scan more efficient

6/10/2019 · Maria Basmanova, Ying Su, Tim Meehan, Orri Erling

What this post added

This post details the Aria Presto initiative to improve PrestoDB efficiency, focusing on optimizing table scans for ORC format. Key technical contributions include subfield pruning for complex data types, adaptive filter ordering to reduce CPU cycles, and efficient row skipping to minimize data reading. The new architecture shifts filter evaluation to the Hive connector and introduces stream readers and record readers to manage filtering and column scanning dynamically. These optimizations aim for a 2-3x decrease in CPU time for Hive queries.

Read the original post ↗