Why is vectorized execution mentioned alongside columnar storage?
Columnar data is naturally amenable to vectorization. Modern CPUs operate on vectors of values efficiently (SIMD instructions). With columns in contiguous memory, a CPU can process 8–64 values per instruction cycle. Row storage scatters related data, preventing vectorization. That's why columnar databases are fast—they align data layout with CPU capabilities, not despite the layout.
Answered in
Columnar Storage: Why Column Stores Beat Row Stores for AnalyticsColumnar storage reads only needed columns, skipping the rest. Dictionary encoding shrinks data 50–100×. Analytics queries go from minutes to milliseconds.
Read the full analysisOther questions this article answers
More system design questions
- Why doesn't Google just run Dijkstra faster?
- What is a shortcut edge and when is it precomputed?
- How much space do shortcut edges take compared to the original graph?
- Can Contraction Hierarchies handle dynamic graphs like traffic or road closure?
- Why contract low-degree nodes first instead of high-degree ones?
- What is a CRDT and why does it matter for real-time collaboration?
- How do CRDTs handle concurrent edits without a central server referee?
- Why did Figma move from operational transforms to CRDTs?
Every answer on Crashtech is written by the editor of the article it comes from — never auto-summarised. Browse all answers or the System Design beat.