Alex Merced's Data, Dev and AI Blog

Topic

database

16 posts tagged “database”.

  1. 9 min read

    Concurrency, Isolation, and MVCC: How Engines Handle Contention

    Databases handle concurrent access using locks, MVCC, or optimistic concurrency control. Here is how each approach works and what tradeoffs each creates.

  2. 8 min read

    Hash, Sort-Merge, Broadcast: How Distributed Joins Work

    Distributed joins move data across the network using shuffle, broadcast, or co-location strategies. Here is how each works and when engines choose which.

  3. 7 min read

    Partitioning, Sharding, and Data Distribution Strategies

    Hash partitioning distributes data evenly. Range partitioning enables fast range scans. Both create tradeoffs.

  4. 8 min read

    Buffer Pools, Caches, and the Memory Hierarchy

    Databases use buffer pools, column caches, and result caches to keep hot data in RAM. Here is how each caching strategy works and what happens when data.

  5. 8 min read

    Volcano, Vectorized, Compiled: How Engines Execute Your Query

    The Volcano model processes one row at a time. Vectorized execution processes batches with SIMD. Code generation fuses operators into compiled code.

  6. 8 min read

    Inside the Query Optimizer: How Engines Pick a Plan

    Query optimizers transform SQL into execution plans using rule-based rewrites, cost-based search, and adaptive runtime adjustments.

  7. 8 min read

    B-Trees, LSM Trees, and the Indexing Tradeoff Spectrum

    B-trees balance reads and writes for OLTP. LSM trees maximize write throughput. Bitmap indexes accelerate OLAP filtering. Here is when to use each.

  8. 8 min read

    How Databases Organize Data on Disk: Pages, Blocks, and File Formats

    Databases structure data on disk as heap files, sorted files, or LSM trees, then wrap it in formats like Parquet with metadata that lets engines skip.

  9. 8 min read

    Row vs. Column: How Storage Layout Shapes Everything

    Row stores keep records together for fast transactions. Column stores keep field values together for fast analytics.

  10. 9 min read

    How Query Engines Think: The Tradeoffs Behind Every Data System

    Every database is a collection of engineering tradeoffs. Learn the 9 design decisions that shape how query engines store, index, and process your data.

  11. 20 min read

    Introduction to ANSI SQL - Understanding the Syntax and Concepts

    Learning the Standard SQL Syntax

  12. 6 min read

    Columnar vs. Row-based Data Structures in OLTP and OLAP Systems

    The Fundamentals of Data Systems

  13. 8 min read

    Building Full CRUD Rest API's with Flask & FastAPI using PsychoPG2

    When it comes to building web applications in the modern era, developers are spoilt for choice with a plethora of frameworks and libraries at their…

  14. 13 min read

    Introduction to The World of Data - (OLTP, OLAP, Data Warehouses, Data Lakes and more)

    An accessible high-level guide for data and non-data professionals

  15. 5 min read

    A 2022 Introduction to SQL

    Learning Structured Query Language

  16. 6 min read

    2022 MongooseJS Cheatsheet

    Details on working with MongooseJS

Browse all posts

Work with Alex

Menu

Search

Type at least two characters.