Understanding the Limitations of Pubsub Systems
Paper: Understanding the Limitations of Pubsub Systems Authors: Atul Adya, Phil Bogle, Colin Meek (Databricks) Venue: HotOS 2025 If you've ever built anything with Kafka, GCP Pub/Sub, RabbitMQ, or AWS

Search for a command to run...
Paper: Understanding the Limitations of Pubsub Systems Authors: Atul Adya, Phil Bogle, Colin Meek (Databricks) Venue: HotOS 2025 If you've ever built anything with Kafka, GCP Pub/Sub, RabbitMQ, or AWS

1. Why Do We Need Caching for Writes? Most caching conversations start and end with reads: "the database is slow, put Redis in front of it." But in any system with real traffic, writes become the bott
A deep dive from first principles to Reed-Solomon math to production architecture Why this matters If you've ever wondered how S3 promises 99.999999999% durability (eleven nines lose 1 object out of

1. Introduction: why reverse image search is a hard engineering problem Here's the trap juniors fall into with this topic: they assume the hard part is "the AI." It isn't. The model that turns an imag

If you ask ten backend engineers how MongoDB Change Streams work, most will answer something like: "You call watch() and MongoDB notifies you whenever a document changes." That answer is technically

When most developers first learn Kafka, they are introduced to a bunch of terms like Producer , Consumer, Partition, Offset, Topics etc. The problem is that none of these explain why Kafka exists. It'

Most engineers learn Kafka before they actually need Kafka. That sentence sounds backwards, but it's the pattern I keep seeing. Teams reach for Kafka to solve problems that a database table and a work
Lessons from building high-scale backend systems When engineers first encounter AsyncLocalStorage, the reaction is often: "Why do we need this? Node.js is single-threaded." It's a reasonable questio