Topic series
Ordered series keep distinct perspectives separate while linking overview → deep dive → production.
Caching
From cache vocabulary through strategies, distributed caching, stampede control, and edge CDNs.
- Caching 101 — Memory Offloading and Latency Reduction (canonical-overview)
- Caching Strategies — Aside, Through, Behind, and Refresh-Ahead (deep-dive)
- Cache Eviction Policies — LRU, LFU, TTL, and Friends (deep-dive)
- Distributed Caching — Sharding and High-Availability Clusters (production-guide)
- Cache Stampede — When Expiry Melts the Database (failure-case)
- Stale Cache After Write — When Your Own Update Disappears (failure-case)
- Content Delivery Networks (CDN) — Edge Acceleration and Caching (production-guide)
- Bloom Filters — Probabilistic Set Membership at Scale (deep-dive)
Rate Limiting
Algorithms, service design, distributed enforcement, resilience coupling, and interview framing.
- Rate Limiting Algorithms — Token Bucket, Windows, and Bursts (foundation)
- Rate Limiter Design — Case Study (deep-dive)
- Distributed Rate Limiting — Shared Quotas Across Many Pods (production-guide)
- Circuit Breakers & Rate Limiting Together (production-guide)
Idempotency & Exactly-Once
API idempotency, retries, duplicate gaps, practical EOS, and payment-domain reconciliation.
- Idempotency, Retries & Backoff (deep-dive)
- Duplicate Requests & the Idempotency Gap (failure-case)
- Exactly-Once Processing vs Practical Deduplication (deep-dive)
- Payment Idempotency and Reconciliation (case-study)
CAP, Consistency & Quorums
Consistency vocabulary from CAP through models, strong vs eventual trade-offs, and quorum practice.
- CAP Theorem — Consistency, Availability, and Partition Tolerance (canonical-overview)
- Consistency Models — From Strong to Eventual (deep-dive)
- Strong vs Eventual Consistency — Trade-offs in Distributed Systems (deep-dive)
- CAP, Consistency & Idempotency (deep-dive)
- Quorum Reads vs Quorum Writes (production-guide)
Payments Architecture
Platform architecture, design cases, idempotent money movement, SEPA lifecycle, and banking platforms.
- Payment Platform Architecture (canonical-overview)
- Payment Platform Design (Interview Sketch) (deep-dive)
- Payment Idempotency and Reconciliation (production-guide)
- SEPA Lifecycle & Reconciliation (case-study)
- Banking Transaction Platform Design (case-study)
Kafka & Event Streaming
Queues to Kafka architecture, ordering, EOS, Streams, outbox/saga, and Spring integration.
- Message Queues — Hand Work Off Without Blocking the User (foundation)
- Kafka Architecture — Brokers, Partitions, ISR, and Consumers (canonical-overview)
- Kafka Partition Ordering and Delivery Guarantees (deep-dive)
- Kafka Exactly-Once Semantics — What Is Actually Guaranteed (deep-dive)
- Transactional Outbox and Saga Patterns — Reliable Multi-Step Work (production-guide)
- Kafka Streams Introduction (deep-dive)
- Kafka with Spring Boot Code Walkthrough (implementation-guide)
Replication
Replication foundations, Postgres failover, DDIA notes, and quorum read/write practice.
- Data Replication — Keeping Copies in Sync (canonical-overview)
- DDIA Notes — Replication & Partitioning (book-notes)
- PostgreSQL Replication & Failover (production-guide)
- Quorum Reads vs Quorum Writes (deep-dive)
Reliability & Failover
Availability, reliability, fault tolerance, failover, multi-region survival, and disaster recovery.
- Availability — Nines, Error Budgets, and Redundancy (canonical-overview)
- Reliability — Correct Results Under Stress (foundation)
- Fault Tolerance — Keep Working When Parts Fail (foundation)
- Failover — Switching to a Healthy Spare (deep-dive)
- Multi-Region Failover — Surviving a Region Outage (case-study)
- Disaster Recovery — RPO, RTO, and Backups That Work (production-guide)
Spring Boot Production
DI core through lifecycle, MVC, JPA transactions, security, and production practices.
- Spring Core Dependency Injection (canonical-overview)
- Spring Bean Lifecycle (deep-dive)
- Spring Boot Startup Lifecycle (deep-dive)
- Spring MVC REST APIs (implementation-guide)
- Spring Data JPA and Transactions — Persistence Without Surprises (production-guide)
- Spring Security OAuth2 Resource Server (JWT) (production-guide)
- Spring Boot Best Practices (production-guide)
Java Collections
Collections overview through HashMap, lists, ConcurrentHashMap, and Streams.
- Java Collections Overview (canonical-overview)
- HashMap Internals (deep-dive)
- ArrayList vs LinkedList — When Contiguous Beats Nodes (deep-dive)
- ConcurrentHashMap Internals — Concurrent Hashing Without Global Locks (deep-dive)
- Java Streams API Deep Dive (deep-dive)
Java Concurrency
Memory model and virtual threads through locks, futures, executors, and race debugging.
- Java Memory Model & Virtual Threads (canonical-overview)
- Synchronization, Locks & Deadlocks (deep-dive)
- CompletableFuture Patterns (deep-dive)
- Java Executor Framework — Thread Pools Done Right (production-guide)
- Race Conditions — Finding and Fixing (failure-case)
CDC, Outbox & Sagas
Dual-write problems, outbox/saga patterns, practical exactly-once, and idempotent handlers.
- CDC vs Dual Writes — Keeping Two Stores in Sync (foundation)
- Transactional Outbox and Saga Patterns — Reliable Multi-Step Work (canonical-overview)
- Exactly-Once Processing vs Practical Deduplication (deep-dive)
Staff Decision Operating System
How Staff engineers make change durable: ADRs, RFCs, quality-attribute scenarios, and architecture reviews.
- Architecture Decision Records (ADRs) (canonical-overview)
- Technical RFCs That Ship (production-guide)
- Quality-Attribute Scenarios (deep-dive)
- Architecture Reviews Without Theater (production-guide)
- Technical Strategy for Staff Engineers (concept)
- Cost-Aware Architecture (deep-dive)
Reliability & SRE Practice
From availability vocabulary to SLOs, capacity, tails, PRRs, and incident command.
- Availability — Nines, Error Budgets, and Redundancy (foundation)
- SLIs, SLOs, and Error Budgets — Measure Reliability Like a Product (canonical-overview)
- Capacity Planning for Backend Services (deep-dive)
- Tail Latency and Load Shedding — Surviving Peak Traffic Overload (deep-dive)
- Production-Readiness Reviews (PRRs) (production-guide)
- Incident Command for Backend Teams (production-guide)
Security Hardening for Backends
Threat modelling through secrets, multi-tenant isolation, privacy, and supply chain.
- Threat Modelling for Backend Services (canonical-overview)
- Secrets and Key Management (production-guide)
- Multi-Tenant Isolation Patterns (deep-dive)
- Privacy Engineering Basics (concept)
- Supply-Chain Security and SBOMs (concept)
Platform as a Product
IDP, GitOps, progressive delivery, and contract testing as platform capabilities.
- Internal Developer Platforms (canonical-overview)
- GitOps Fundamentals (concept)
- Progressive Delivery and Feature Flags (production-guide)
- Contract Testing for Services (production-guide)
- CI/CD & Developer Experience (production-guide)