01 / Cloud architecture
AWS High Availability: Design for Failure, Then Test It
A practical guide to failure domains, application state, database recovery and the capacity your surviving zone actually needs.
Read the articleThe engineering notebook / 01—05
Notes on building reliable infrastructure, writing better backends and making automation earn its place.
01 / Cloud architecture
A practical guide to failure domains, application state, database recovery and the capacity your surviving zone actually needs.
Read the article02 / FinOps
Move beyond the monthly invoice: choose a meaningful unit, find unnecessary work and verify savings without sacrificing reliability.
Read the article03 / Observability
Connect metrics, logs and traces to user-visible outcomes, then build an investigation path that works during a real failure.
Read the article04 / Backend engineering
Make caching predictable: define freshness, protect data boundaries and understand the races that annotations cannot solve alone.
Read the article05 / AI & infrastructure
Let a model interpret the request while code controls validation, authorization, change review and cluster execution.
Read the article