Senior SRE
Keeping distributed systems reliable — from Kubernetes networking to multi-cloud infrastructure as code.
Site reliability engineer based in Taipei, working where platform, network and operations meet.
I am a site reliability engineer with more than twelve years in the software industry, split between back-end development and the infrastructure that keeps it running. My work centres on making distributed systems observable, predictable and cheap to operate — defining service levels that mean something, automating the paths that used to need a human, and understanding a platform deeply enough to debug it at three in the morning.
Most of my recent time goes into Kubernetes platform engineering: cluster networking with Cilium and eBPF, certificate and identity automation, and describing multi-cloud infrastructure entirely in Terraform and Terragrunt. I write up most of what I learn — the blog is where those working notes end up.
Grouped by the problem they solve, rather than by how long the list can get.
Working notes from the platform — over a hundred posts, still going.
Employment history.
Senior Software Engineer
DevOps Engineer
DevOps Engineer
DevOps Engineer
DevOps Engineer
Senior Program Analyst
Postgraduate research · Teaching assistant