Practical articles on AI, DevOps, Cloud, Linux, and infrastructure engineering.
A condensed checklist of the systemd unit-file patterns we now use everywhere, with the production reasons each one matters.
A systematic approach to debugging Linux network issues. The tools that earn their place and the order I use them in.
A practical Linux performance tuning playbook for production servers. The kernel parameters, disk and network tweaks that earn their place, and the ones that turned out to be folklore.
A practical guide to writing and managing systemd services for production. The unit file features that earn their place, plus the operational workflows.
Run services reliably with systemd: units, dependencies, and resource limits.
We use CloudFront + Lambda@Edge for specific patterns. The wins, the production gotchas, and where we hit Lambda@Edge's limits.
Postgres, DynamoDB, Redis, Elasticsearch, Snowflake. We use all five for different workloads. The decision criteria, not the marketing comparison.
We've executed real disaster recoveries twice. The plan that survived contact with reality, and what was wrong about the plans we had before that.
VPCs, subnets, route tables, gateways. The mental model that finally made cloud networking click after I stopped trying to map it 1:1 to physical networks.
We run both ECS and EKS in production. Which we use for what, and the actual decision criteria — not the marketing comparison.
Shift-left security with image scanning. Trivy, policy gates, and runtime integration.
A working AWS security baseline, derived from the actual incidents we've had and the audit findings we've cleared.