Skip to main content

Practical DevOps, Cloud, AI & Linux engineering guides

Featured Article

GitLab's New Rate Limits: What to Fix Before Oct 19

GitLab is capping unauthenticated API calls at 60 an hour starting October 19, and the preview windows land before most teams will have noticed.

KU
Kiril UrbonasAI Engineer
|Oct 4, 2026
GitLab's New Rate Limits: What to Fix Before Oct 19

Most read

  1. 01
  2. 02
  3. 03
  4. 04
  5. 05

Topics

Latest Articles

View All →
Practical patterns for Terraform modules at scale: versioning, composition, testing, and avoiding the monolith trap.
••6 months ago

Terraform Modules Done Right: Lessons from Managing 50+ Services

Practical patterns for Terraform modules at scale: versioning, composition, testing, and avoiding the monolith trap.

KU
Kiril Urbonas·1 min read·30
Read article
Step-by-step debugging of a production Linux server hitting 100% CPU. From top to perf to the actual fix.
••6 months ago

Linux Performance Troubleshooting: A Real Incident Walkthrough

Step-by-step debugging of a production Linux server hitting 100% CPU. From top to perf to the actual fix.

KU
Kiril Urbonas·1 min read·24
Read article
Battle-tested prompt patterns from running LLM features in production: structured output, chain-of-thought, and graceful failure handling.
••6 months ago

Prompt Engineering Patterns That Actually Work in Production

Battle-tested prompt patterns from running LLM features in production: structured output, chain-of-thought, and graceful failure handling.

KU
Kiril Urbonas·1 min read·47
Read article
A real cost audit uncovered idle load balancers, oversized RDS instances, and forgotten snapshots. Here's what we found and how we fixed each one.
••6 months ago

AWS Cost Audit: 7 Things We Found Wasting Money Every Month

A real cost audit uncovered idle load balancers, oversized RDS instances, and forgotten snapshots. Here's what we found and how we fixed each one.

KU
Kiril Urbonas·1 min read·22
Read article
A real walkthrough of shrinking bloated Docker images from 1.2GB to 240MB using multi-stage builds, Alpine, and dependency auditing.
••6 months ago

How We Cut Our Docker Image Size by 80% and Why It Matters

A real walkthrough of shrinking bloated Docker images from 1.2GB to 240MB using multi-stage builds, Alpine, and dependency auditing.

KU
Kiril Urbonas·1 min read·22
Read article
A real-world model fallback guide for customer-facing AI systems, covering how one team preserved response quality and support SLAs during a partial provider degradation.
••6 months ago

Model Fallback Policies for Customer-Facing AI: The Routing Rules That Kept SLA Intact

A real-world model fallback guide for customer-facing AI systems, covering how one team preserved response quality and support SLAs during a partial provider degradation.

KU
Kiril Urbonas·2 min read·47
Read article
A practical artifact promotion guide for CI/CD teams that were tired of hearing 'it passed in staging' after production behaved differently because the release was rebuilt.
••6 months ago

Artifact Promotion Instead of Rebuilds: The Release Control Pattern That Stopped Drift

A practical artifact promotion guide for CI/CD teams that were tired of hearing 'it passed in staging' after production behaved differently because the release was rebuilt.

KU
Kiril Urbonas·2 min read·114
Read article
A hands-on RDS restore drill guide for small cloud teams that thought backups were covered until a timed restore test exposed missing steps, DNS confusion, and stale credentials.
••6 months ago

RDS Restore Drills for Busy Teams: The Recovery Workflow That Surfaced Real Gaps

A hands-on RDS restore drill guide for small cloud teams that thought backups were covered until a timed restore test exposed missing steps, DNS confusion, and stale credentials.

KU
Kiril Urbonas·2 min read·30
Read article
A practical systemd drop-in guide built from a real operations problem: vendor unit files kept changing, but the team still needed consistent restart, environment, and logging behavior.
••6 months ago

Systemd Drop-In Overrides for Vendor Services: The Supportable Linux Ops Pattern

A practical systemd drop-in guide built from a real operations problem: vendor unit files kept changing, but the team still needed consistent restart, environment, and logging behavior.

KU
Kiril Urbonas·2 min read·28
Read article
A real-world Terraform module version pinning guide for platform teams that want safer upgrades, clearer ownership, and fewer broken pipelines after shared module releases.
••6 months ago

Terraform Module Version Pinning: How One Platform Team Stopped Surprise Breakage

A real-world Terraform module version pinning guide for platform teams that want safer upgrades, clearer ownership, and fewer broken pipelines after shared module releases.

KU
Kiril Urbonas·2 min read·28
Read article
A practical embedding model upgrade guide for RAG systems, built from a real support-search migration that initially reduced answer quality instead of improving it.
••6 months ago

Embedding Model Upgrades Without Search Chaos: A Safer RAG Rollout Pattern

A practical embedding model upgrade guide for RAG systems, built from a real support-search migration that initially reduced answer quality instead of improving it.

KU
Kiril Urbonas·2 min read·80
Read article
A real-world multi-cluster traffic routing guide for SaaS teams that have outgrown a single Kubernetes cluster and need safer rollout control without a service-mesh science project.
••6 months ago

Multi-Cluster Traffic Routing Strategies: A Pragmatic Rollout Pattern for Growing SaaS Teams

A real-world multi-cluster traffic routing guide for SaaS teams that have outgrown a single Kubernetes cluster and need safer rollout control without a service-mesh science project.

KU
Kiril Urbonas·1 min read·34
Read article
Page 36 of 47 · 559 posts