Skip to main content

Practical DevOps, Cloud, AI & Linux engineering guides

Featured Article

GitLab's New Rate Limits: What to Fix Before Oct 19

GitLab is capping unauthenticated API calls at 60 an hour starting October 19, and the preview windows land before most teams will have noticed.

KU
Kiril UrbonasAI Engineer
|Oct 4, 2026
GitLab's New Rate Limits: What to Fix Before Oct 19

Most read

  1. 01
  2. 02
  3. 03
  4. 04
  5. 05

Topics

Latest Articles

View All →
We started with a single Celery worker handling everything. Eight months and three architecture changes later, here's what scaled and what we learned about queue design.
••October 7, 2025

Architecture Review: Python Worker Queue Scaling Patterns

We started with a single Celery worker handling everything. Eight months and three architecture changes later, here's what scaled and what we learned about queue design.

KU
Kiril Urbonas·3 min read·32
Read article
We cut our average CI build time from 28 minutes to 6 minutes. The changes that mattered, ranked by impact.
••October 6, 2025

CI/CD Pipeline Optimization: Speeding Up Your Builds

We cut our average CI build time from 28 minutes to 6 minutes. The changes that mattered, ranked by impact.

KU
Kiril Urbonas·3 min read·34
Read article
We scan every container image in CI and at runtime. Trivy + Cosign + admission controllers. The setup that earns its place and what we wish we'd known.
••October 2, 2025

Container Security Scanning: Protecting Your Docker Images

We scan every container image in CI and at runtime. Trivy + Cosign + admission controllers. The setup that earns its place and what we wish we'd known.

KU
Kiril Urbonas·3 min read·32
Read article
We migrated 40+ services to GitOps with Argo CD. Two years in, here's what works and what required workarounds.
••September 28, 2025

GitOps with ArgoCD: Automating Kubernetes Deployments

We migrated 40+ services to GitOps with Argo CD. Two years in, here's what works and what required workarounds.

KU
Kiril Urbonas·3 min read·15
Read article
How a packet actually gets from the internet to a pod, walked layer by layer. Plus the things that surprise people the first time they hit them.
••September 25, 2025

Kubernetes Networking Deep Dive: Understanding Pods, Services, and Ingress

How a packet actually gets from the internet to a pod, walked layer by layer. Plus the things that surprise people the first time they hit them.

KU
Kiril Urbonas·4 min read·43
Read article
Design serverless apps for reliability, cold start, and cost. Event-driven patterns and observability.
••September 22, 2025

AWS Lambda and Serverless Best Practices for Production

Design serverless apps for reliability, cold start, and cost. Event-driven patterns and observability.

KU
Kiril Urbonas·1 min read·41
Read article
We've shipped three end-to-end ML systems. The pieces that look obvious in slides and turn out to be the actual work.
••September 21, 2025

Production AI Pipelines: Building End-to-End ML Systems

We've shipped three end-to-end ML systems. The pieces that look obvious in slides and turn out to be the actual work.

KU
Kiril Urbonas·3 min read·16
Read article
We started routing 90% of LLM traffic through a small internal gateway. The gateway wasn't planned — it emerged from solving the same problem in 5 places. Here's the shape it took.
••September 20, 2025

Architecture Review: LLM Gateway Design for Multi-Provider Inference

We started routing 90% of LLM traffic through a small internal gateway. The gateway wasn't planned — it emerged from solving the same problem in 5 places. Here's the shape it took.

KU
Kiril Urbonas·3 min read·45
Read article
Prompt injection, data leakage, jailbreaks, and the boring controls that actually keep production AI features safe. The threat model that matters once you ship.
••September 18, 2025

AI Security and Safety: Protecting Your AI Applications

Prompt injection, data leakage, jailbreaks, and the boring controls that actually keep production AI features safe. The threat model that matters once you ship.

KU
Kiril Urbonas·4 min read·18
Read article
We benchmarked six embedding models on the same retrieval task. The results that surprised us, and how we'd pick today.
••September 14, 2025

Embedding Models Comparison: Choosing the Right Model for Your Use Case

We benchmarked six embedding models on the same retrieval task. The results that surprised us, and how we'd pick today.

KU
Kiril Urbonas·3 min read·87
Read article
We cut our monthly LLM bill from $11,200 to $2,300 with seven specific changes. The ones that worked, the ones that didn't, and what we'd do first.
••September 10, 2025

AI Cost Optimization: Reducing LLM Inference Costs by 80%

We cut our monthly LLM bill from $11,200 to $2,300 with seven specific changes. The ones that worked, the ones that didn't, and what we'd do first.

KU
Kiril Urbonas·3 min read·40
Read article
Fine-tuning is rarely the right answer. We've fine-tuned three times in two years; few-shot or RAG was correct for everything else. The decision criteria.
••September 7, 2025

Fine-tuning vs Few-Shot Learning: When to Use Each Approach

Fine-tuning is rarely the right answer. We've fine-tuned three times in two years; few-shot or RAG was correct for everything else. The decision criteria.

KU
Kiril Urbonas·3 min read·45
Read article
Page 42 of 47 · 559 posts