Insights

Notes from the engineering desk.

Opinionated, evidence-led writing on the systems we build — AI engineering, cloud and FinOps, modernization, and architecture.

What Is a Vector Database? How It Works with LLMs, RAG, and AI
Prasad Waghamare

What Is a Vector Database? How It Works with LLMs, RAG, and AI

Learn what a vector database is, how it works, and why it has become a core component of modern AI applications

Read more
Containers vs Virtual Machines: Differences and Which One to Choose
Saurabh Sawant

Containers vs Virtual Machines: Differences and Which One to Choose

Containers vs Virtual Machines: Compare architecture, performance, security trade-offs, and how to choose the right solution for your workload.

Read more
Terraform vs OpenTofu vs Pulumi: What Each Tool Does and How to Choose
Sahil Deshmukh

Terraform vs OpenTofu vs Pulumi: What Each Tool Does and How to Choose

What is Terraform, OpenTofu, and Pulumi. How each tool works, what problem it solves, the real differences, and how to pick the right one for your team.

Read more
OpenTelemetry on Kubernetes Without the Surprise Bill
Subhendu Nayak

OpenTelemetry on Kubernetes Without the Surprise Bill

Kubernetes autoscaling quietly inflates observability bills. See how OpenTelemetry cuts costs without losing visibility into your systems.

Read more
What Is Lift and Shift Cloud Migration? Benefits & Best Practices
Mahesh Bahir

What Is Lift and Shift Cloud Migration? Benefits & Best Practices

Understand Lift and Shift Cloud Migration, how it works, its benefits, challenges, best practices, and post-migration cost optimization.

Read more
What Is Terraform in DevOps? Learn Infrastructure as Code with Examples
Prasad Waghamare

What Is Terraform in DevOps? Learn Infrastructure as Code with Examples

Master Terraform in DevOps with this beginner-friendly guide. Learn Infrastructure as Code (IaC), providers, resources, state, modules, workspaces, and AWS automation.

Read more
Running LLMs on Kubernetes: What Actually Breaks in Production
Saurabh Sawant

Running LLMs on Kubernetes: What Actually Breaks in Production

A practical guide to running LLMs on Kubernetes: GPU scheduling, model loading, serving frameworks, and the autoscaling tradeoffs most teams get wrong.

Read more
Designing Multi-Region Kubernetes for Rapid Disaster Recovery
Subhendu Nayak

Designing Multi-Region Kubernetes for Rapid Disaster Recovery

Backing up YAML with Velero is not enough for zero data loss. A practical guide to multi-region Kubernetes disaster recovery, RPO=0, and RTO under 5 minutes.

Read more
Designing AI Gateways for Multi-Model Cloud Applications
Mahesh Bahir

Designing AI Gateways for Multi-Model Cloud Applications

Design AI Gateways for multi-model cloud applications with intelligent routing, token governance, orchestration, scalability, and secure AI traffic management.

Read more
Begin a conversation