Kubernetes GPU Scheduling Explained
Fractional GPUs, time-slicing and MIG — how the scheduler decides which pod gets silicon, and how to stop wasting it.
Engineering the Future of Infrastructure
Principal Editor, Cloud & Kubernetes
Priya writes about production Kubernetes and cloud-native architecture. Previously a staff SRE running multi-region clusters at scale.
5 stories
Fractional GPUs, time-slicing and MIG — how the scheduler decides which pod gets silicon, and how to stop wasting it.
A practical production checklist covering security, reliability, networking, observability and cost — the configuration that separates a demo from a system that carries revenue.
One standard for metrics, logs and traces. How OpenTelemetry actually fits together, and how to adopt it without a big-bang migration.
Declarative delivery is easy in the demo. Secrets, drift, multi-cluster and rollbacks are where GitOps earns or loses trust.
The honest decision tree. When a mesh earns its operational cost, and when a good ingress and library get you there cheaper.
Practical DevOps, cloud, AI infrastructure and engineering insights — delivered weekly. Read by engineers and engineering leaders.