Skip to content
DevOpsSociety

Kubernetes GPU Scheduling Explained

Fractional GPUs, time-slicing and MIG — how the scheduler decides which pod gets silicon, and how to stop wasting it.

Priya Nair9 min read
Share

Published your local timeupdated

Fractional GPUs, time-slicing and MIG — how the scheduler decides which pod gets silicon, and how to stop wasting it.

Written by

Priya Nair

Principal Editor, Cloud & Kubernetes

Priya writes about production Kubernetes and cloud-native architecture. Previously a staff SRE running multi-region clusters at scale.

More from Priya
The Infrastructure Briefing

Get the infrastructure briefing.

Practical DevOps, cloud, AI infrastructure and engineering insights — delivered weekly. Read by engineers and engineering leaders.

No spam. Unsubscribe anytime.