GPU Scheduling Landscape and mutex Semantics
Why GPU scheduling matters: how the default scheduler misses node, card, and topology placement, how HAMi v2.10 fills the gaps, and what mutex really means on kind.
GPU Scheduling Landscape and mutex Semantics
Why GPU scheduling matters: how the default scheduler misses node, card, and topology placement, how HAMi v2.10 fills the gaps, and what mutex really means on kind.
GPU sharing is moving from soft allocation to governance—scheduling decisions and runtime isolation can finally be reconciled.
Kubernetes as the GPU Control Plane for AI
Observations on the evolution of AI infrastructure control planes, focusing on HAMi v2.9, GPU scheduling, and Kubernetes resource models.
When GPUs Move Toward Open Scheduling: Structural Shifts in AI Native Infrastructure
A CTO/VP view on open GPU scheduling: CDI, Kubernetes DRA, virtualization data planes, ecosystem governance, and lock-in risk.
AI 2026: Infrastructure, Agents, and the Next Cloud-Native Shift
2026 AI’s turning point: not models, but infrastructure, agentic runtimes, GPU efficiency, and new organizational forms.