Osman Goni Nahid
Blog
About
Hire me
Kubernetes
2026/10/03
The GPU allocation model: why the number lies
2026/09/30
Shipping one platform to many clusters: release engineering for an AI platform
2026/09/29
Stuck in Undeploying: when your status and your cluster disagree
2026/09/29
Why your LLM endpoint returns 503 (it’s rarely the model)
2026/09/23
Four numbers that decide if your model will serve
2026/09/19
We were hashing every byte twice: making 500 GiB volume imports fast