Skip to main content

Running Multiple AI Workloads on One GPU with HAMi: Architecture and Gotchas

August 9, 2026Taipei, Taiwan11:20 - 11:50TR214

GPUs are expensive. Kubernetes doesn't share them well yet, and DRA is still work in progress. HAMi brings heterogeneous GPU sharing to Kubernetes. This talk covers how HAMi hijacks CUDA calls without touching your application code, why memory isolation matters, and real production use cases where teams cut GPU costs by 40-60%.

Join the Community

Get involved with the HAMi open-source project. Connect with maintainers and the community.

CNCFHAMi is a CNCF Incubating project