As an ML engineer,
I want to be able to explicitly define the GPU partition in the deployment charts manifest so that:
- scheduling aligns with the model’s resource requirements
- larger GPU partitions are not allocated unnecessarily & GPU resources remain available for other workloads