TCH AI DE Devashish Gupta The Hidden Cost of “Cold Starts”: Defeating EBS Lazy Loading in AI Pipelines In the world of MLOps, we often obsess over model inference time. We spend weeks optimizing PyTorch code to shave off 50 milliseconds per…