Open Internet by MindsNetArticle: Model caching for AI workloads on GKE/Kubernetes without re-downloading weights