{"enrichment":{"faq":[{"a":"ml-cloud-deployment guides you through SageMaker model deployment by covering endpoint creation, inference instance selection, and traffic routing. The skill walks you through packaging your model, registering it in the SageMaker Model Registry, and configuring real-time or batch transform jobs. You'll learn to choose between managed endpoints and serverless inference based on your latency and cost requirements.","q":"How do I deploy a machine learning model to AWS SageMaker?"},{"a":"ml-cloud-deployment covers autoscaling configuration and cost optimization for Vertex AI workloads. The skill explains how to set up custom training jobs with automatic scaling, configure prediction endpoints with traffic-based scaling policies, and use Vertex AI's hyperparameter tuning to optimize resource allocation. You'll learn to balance performance against spending through instance type selection and scaling thresholds.","q":"What are the best practices for scaling ML workloads on GCP Vertex AI?"},{"a":"ml-cloud-deployment teaches real-time endpoint setup across AWS, GCP, and Azure platforms. The skill covers containerizing your model, configuring endpoint autoscaling policies, setting up health checks and traffic routing, and monitoring latency and throughput. You'll learn deployment patterns including blue-green deployments and canary rollouts to safely update models in production.","q":"How do I set up a real-time inference endpoint?"},{"a":"ml-cloud-deployment helps you select appropriate compute resources by analyzing your workload's requirements. The skill covers GPU vs. TPU trade-offs, spot instance strategies for cost savings, and infrastructure decisions for batch vs. real-time inference. You'll learn how to profile your model, estimate throughput needs, and match them to cloud provider offerings across AWS, GCP, and Azure.","q":"What hardware and infrastructure should I choose for ML training and inference?"},{"a":"ml-cloud-deployment provides strategies for reducing ML training costs across managed platforms. The skill covers using spot instances and preemptible VMs, right-sizing compute resources, scheduling jobs during off-peak hours, and leveraging multi-region deployments. You'll learn to configure autoscaling policies that balance speed against expense and monitor spending through cloud provider cost analysis tools.","q":"How can I cost-optimize my ML training jobs on cloud platforms?"},{"a":"ml-cloud-deployment covers Kubernetes-based model serving using KServe and similar frameworks. The skill explains containerizing models with Docker, deploying inference servers to Kubernetes clusters, configuring autoscaling based on request volume, and managing model versioning. You'll learn to orchestrate multi-model deployments and implement canary updates for safe production rollouts.","q":"What deployment options exist for containerized ML models on Kubernetes?"}],"shadow_tags":["managed-platforms","gpu-acceleration","model-serving","infrastructure-as-code","multi-cloud","cost-efficiency","production-deployment","autoscaling-strategy","containerization","model-registry"],"summary_rewrite":"This skill guides you through deploying machine learning workloads on managed cloud platforms, Kubernetes clusters, and serverless systems. It covers platform selection across AWS, GCP, Azure, Databricks, and specialized providers, plus practical patterns for endpoint configuration, training job orchestration, and scaling decisions based on your workload's latency, throughput, and compliance needs."},"files":[{"bytes":18429,"path":"plugins/ml-master/skills/ml-cloud-deployment/SKILL.md","sha256":"c47bd716c6ad5799ef92dc7a5a4416bc7d8783da66639c189ed545d6fc8766d3","url":"https://skillfed.io/files/JosiahSiegel/claude-plugin-marketplace/ml-cloud-deployment/442e1dcc/SKILL.md"}],"id":"JosiahSiegel/claude-plugin-marketplace/ml-cloud-deployment","links":{"html":"https://skillfed.io/JosiahSiegel/claude-plugin-marketplace/ml-cloud-deployment","md":"https://skillfed.io/JosiahSiegel/claude-plugin-marketplace/ml-cloud-deployment.md","repo":"https://github.com/JosiahSiegel/claude-plugin-marketplace"},"meta":{"agents_supported":[],"first_seen":"2026-07-28","forks":10,"language":"Shell","last_updated":"2026-06-18","license":"MIT","name":"ml-cloud-deployment","publisher":"JosiahSiegel","stars":49},"relations":{"similar":[{"id":"JosiahSiegel/claude-plugin-marketplace/azure-ml-foundry-workspace"},{"id":"ancoleman/ai-design-components/implementing-mlops"},{"id":"JosiahSiegel/claude-plugin-marketplace/ml-mlops"},{"id":"synthetic-sciences/openscience/mlflow"},{"id":"Orchestra-Research/AI-Research-SKILLs/mlflow"},{"id":"OpenLAIR/dr-claw/mlflow"},{"id":"ancoleman/ai-design-components/deploying-on-gcp"},{"id":"foryourhealth111-pixel/Vibe-Skills/ml-pipeline-workflow"},{"id":"HermeticOrmus/LibreUIUX-Claude-Code/ml-pipeline-workflow"},{"id":"wshobson/agents/ml-pipeline-workflow"}]},"slug":{"owner":"JosiahSiegel","repo":"claude-plugin-marketplace","skill":"ml-cloud-deployment"},"version":"442e1dcc"}
