
Hf Cloud Sagemaker Deployment Planner
by huggingfaceabc20ae526d8No license11K starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated today
Plan and coordinate the deployment of a model to Amazon SageMaker AI. Use this skill whenever the user wants to deploy, host, serve, or expose a model on SageMaker or AWS — including phrases like "deploy a model", "host this LLM on AWS", "serve this embedding model", "deploy a reranker", "deploy a text-to-image / diffusion model", "host this for async inference", "create an endpoint", "serve my fine-tuned model", or any request that involves making a model available for inference on AWS. Use this even when the user is vague (e.g. "I just want to get this running on AWS, you figure it out"). Works for text-generation LLMs, embedding models, rerankers, classifiers, text-to-image / diffusion models — picks the right serving stack and chooses between real-time and async inference. This is the entry-point skill for SageMaker deployment work — it asks clarifying questions, picks a deployment pathway, and coordinates the other deployment skills.
Only the file list is public. File contents are available once the skill is installed in a workspace.
| Path | Size | Type |
|---|---|---|
| SKILL.md | 8.4 KB | text/markdown |
Source and attribution
Source:huggingface/skillsinskills/hf-cloud-sagemaker-deployment-plannerat commitabc20ae
License: No license
Content belongs to its original authors. SourceWeft indexes it from a public repository.
More from huggingface/skills

Trl Training
huggingface
Guides fine-tuning and aligning transformer language models with TRL CLI commands for SFT, DPO, GRPO, KTO, RLOO and reward models.

Huggingface Zerogpu
huggingface
Guidance for writing and reviewing Gradio apps that run on Hugging Face Spaces ZeroGPU hardware.

Huggingface Trackio
huggingface
Guides logging, alerting on, and retrieving ML training metrics with the Trackio experiment tracking library.

Huggingface Papers
huggingface
Fetches Hugging Face paper pages as markdown and queries the papers API for metadata, links, search and daily papers.

Huggingface Local Models
huggingface
Guides selecting and running GGUF models locally with llama.cpp, covering Hub search, quant choice, and local serving.

Huggingface Gradio
huggingface
Reference guide for building Gradio web UIs and ML demos in Python, covering components, layouts, events and chatbots.
More in DevOps & Cloud

M5 Onboard
anthropics
Provisions M5Stack ESP32 boards by detecting them on USB, flashing UIFlow 2.0 firmware, and installing a MicroPython app bundle.

Spanner Basics
Guides Google Cloud Spanner administration, schema design, querying and performance diagnosis.

Secops Cases
Manages Google Security Operations SOAR cases across their lifecycle via MCP tools.

Managed Airflow Migrations
Guides migration of Apache Airflow DAGs to Airflow 2.11.1 or Airflow 3 in Managed Service for Apache Airflow.

Managed Airflow Dag Troubleshooting
Guides deterministic troubleshooting of failed Managed Service for Apache Airflow DAG runs and task instances using gcloud commands.

Iam Helper For Privileged Access Management
Guides Google Cloud Privileged Access Manager entitlement CRUD, temporary access requests, and grant approval workflows.