Hf Cloud Sagemaker Deployment Planner

作者 huggingfaceabc20ae526d8無授權條款11K 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫今天更新

Plan and coordinate the deployment of a model to Amazon SageMaker AI. Use this skill whenever the user wants to deploy, host, serve, or expose a model on SageMaker or AWS — including phrases like "deploy a model", "host this LLM on AWS", "serve this embedding model", "deploy a reranker", "deploy a text-to-image / diffusion model", "host this for async inference", "create an endpoint", "serve my fine-tuned model", or any request that involves making a model available for inference on AWS. Use this even when the user is vague (e.g. "I just want to get this running on AWS, you figure it out"). Works for text-generation LLMs, embedding models, rerankers, classifiers, text-to-image / diffusion models — picks the right serving stack and chooses between real-time and async inference. This is the entry-point skill for SageMaker deployment work — it asks clarifying questions, picks a deployment pathway, and coordinates the other deployment skills.

僅含說明DevOps & Cloud
  1. abc20ae526d8目前提交 abc20ae發布於 2026年10月8日

來源與署名

來源:huggingface/skills位於skills/hf-cloud-sagemaker-deployment-planner提交abc20ae

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架