Render Scaling

作者 render-osse8f889396634MIT收錄於 2026年10月8日更新於 2026年10月8日

Scales Render services—configures autoscaling targets, chooses instance types, sets manual instance counts, and optimizes cost. Use when the user needs to handle more traffic, set up autoscaling, pick the right instance type, reduce costs, or troubleshoot scaling behavior like slow scale-down or stuck instances.

僅含說明DevOps & Cloud
AI 產生的概覽

指導 Render 服務的擴縮容,涵蓋執行個體數量、自動擴縮與成本最佳化。

功能
此技能提供在 Render 上擴縮容 Web Service、Private Service 與 Background Worker 的操作說明。內容涵蓋手動設定執行個體數量、Professional 以上工作區的自動擴縮(CPU 與記憶體目標)、擴縮計算公式、擴縮限制、縮容行為、執行個體類型(plan)選擇、成本模式,以及 numInstances、scaling、plan 等 Blueprint 欄位。它也會指向包含執行個體類型表與自動擴縮調校建議的參考檔案。
適用情境
適用於需要因應更大流量、設定或調整自動擴縮、在垂直與水平擴縮之間做選擇、挑選執行個體類型、降低成本,或排解縮容緩慢、執行個體卡住等擴縮行為問題時。
執行需求
不含指令碼,僅為說明性內容。前提是使用 Render 的 Web Service、Private Service 或 Background Worker;自動擴縮需要 Professional 或更高等級的工作區。參考檔案由模型讀取。

Render Scaling

This skill covers how to scale Web Services, Private Services, and Background Workers on Render: manual instance counts, Professional+ autoscaling, plan (instance type) choices, and platform limits. Deeper tables and tuning guidance live under references/.

When to Use

  • Setting or changing instance count (Dashboard, CLI, API, or Blueprint)
  • Configuring autoscaling (min/max, CPU and memory targets)
  • Choosing vertical (plan) vs horizontal (more instances) scaling
  • Understanding constraints (disks, static sites, cron/workflows, 100-instance cap)
  • Cost implications of multi-instance and per-second billing
  • Blueprint fields: numInstances, scaling, plan

Manual Scaling

  • Set instance count from 1 to 100 via the Dashboard, CLI, or API.
  • All instances share the same instance type (plan); you cannot mix plans on one service.
  • Changes apply immediately: Render provisions new instances and deprovisions excess capacity as needed.

Autoscaling

  • Available on Professional and higher workspaces only.
  • Configure minimum and maximum instances and targets for CPU and/or memory utilization (1–90% each).
  • At least one metric must be enabled (CPU or memory). If both CPU and memory autoscaling toggles are off, autoscaling is disabled.
  • If both manual instance settings and autoscaling are configured, autoscaling wins—manual count does not override the scaling policy in effect.

Autoscaling Formula

Render computes a candidate instance count from utilization vs target:

new_instances = ceil(current_instances * (current_utilization / target_utilization))

  • When both CPU and memory targets are set, the platform uses the larger of the two new_instances values (the more conservative scale-out).

Scaling Constraints

ConstraintBehavior
Per serviceMaximum 100 instances
Persistent diskCannot scale to multiple instances—single instance only
Static sitesNot scalable (served by CDN)
Cron jobs & WorkflowsScaling model does not apply (different execution model)

Scale-Down Behavior

  • Scale-up is immediate when utilization supports it.
  • Scale-down waits a few minutes after conditions allow reduction (spike protection). This reduces flapping from brief load spikes.

Instance Types

  • In Blueprints, the instance type is the plan field (e.g. standard, pro).
  • Options span free / starter through standard, pro, pro_plus, pro_max, pro_ultra—each with defined CPU and RAM (see references/instance-types.md).

Vertical vs Horizontal

NeedApproachWhen
More throughputHorizontal (add instances)Stateless services, request-based workloads
More RAM/CPU per processVertical (upgrade plan)Memory-intensive or single-threaded apps
BothCombineRight-size plan, then scale out for traffic

Cost Patterns

  • Per-second billing; no separate fee for scaling actions.
  • You pay roughly for compute time × number of running instances (see Render pricing for current rates).
  • Right-size by monitoring CPU and memory utilization (see render-monitor).

Blueprint Configuration

Manual instance count:

yaml
numInstances: 3

Autoscaling:

yaml
scaling:  minInstances: 1  maxInstances: 10  targetCPUPercent: 70  targetMemoryPercent: 80

Instance type (plan):

yaml
plan: standard

Do not rely on numInstances to cap autoscaling when a scaling block is present—autoscaling takes precedence. Preview behavior for scaling is detailed in references/autoscaling-guide.md.

References

TopicFile
Plan names, CPU/RAM, flexible vs non-flexible, free tierreferences/instance-types.md
Enabling autoscaling, targets, min/max, mistakes, previewsreferences/autoscaling-guide.md

Related Skills

  • render-web-services — Web Service settings, disks, deploy lifecycle
  • render-background-workers — Worker-specific configuration and scaling context
  • render-blueprints — Full Blueprint schema and field reference
  • render-monitor — Metrics, logs, and utilization for right-sizing

來源與署名

來源:render-oss/render-plugin-claude-code位於skills/render-scaling提交e8f8893

授權條款: MIT

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架