
Benchmark Model
modular/skills/benchmark-model作者 modularb9b3a8e86700無授權條款200 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫今天更新
Benchmark a model served on MAX with the `max benchmark` command: measure throughput (tokens/sec), latency (TTFT, TPOT, inter-token latency), and GPU utilization by driving load against a running `max serve` endpoint. Use this whenever the user wants to benchmark, load-test, or measure the performance of a MAX model, get tokens-per-second / TTFT / TPOT numbers, run a concurrency or request-rate sweep, compare latency vs throughput, size a deployment, or produce benchmark JSON, even if they don't say "benchmark" by name. Also use when a `max benchmark` run fails to connect or reports zero/garbage numbers.
- b9b3a8e86700目前提交 b9b3a8e發布於 2026年10月8日
更多DevOps & Cloud技能

Playwright Devops
microsoft
Playwright 的 DevOps 工作流程:分析 main 分支最新提交的 GitHub Actions 失敗並下載失敗作業日誌。

M5 Onboard
anthropics
透過 USB 偵測 M5Stack ESP32 開發板,燒錄 UIFlow 2.0 韌體並安裝 MicroPython 應用程式套件。
Yeet
openai
使用 GitHub CLI 一次完成暫存、提交、推送並建立或更新 GitHub 拉取請求。

Runbook
anthropics
為重複性任務建立或更新逐步操作手冊,包含疑難排解、回復與升級流程。

Incident Response
anthropics
引導事故回應流程:嚴重程度分級、狀態更新、緩解追蹤,以及無責事後檢討。

Deploy Checklist
anthropics
產生部署前就緒檢查清單,涵蓋部署前、部署、部署後與回復觸發條件。