
Benchmark Model
modular/skills/benchmark-model作者 modularb9b3a8e86700无许可证200 个星标收录于 2026年10月8日更新于 2026年10月8日仓库今天更新
Benchmark a model served on MAX with the `max benchmark` command: measure throughput (tokens/sec), latency (TTFT, TPOT, inter-token latency), and GPU utilization by driving load against a running `max serve` endpoint. Use this whenever the user wants to benchmark, load-test, or measure the performance of a MAX model, get tokens-per-second / TTFT / TPOT numbers, run a concurrency or request-rate sweep, compare latency vs throughput, size a deployment, or produce benchmark JSON, even if they don't say "benchmark" by name. Also use when a `max benchmark` run fails to connect or reports zero/garbage numbers.
- b9b3a8e86700当前提交 b9b3a8e发布于 2026年10月8日
更多DevOps & Cloud技能

Playwright Devops
microsoft
面向 Playwright 的 DevOps 工作流:分析 main 分支最新提交的 GitHub Actions 失败并下载失败作业日志。

M5 Onboard
anthropics
通过 USB 检测 M5Stack ESP32 开发板,刷写 UIFlow 2.0 固件并安装 MicroPython 应用包。
Yeet
openai
使用 GitHub CLI 一次性完成暂存、提交、推送并创建或更新 GitHub 拉取请求。

Runbook
anthropics
为重复性任务创建或更新分步运维手册,包含故障排查、回滚和升级流程。

Incident Response
anthropics
指导事故响应流程:严重级别判定、状态更新、缓解跟踪以及无责事后复盘。

Deploy Checklist
anthropics
生成部署前就绪检查清单,涵盖部署前、部署、部署后和回滚触发条件。