Local Llm Ops

作者 bobmatnyc718070a7d622MIT77 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫2 個月前更新

Local LLM operations with Ollama on Apple Silicon, including setup, model pulls, chat launchers, benchmarks, and diagnostics.

僅含說明DevOps & Cloud
AI 產生的概覽

在 Apple Silicon 上使用 Ollama 執行本機大型語言模型的維運指南,涵蓋安裝、拉取模型、對話、基準測試與診斷。

功能
此技能說明在 Apple Silicon 上使用 Ollama 執行本機大型語言模型的維運流程。內容涵蓋安裝 Ollama、啟動服務、初始化虛擬環境、拉取模型,以及啟動對話工作階段或基準測試。也包含診斷步驟與安裝失敗時的常見修正方式。
適用情境
適用於在 macOS 上執行本機大型語言模型、進行以 Ollama 為基礎的對話工作階段,或對模型的速度與延遲進行基準測試。也適合用來排查無法正常運作的本機 Ollama 環境。
執行需求
需要在 macOS 上安裝並執行 Ollama 服務,具備可執行所引用之安裝、對話、基準測試與診斷指令碼的 shell 環境,以及拉取模型所需的網路存取。此技能本身僅為說明文件,不隨附指令碼。

Local LLM Ops (Ollama)

Overview

Your localLLM repo provides a full local LLM toolchain on Apple Silicon: setup scripts, a rich CLI chat launcher, benchmarks, and diagnostics. The operational path is: install Ollama, ensure the service is running, initialize the venv, pull models, then launch chat or benchmarks.

Quick Start

bash
./setup_chatbot.sh./chatllm

If no models are present:

bash
ollama pull mistral

Setup Checklist

  1. Install Ollama: brew install ollama
  2. Start the service: brew services start ollama
  3. Run setup: ./setup_chatbot.sh
  4. Verify service: curl http://localhost:11434/api/version

Chat Launchers

  • ./chatllm (primary launcher)
  • ./chat or ./chat.py (alternate launchers)
  • Aliases: ./install_aliases.sh then llm, llm-code, llm-fast

Task modes:

bash
./chat -t coding -m codellama:70b./chat -t creative -m llama3.1:70b./chat -t analytical

Benchmark Workflow

Benchmarks are scripted in scripts/run_benchmarks.sh:

bash
./scripts/run_benchmarks.sh

This runs bench_ollama.py with:

  • benchmarks/prompts.yaml
  • benchmarks/models.yaml
  • Multiple runs and max token limits

Diagnostics

Run the built-in diagnostic script when setup fails:

bash
./diagnose.sh

Common fixes:

  • Re-run ./setup_chatbot.sh
  • Ensure ollama is in PATH
  • Pull at least one model: ollama pull mistral

Operational Notes

  • Virtualenv lives in .venv
  • Chat configs and sessions live under ~/.localllm/
  • Ollama API runs at http://localhost:11434

Related Skills

  • toolchains/universal/infrastructure/docker

來源與署名

來源:bobmatnyc/claude-mpm-skills位於toolchains/ai/ops/local-llm-ops提交718070a

授權條款: MIT

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架