Local Llm Ops

作者 bobmatnyc718070a7d622MIT77 个星标收录于 2026年10月8日更新于 2026年10月8日仓库2个月前更新

Local LLM operations with Ollama on Apple Silicon, including setup, model pulls, chat launchers, benchmarks, and diagnostics.

仅含说明DevOps & Cloud
AI 生成的概览

在 Apple Silicon 上使用 Ollama 运行本地大模型的运维指南,涵盖安装、拉取模型、对话、基准测试与诊断。

功能
该技能记录了在 Apple Silicon 上使用 Ollama 运行本地大语言模型的运维流程。它涵盖安装 Ollama、启动服务、初始化虚拟环境、拉取模型,以及启动对话会话或基准测试运行。它还包含诊断步骤和安装失败时的常见修复方法。
适用场景
适用于在 macOS 上运行本地大模型、进行基于 Ollama 的对话会话,或对模型的速度与延迟进行基准测试。也适用于排查无法正常工作的本地 Ollama 环境。
运行要求
需要在 macOS 上安装并运行 Ollama 服务,具备可执行所引用的安装、对话、基准测试和诊断脚本的 shell 环境,以及拉取模型所需的网络访问。该技能本身仅为说明文档,不附带脚本。

Local LLM Ops (Ollama)

Overview

Your localLLM repo provides a full local LLM toolchain on Apple Silicon: setup scripts, a rich CLI chat launcher, benchmarks, and diagnostics. The operational path is: install Ollama, ensure the service is running, initialize the venv, pull models, then launch chat or benchmarks.

Quick Start

bash
./setup_chatbot.sh./chatllm

If no models are present:

bash
ollama pull mistral

Setup Checklist

  1. Install Ollama: brew install ollama
  2. Start the service: brew services start ollama
  3. Run setup: ./setup_chatbot.sh
  4. Verify service: curl http://localhost:11434/api/version

Chat Launchers

  • ./chatllm (primary launcher)
  • ./chat or ./chat.py (alternate launchers)
  • Aliases: ./install_aliases.sh then llm, llm-code, llm-fast

Task modes:

bash
./chat -t coding -m codellama:70b./chat -t creative -m llama3.1:70b./chat -t analytical

Benchmark Workflow

Benchmarks are scripted in scripts/run_benchmarks.sh:

bash
./scripts/run_benchmarks.sh

This runs bench_ollama.py with:

  • benchmarks/prompts.yaml
  • benchmarks/models.yaml
  • Multiple runs and max token limits

Diagnostics

Run the built-in diagnostic script when setup fails:

bash
./diagnose.sh

Common fixes:

  • Re-run ./setup_chatbot.sh
  • Ensure ollama is in PATH
  • Pull at least one model: ollama pull mistral

Operational Notes

  • Virtualenv lives in .venv
  • Chat configs and sessions live under ~/.localllm/
  • Ollama API runs at http://localhost:11434

Related Skills

  • toolchains/universal/infrastructure/docker

来源与署名

来源:bobmatnyc/claude-mpm-skills位于toolchains/ai/ops/local-llm-ops提交718070a

许可证: MIT

内容归原作者所有。SourceWeft 从公开仓库中收录这些内容。

举报或申请下架