Finetuning Method Selection

作者 wshobson46891e7e60da無授權條款收錄於 2026年10月8日更新於 2026年10月8日

Decide whether to fine-tune at all, and route to the right method (SFT, DPO/ORPO/KTO, GRPO/RLVR, continued pretraining) and base model. Use when starting any fine-tuning effort, when unsure whether RAG or prompting would suffice, or when choosing between preference-optimization and reinforcement methods.

僅含說明AI & Agents
  1. 46891e7e60da目前提交 46891e7發布於 2026年10月8日

來源與署名

來源:wshobson/agents位於plugins/llm-finetuning/skills/finetuning-method-selection提交46891e7

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架