Finetuning Method Selection

作者 wshobson46891e7e60da無授權條款收錄於 2026年10月8日更新於 2026年10月8日

Decide whether to fine-tune at all, and route to the right method (SFT, DPO/ORPO/KTO, GRPO/RLVR, continued pretraining) and base model. Use when starting any fine-tuning effort, when unsure whether RAG or prompting would suffice, or when choosing between preference-optimization and reinforcement methods.

僅含說明AI & Agents

僅公開檔案列表。將技能安裝到工作區後即可檢視檔案內容。

路徑大小類型
references/memory-math.md5.3 KBtext/markdown
references/model-catalog.md2.7 KBtext/markdown
SKILL.md7.7 KBtext/markdown

來源與署名

來源:wshobson/agents位於plugins/llm-finetuning/skills/finetuning-method-selection提交46891e7

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架