ml-stack-training
Design advanced ML training when a user asks for SFT, DPO, GRPO, reward modeling, LoRA, QLoRA, sentence-transformer, vision, checkpointing, Trackio, or HF persistence; emit an executable specification unless an equivalent runtime operation exists.
Pinned to revision 6829bde67fb6, so it is the text this page describes rather than whatever the author pushed since.
Files
- skills/ml-stack-training/SKILL.md
- skills/ml-stack-training/agents/openai.yaml
- skills/ml-stack-training/references/checkpointing.md
- skills/ml-stack-training/references/job-handoff.md
- skills/ml-stack-training/references/method-selection.md
- skills/ml-stack-training/references/sft-dpo-grpo.md
- skills/ml-stack-training/references/vision-training.md
Every link opens the file at its source, pinned to the revision this page describes.