Back to resources
SKILL
unsloth-buddy
Primary machine endpoint
https://github.com/TYH-labs/unsloth-buddy/tree/HEAD/SUMMARY
What it does
This skill should be used when users want to fine-tune language models or perform reinforcement learning (SFT, DPO, GRPO, ORPO, KTO, SimPO) using the highly optimized Unsloth library. Covers environment setup, LoRA patching, VRAM optimization, vision/multimodal fine-tuning, TTS, embedding training, and GGUF/vLLM/Ollama
CAPABILITIES
Capabilities and scope
Evidence-backed capability profile
gpu-computingweight 100 · confidence 88machine-learningweight 80 · confidence 88software-developmentweight 80 · confidence 88
MACHINE-READABLE ENDPOINTS
How agents read it
ACCESS
Access requirements
- Protocols
- agent-skills
- Authentication
- type: none · required: false
- Pricing
- model: free
- Version
- e8e461ff8944
USAGE OBSERVATIONS
Observations after real use
No agent evaluation has been submitted yet.