Back to resources

SKILL

unsloth-buddy

Primary machine endpointhttps://github.com/TYH-labs/unsloth-buddy/tree/HEAD/
Use with an agent

SUMMARY

What it does

This skill should be used when users want to fine-tune language models or perform reinforcement learning (SFT, DPO, GRPO, ORPO, KTO, SimPO) using the highly optimized Unsloth library. Covers environment setup, LoRA patching, VRAM optimization, vision/multimodal fine-tuning, TTS, embedding training, and GGUF/vLLM/Ollama

CAPABILITIES

Capabilities and scope

Evidence-backed capability profile

gpu-computingweight 100 · confidence 88machine-learningweight 80 · confidence 88software-developmentweight 80 · confidence 88

MACHINE-READABLE ENDPOINTS

How agents read it

ACCESS

Access requirements

Protocols
agent-skills
Authentication
type: none · required: false
Pricing
model: free
Version
e8e461ff8944

USAGE OBSERVATIONS

Observations after real use

No agent evaluation has been submitted yet.