{"repo":"TYH-labs/unsloth-buddy","free":true,"listed":false,"github":"https://github.com/TYH-labs/unsloth-buddy","clone":"git clone https://github.com/TYH-labs/unsloth-buddy.git","description":"Zero-friction LLM fine-tuning skill for Claude Code, Gemini CLI & any ACP agent. Unsloth on NVIDIA · TRL+MPS/MLX on Apple Silicon. Automates env setup, LoRA training (SFT, DPO, GRPO, vision), post-hoc GRPO log diagnostics, evaluation, and export end-to-end. Part of the Gaslamp AI platform.","language":"Python","stars":273,"topics":["apple-silicon","claude-code","dpo","fine-tuning","grpo","huggingface","lora","qlora","rlhf","sft"],"license":"MIT","category":"machine-learning","readme_excerpt":"unsloth-buddy /unsloth-buddy I have 500 customer support Q&As and want to fine-tune a summarization model. I only have a MacBook Air. English 简体中文 繁體中文 --- What is this? The self-evolving fine-tuning agent. It talks like a colleague, learns your setup's quirks over time, and orchestrates the full lifecycle: from data formatting and model selection to training, validation, and deployment. Runs on NVIDIA GPUs via Unsloth, natively on Apple Silicon via mlx-tune, and on free cloud GPUs via colab-mcp. Part of the Gaslamp AI development platform — docs. --- One sentence, one fine-tuned model. One run, one step smarter. One conversation, eight phases, one deployable model — and a smarter agent next time. --- Quick Start This skill includes sub-skills and utility scripts — install the full repository, not a single file. Claude Code (recommended) Then describe what you want to fine-tune. The skill activates automatically. Gemini CLI Any agent supporting the Agent Skills standard --- How is it different? Most tools assume you already know what to do. This one doesn't — and it learns from every project you run. Your concern What actually happens --- --- \"I don't know where to start\" A 2-question interview locks in task, audience, and data — then recommends the right model, hardware, and method \"I don't have data, or it's in the wrong format\" A dedicated data phase acquires, generates, or reformats data to exactly match the trainer's required schema \"SFT? DPO? GRPO? Which one?\" Maps your","default_branch":null,"files":null,"tree":[],"storefront":"/r/TYH-labs","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/TYH-labs/unsloth-buddy/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}