No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →
by TYH-labs · Claude Skill · ★ 270
Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h
🔒 Is unsloth-buddy safe to install? View the security audit →
unsloth-buddy /unsloth-buddy I have 500 customer support Q&As and want to fine-tune a summarization model. I only have a MacBook Air. <img src="https://img.shields.io/badge/Try%20It-1%20minute-black?style=for-the-badge" alt="Tr
| Stars | 270 |
| Forks | 13 |
| Language | Python |
| Category | Claude Skill |
| License | MIT |
| Quality Score | 70.2861770522735/100 |
| Open Issues | 1 |
| Last Updated | 2026-06-28 |
| Created | 2026-03-15 |
| Platforms | claude-code, cli, gemini, python |
| Est. Tokens | ~199k |
Looking for a unsloth-buddy alternative? If you're comparing unsloth-buddy with other claude skill tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.
动手学大模型全栈:CS336 中文精讲 · PyTorch 手搓 Transformer · 单卡复现 Pretrain/SFT/LoRA/DPO/GRPO/RLVR · PEFT/Agent/RAG 落地(含真实实验、
A framework for agentic tool use training with reinforcement learning
A travel agent based on Qwen2.5, fine-tuned by SFT + DPO/PPO/GRPO using traveling question-answer dataset, a m
Toolkit for fine-tuning, ablating and unit-testing open-source LLMs.
Bilingual (中文+EN) ML / LLM / diffusion / agent interview cheat sheets for AI 秋招 — generated by ARIS /interview
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
Explore other popular claude skill tools:
unsloth-buddy is Zero-friction LLM fine-tuning skill for Claude Code, Gemini CLI & any ACP agent. Unsloth on NVIDIA · TRL+MPS/MLX on Apple Silicon. Automates env setup, LoRA training (SFT, DPO, GRPO, vision), post-hoc. It is categorized as a Claude Skill with 270 GitHub stars.
unsloth-buddy is primarily written in Python. It covers topics such as apple-silicon, claude-code, dpo.
You can find installation instructions and usage details in the unsloth-buddy GitHub repository at github.com/TYH-labs/unsloth-buddy. The project has 270 stars and 13 forks, indicating an active community.
unsloth-buddy is released under the MIT license, making it free to use and modify according to the license terms.
The top alternatives to unsloth-buddy on Agent Skills Hub include hands-on-llm, ToolBrain, Travel-Agent-based-on-Qwen2-RLHF. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.
Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.
The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.
Sources & who's responsible: