Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods. Use when fine-tuning large models (7B-70B) with limited GPU memory, when you need to train <1% of parameters with minimal accuracy loss, or for multi-adapter serving. HuggingFace's official library integrated with transformers ecosystem.
fine-tuning-expert alternatives
6 skills from other authors that do a similar job to fine-tuning-expert by Jeffallan. The most recently updated is peft-fine-tuning (Jun 16, 2026).
Side-by-side comparison
| Skill | Stars | License | Agents | Last update | Category | Install |
|---|---|---|---|---|---|---|
| fine-tuning-expert by Jeffallan · this skill | 11,060 | MIT | Claude Code | Aug 7, 2026 | Data | Plugin marketplace |
| peft-fine-tuning by Orchestra-Research | 11,807 | MIT | Claude Code, Codex | Jun 16, 2026 | Workflow | Plugin marketplace |
| fine-tuning-with-trl by Orchestra-Research | 11,807 | MIT | Claude Code, Codex | Jun 16, 2026 | Productivity | Plugin marketplace |
| axolotl by Orchestra-Research | 11,807 | MIT | Claude Code, Codex | Jun 16, 2026 | Data | Plugin marketplace |
| grpo-rl-training by Orchestra-Research | 11,807 | MIT | Claude Code, Codex | Jun 16, 2026 | Productivity | Plugin marketplace |
| llama-factory by Orchestra-Research | 11,807 | MIT | Claude Code, Codex | Jun 16, 2026 | Frontend | Plugin marketplace |
| unsloth by Orchestra-Research | 11,807 | MIT | Claude Code, Codex | Jun 16, 2026 | Workflow | Plugin marketplace |
Alternatives are matched on overlapping names, descriptions, tags and SKILL.md sections, weighted by how rare each term is across the catalog. Data refreshed weekly from GitHub.
The alternatives
Fine-tune LLMs using reinforcement learning with TRL - SFT for instruction tuning, DPO for preference alignment, PPO/GRPO for reward optimization, and reward model training. Use when need RLHF, align model with preferences, or train from human feedback. Works with HuggingFace Transformers.
Expert guidance for fine-tuning LLMs with Axolotl - YAML configs, 100+ models, LoRA/QLoRA, DPO/KTO/ORPO/GRPO, multimodal support
Expert guidance for GRPO/RL fine-tuning with TRL for reasoning and task-specific model training
Expert guidance for fine-tuning LLMs with LLaMA-Factory - WebUI no-code, 100+ models, 2/3/4/5/6/8-bit QLoRA, multimodal support
Expert guidance for fast fine-tuning with Unsloth - 2-5x faster training, 50-80% less memory, LoRA/QLoRA optimization
Custom skills
None of these quite fit? Get a skill built for your exact workflow.
Tell us what you want your AI agent to do. We’ll help turn your workflow, standards or internal tools into a production-ready skill for Claude Code, Codex, Cursor or Gemini.
Frequently asked questions
- What are the best alternatives to fine-tuning-expert?
- The closest matches in the catalog are peft-fine-tuning by Orchestra-Research, fine-tuning-with-trl by Orchestra-Research, axolotl by Orchestra-Research. They were picked because their names, descriptions, tags and SKILL.md sections overlap most with fine-tuning-expert’s, and they come from other repositories.
- Which fine-tuning-expert alternative is most popular?
- peft-fine-tuning by Orchestra-Research has the most GitHub stars of the alternatives listed here (11,807). Stars belong to the whole repository, so compare how many skills each repository contains too.
- Are these alternatives free?
- Yes. All of them are free to install from GitHub, and 6 of the 6 alternatives declare a recognized open-source license.