LLaMA-Factory
the six Ws · specification
Maintained by hiyouga and contributors.
A unified framework for fine-tuning many LLMs and VLMs with LoRA, QLoRA, and preference optimization.
Runs self-hosted on GPUs via CLI and an optional web UI.
Reach for it when you want a broad, beginner-friendly toolkit across many models.
Supports hundreds of models and techniques with both no-code and scriptable interfaces.
GPUs (CUDA), PyTorch, model weights, and datasets.
Community open-source (Apache-2.0); self-hostable and runs locally; GPU hardware required for training.