Skip to content

Fine-tuningAxolotl AI

Axolotl

Open-source post-training for LLMs — LoRA to RL, all from one YAML config.

Category
Fine-tuning
Pricing
FREE
Platforms
CLILinux
Models
Multi-model
Verified
Jun 10, 2026

An open-source (Apache-2.0) framework that streamlines post-training for open-weight models: full fine-tuning, LoRA/QLoRA, preference tuning (DPO, IPO, KTO, ORPO), reinforcement learning (GRPO), reward modeling and quantization-aware training, configured through a single YAML file with no scripting. Wraps Hugging Face Transformers, PEFT, TRL and DeepSpeed, and supports dozens of model families including multimodal vision and audio models.

Capabilities 1

What it actually does — grouped by capability family.

  • Fine-tuning / training (primary capability)

Pros & cons

  • Apache-2.0 with 12k+ GitHub stars
  • One YAML config, no scripting needed
  • DPO/GRPO/QAT and multimodal support
  • Wraps Transformers, PEFT, TRL, DeepSpeed
  • Needs your own GPUs or cloud compute
  • Config surface can overwhelm beginners

Tags

Further reading

View all Fine-tuning
  • View Unsloth details
    Fine-tuningFREEMIUMOpen core

    Unsloth

    Unsloth AI

    Fine-tune open LLMs faster with far less VRAM.

    An open-source (Apache-2.0) framework for fine-tuning and running open-weight models with custom CUDA kernels — roughly 2x faster training and large VRAM savings, so 7B–13B models fit on a single consumer GPU. Free tier runs on Colab/Kaggle or locally; Pro and Enterprise tiers add multi-GPU and multi-node speedups. Exports to GGUF/Safetensors for llama.cpp, vLLM, and Ollama.

    LoRA, QLoRA, and full fine-tuning
    Multi-GPU speedups are paid tiers
    • fine-tuning
    • lora
    • open-source
    • training
  • View OpenPipe details
    Fine-tuningFREEMIUM

    OpenPipe

    OpenPipe

    Replace frontier-model spend with a fine-tuned small model.

    Captures your production OpenAI / Anthropic calls, builds a dataset, fine-tunes a small open-weights model on your traffic, then serves the swap behind your existing SDK. The pitch: 10x cost reduction at parity.

    Uses your production logs as training data
    Needs enough quality traffic to distill
    • fine-tuning
    • cost-reduction
    • drop-in
    • open-weights