A parameter-efficient fine-tuning method that freezes the base model and trains small rank-decomposed matrices, cutting VRAM needs by 10-100x versus full fine-tuning.
fine-tuning
LoRA (Low-Rank Adaptation)
Related terms
Tools that implement LoRA (Low-Rank Adaptation)
Together AI
FeaturedFine-tuning · Llama / Mistral / Qwen / DeepSeek and others
8.6
Fine-tune & serve open-weight models (Llama, Mistral, DeepSeek).
Paid· Pay-per-token; fine-tuning per-tokenopen modelsfine-tuning
Fireworks AI
Fine-tuning · Multi-model (DeepSeek, Qwen, GLM, Kimi, Gemma, Minimax, others)
7.9
Production inference and fine-tuning platform for open-source LLMs, tuned for speed and enterprise economics.
Freemium· Free signup credits; pay-per-token from ~$0.14/M in; enterprise reserved capacity on requestllm-fine-tuningserverless-inference
Unsloth
Fine-tuning · Llama, Mistral, Gemma, Qwen, GLM (multi-model)
8.2
Open-source LLM fine-tuning toolkit with custom kernels that train 2-30x faster and use up to 90% less VRAM.
Freemium· Free open-source; Pro and Enterprise contact saleslora-finetuningqlora