Skip to main content
📖 The AI Tool Bible

Best AI tools for distillation

37 tools in the Fine-tuning category, filtered to distillation.

All Fine-tuning
Together AI preview image
Together AI logo

Together AI

Featured
Fine-tuning · Llama / Mistral / Qwen / DeepSeek and others
8.6

Fine-tune & serve open-weight models (Llama, Mistral, DeepSeek).

Paid· Pay-per-token; fine-tuning per-tokenopen modelsfine-tuning
Modal preview image
Modal logo

Modal

Fine-tuning · Infrastructure (any model you can host)
8.7

Serverless GPUs and infra for training & serving ML.

Freemium· $30/mo free credits; pay-as-you-go GPU ratesserverless GPUfine-tuning
Replicate preview image
Replicate logo

Replicate

Fine-tuning · Thousands of community + first-party models
8.5

One-API platform for running and fine-tuning open-source models.

Paid· Pay-per-second of GPU timemodel hostingfine-tuning
OpenAI Fine-tuning preview image
OpenAI Fine-tuning logo

OpenAI Fine-tuning

Fine-tuning · GPT-4o-mini / GPT-3.5
8.4

Fine-tune GPT-4o-mini and friends on your own data.

Paid· Basic: $10 · Pro: $25 · Enterprise: Contact salesstyleformat
Llama preview image
Llama logo

Llama

Fine-tuning · Llama 4 (Maverick, Scout), Llama 3.3/3.2/3.1
8.3

Meta's open-weight LLM family covering 1B mobile models up to 405B frontier and natively multimodal 10M-context Llama 4 variants.

Freemium· Basic: $15 · Pro: $30 · Enterprise: $100self-hosted-llmfine-tuning
RunPod preview image
RunPod logo

RunPod

Fine-tuning · Bring-your-own (any open-weight or custom model)
8.3

On-demand GPU cloud and serverless inference platform built specifically for AI workloads.

Paid· Pod: $7.39/hr · Pod: $4.39/hr · Pod: $5.89/hr · Pod: $1.99/hr · Pod: $3.19/hrllm-fine-tuninggpu-rental
vLLM preview image
vLLM logo

vLLM

Fine-tuning · Multi-model (open-weight LLMs: Llama, Qwen, DeepSeek, Mistral, Gemma, Phi, etc.)
8.3

Open-source high-throughput inference engine for serving LLMs with PagedAttention and continuous batching.

Free· Free and open-source (Apache 2.0); self-hosted infrastructure costs applyllm-servingself-hosted-inference
CoreWeave preview image
CoreWeave logo

CoreWeave

Fine-tuning · DeepSeek
8.2

AI-native GPU cloud built for large-scale training, fine-tuning, and inference on NVIDIA hardware.

Enterprise· NVIDIA GB300 NVL72: Contact sales · NVIDIA GB200 NVL72: $42.00 · NVIDIA HGX B300: Contact sales · NVIDIA HGX B200: $68.80 · NVIDIA RTX PRO 6000 Blackwell Server Edition: $20.00model-trainingfine-tuning
Ludwig preview image
Ludwig logo

Ludwig

Fine-tuning · Multi-model (PyTorch + HuggingFace Transformers)
8.2

Declarative, YAML-driven deep learning framework for fine-tuning LLMs and multi-modal models without writing training loops.

Free· Free, Apache 2.0 open sourcellm-fine-tuningmulti-modal-training
OpenPipe preview image
OpenPipe logo

OpenPipe

Fine-tuning · Llama, Mistral, Qwen and other open-weight base models
8.2

Fine-tuning and reinforcement learning platform for turning expensive prompts into cheap, fast, task-specific models.

Freemium· Free tier available; usage-based pricing for training and hosted inference; enterprise plans on requestllm-cost-reductionfine-tuning
SGLang preview image
SGLang logo

SGLang

Fine-tuning · Multi-model (DeepSeek, Qwen, Llama, Mistral, GLM, GPT-OSS)
8.2

Open-source high-throughput inference engine for LLMs and multimodal models with OpenAI-compatible serving.

Free· Free, open-source (Apache 2.0); self-hosted infra cost onlyllm-servingmultimodal-inference
Unsloth preview image
Unsloth logo

Unsloth

Fine-tuning · Llama, Mistral, Gemma, Qwen, GLM (multi-model)
8.2

Open-source LLM fine-tuning toolkit with custom kernels that train 2-30x faster and use up to 90% less VRAM.

Freemium· Free open-source; Pro and Enterprise contact saleslora-finetuningqlora
Hugging Face AutoTrain preview image
Hugging Face AutoTrain logo

Hugging Face AutoTrain

Fine-tuning · Multi-model (Hugging Face Hub)
8.1

No-code fine-tuning and training pipeline that spins up state-of-the-art models on the Hugging Face Hub.

Paid· Per-minute billing based on hardware tier; self-hosted OSS version is freellm-fine-tuningtext-classification
Lambda preview image
Lambda logo

Lambda

Fine-tuning · NVIDIA VR200 NVL72, NVIDIA GB300 NVL72, NVIDIA HGX B200, NVIDIA HGX B300, NVIDIA H100
8.1

On-demand NVIDIA GPU cloud built specifically for training, fine-tuning, and serving large AI models.

Paid· Basic: $10 · Pro: $20 · Enterprise: Contact salesllm-trainingfine-tuning
Optuna preview image
Optuna logo

Optuna

Fine-tuning
8.1

Open-source Python framework for automated hyperparameter optimization across any ML stack.

Free· Free and open source (MIT)hyperparameter-tuningml-experiment-tracking
Ray Tune preview image
Ray Tune logo

Ray Tune

Fine-tuning
8.1

Open-source Python library for distributed hyperparameter tuning at any scale.

Free· Open-source (Apache 2.0); managed via Anyscale offers a $100 starting credithyperparameter-tuningdistributed-training
Together AI Fine-tuning preview image
Together AI Fine-tuning logo

Together AI Fine-tuning

Fine-tuning · Multi-model (any Hugging Face open-source model)
8.1

Managed fine-tuning platform for open-source LLMs and vision models with LoRA, full fine-tuning, and RL support.

Paid· Usage-based; cost estimator in-product, no public price listllm-fine-tuningvision-fine-tuning
Edge Impulse preview image
Edge Impulse logo

Edge Impulse

Fine-tuning · Multi-model (TF Lite Micro, custom DSP blocks)
8.0

End-to-end platform for training and deploying ML models on microcontrollers, sensors, and other edge hardware.

Freemium· Developer: $0edge-aitinyml
Anyscale preview image
Anyscale logo

Anyscale

Fine-tuning · Infrastructure (any model)
7.9

Ray-powered platform for training, serving, and scaling LLMs.

Paid· Enterprise / contact salesdistributed trainingRay
Fireworks AI preview image
Fireworks AI logo

Fireworks AI

Fine-tuning · Multi-model (DeepSeek, Qwen, GLM, Kimi, Gemma, Minimax, others)
7.9

Production inference and fine-tuning platform for open-source LLMs, tuned for speed and enterprise economics.

Freemium· Free signup credits; pay-per-token from ~$0.14/M in; enterprise reserved capacity on requestllm-fine-tuningserverless-inference
Lamini preview image
Lamini logo

Lamini

Fine-tuning · Lamini (built on open base models)
7.7

Memory-tuning platform for grounding LLMs in your facts.

Paid· Enterprise / contact salesenterprise FTfactual recall
FedML preview image
FedML logo

FedML

Fine-tuning · Bring-your-own (PyTorch, Hugging Face)
7.3

Distributed training, fine-tuning, and serving platform with federated learning roots.

Freemium· Open-source library free; managed GPU usage pay-as-you-gofine-tuningdistributed-training
Pachyderm preview image
Pachyderm logo

Pachyderm

Fine-tuning
7.3

Kubernetes-native data versioning and pipeline engine for reproducible ML at petabyte scale.

Freemium· Basic: $10 · Pro: $30 · Enterprise: Contact salesdata-versioningml-pipelines
LLaMA Factory preview image
LLaMA Factory logo

LLaMA Factory

Fine-tuning · Multi-model (LLaMA, Mistral, Qwen, Gemma, Phi, LLaVA, ChatGLM, Yi)
7.2

Open-source, no-code WebUI for fine-tuning 100+ open LLMs with LoRA, QLoRA, DPO, and PPO.

Free· Free, open-source (Apache-2.0); self-hostedlora-fine-tuningqlora
Paperspace Gradient preview image
Paperspace Gradient logo

Paperspace Gradient

Fine-tuning · Bring-your-own (PyTorch, TensorFlow, Hugging Face)
7.2

End-to-end MLOps platform with GPU notebooks, training jobs, and model deployment, now folded into DigitalOcean.

Freemium· Free: $0 · Pro: $8 · Growth: $39 · T0: $0 · T1: $12model-trainingfine-tuning
H2O AutoML preview image
H2O AutoML logo

H2O AutoML

Fine-tuning · H2O-3 (GBM, XGBoost, GLM, DRF, Deep Learning, Stacked Ensembles)
7.1

Open-source automated machine learning that handles feature engineering, model selection, and stacked ensembling out of the box.

Free· Free and open-source (Apache 2.0); paid Driverless AI sold separatelyautomltabular-ml
Scale GenAI Platform preview image
Scale GenAI Platform logo

Scale GenAI Platform

Fine-tuning · Multi-model (OpenAI, Google, Meta, Mistral)
7.1

Enterprise agent platform from Scale AI that connects your data, orchestrates multi-agent workflows, and learns from human feedback inside your own VPC.

Enterprise· Contact sales; enterprise contracts onlyenterprise-agentsrag-over-internal-data
W&B Sweeps preview image
W&B Sweeps logo

W&B Sweeps

Fine-tuning · Multi-model (Llama, DeepSeek, Qwen, Kimi)
7.1

Hyperparameter optimization from Weights & Biases with Bayesian search and Hyperband early stopping.

Freemium· Free: $0/mo · Pro: $60/month, billed monthly · Enterprise: Custom plans · Personal: $0/mo · Advanced Enterprise: Custom planhyperparameter-tuningbayesian-optimization
Forefront preview image
Forefront logo

Forefront

Fine-tuning · Multi-model (Mistral-7B, Mixtral, Phi-2)
7.0

Fine-tune and serve open-source LLMs on your own data without managing GPUs.

Paid· Basic: $20 · Pro: $50 · Enterprise: Contact salesfine-tuningopen-source-llms
ONNX preview image
ONNX logo

ONNX

Fine-tuning
7.0

Open standard for representing and exchanging machine learning models across frameworks and runtimes.

Free· Free and open source (Apache-2.0); Linux Foundation AI projectmodel-interchangeedge-deployment
Apache SINGA preview image
Apache SINGA logo

Apache SINGA

Fine-tuning
6.9

Apache-licensed distributed deep learning library focused on scalable training across GPUs and nodes.

Free· Free, Apache 2.0 licenseddistributed trainingdeep learning research
DagsHub preview image
DagsHub logo

DagsHub

Fine-tuning
6.8

GitHub-style collaboration platform for ML datasets, experiments, and models with MLflow and DVC under the hood.

Freemium· Individual: $0 per user/month · Team: $119 per user/month · Enterprise: Custom quoteexperiment-trackingdata-versioning
Velda preview image
Velda logo

Velda

Fine-tuning
6.7

Serverless GPU orchestration that runs AI training and batch jobs without Docker or Kubernetes.

Freemium· Free monthly credits on Velda Cloud; Enterprise contact salesdistributed-trainingbatch-inference
AutotuneLLM preview image
AutotuneLLM logo

AutotuneLLM

Fine-tuning

An open-source optimization layer that sits between your app and Ollama to squeeze more performance out of local LLMs.

Free· Free and open source (MIT licensed).Local LLM inference on Apple SiliconReducing KV cache RAM for Ollama models
Colossal-AI preview image
Colossal-AI logo

Colossal-AI

Fine-tuning · Framework-agnostic; used with LLaMA, GPT, Stable Diffusion, ViT, and other PyTorch-based open-weight models

Making large AI models cheaper, faster, and more accessible through distributed training

Free· Open-source (Apache 2.0). Enterprise support, consulting, and managed training services available from HPC-AI Technology on request.LLM pretraining across multi-node GPU clustersFull-parameter and LoRA fine-tuning of open-weight LLMs
Language Model Builder preview image
Language Model Builder logo

Language Model Builder

Fine-tuning · In-house small transformer models trained by the user; exports to safetensors

Learn how LLMs work by building one on your Mac

Free· Free macOS download. No account, subscription, or fees. Mac App Store version listed as coming soon.Learn transformer internals hands-onPre-train a small language model locally
PyTorch Lightning preview image
PyTorch Lightning logo

PyTorch Lightning

Fine-tuning · Framework-agnostic — trains any PyTorch model (transformers, CNNs, diffusion, RL nets, etc.)

The deep learning framework for professional AI researchers and ML engineers

Free· Free and open source (Apache 2.0). Optional paid compute available via the Lightning AI Studio platform.Multi-GPU LLM fine-tuningComputer vision model training