AI model distillation
Editorial picks for "llm distillation".
48 tools

AWS Bedrock
Build and scale generative AI applications with foundation models

OpenPipe
Fine-tuning and reinforcement learning platform for turning expensive prompts into cheap, fast, task-specific models.

Fabric
An open-source framework for augmenting humans with AI, one composable prompt at a time.

Mem0
Persistent memory layer for AI agents and LLM apps

Warp
The agentic development environment, from the terminal up

Cline
Open-source agentic coding assistant that plans, edits, and runs code inside your IDE

Elasticsearch Vector Search
Hybrid vector + keyword search in the enterprise-grade Elasticsearch engine

Snowflake Cortex
Generative AI and RAG built into the Snowflake data cloud

MongoDB Atlas Vector Search
Vector search built into the operational database you're already using.

Glean
Work AI platform that unifies enterprise knowledge, search, and agents

Replicate
One-API platform for running and fine-tuning open-source models.

ChatGLM (Zhipu Qingyan)
Zhipu AI's bilingual ChatGLM assistant with agents, code, image and video generation

Palantir AIP
Enterprise AI platform that grounds LLMs in your operational data and runs agents against real business systems.

PyCaret
Low-code Python AutoML library that wraps scikit-learn, XGBoost, LightGBM and friends behind a few-line API.

Quivr
Open-source RAG framework for building custom AI assistants over your own documents in a few lines of Python.

Semantic Kernel
Microsoft's open-source SDK for wiring LLMs, plugins, and agents into enterprise .NET, Python, and Java apps.

Yi (01.AI)
Foundation models from 01.AI — open-weight Yi family plus frontier Yi-Lightning and Yi-Large

ClearML
End-to-end MLOps and GenAI platform with open-source experiment tracking and enterprise GPU orchestration.

Dataiku
Enterprise AI platform unifying data, ML, LLMs, and agents under one governed workflow.

H2O.ai
Enterprise AI platform combining AutoML, generative AI, and vertical agents for regulated industries.

Llama
Meta's open-weight LLM family covering 1B mobile models up to 405B frontier and natively multimodal 10M-context Llama 4 variants.

Llama 3
Meta's open-weights LLM family that put serious frontier-adjacent models in everyone's hands.

LM Studio
Desktop app for discovering, downloading, and running open-weight LLMs locally with an OpenAI-compatible server.

n8n
Source-available workflow automation with first-class AI-agent and RAG building blocks.

Qwen
Alibaba's open-weight foundation model family covering chat, vision, image generation, translation, and safety classification.

RunPod
On-demand GPU cloud and serverless inference platform built specifically for AI workloads.

Skyvern
AI browser agent that automates web workflows from natural-language instructions, with CAPTCHA and 2FA handling built in.

Vanna.ai
Open-source text-to-SQL agent that learns your schema and writes queries against your real warehouse.

vLLM
Open-source high-throughput inference engine for serving LLMs with PagedAttention and continuous batching.

BentoML
Open-source framework and managed platform for serving and scaling AI models in production.

Browser Use
Open-source browser automation harness and cloud platform for LLM agents that drive real websites.

Google AI Studio
Browser-based playground and API console for prototyping with Google's Gemini models.

Great Expectations
Open-source data quality framework for validating the datasets that feed your ML and analytics pipelines.

HyperFrames
Open-source video composition framework that lets AI agents build videos by writing HTML, CSS, and JS.

Jan
Open-source desktop ChatGPT alternative that runs local LLMs and routes to cloud providers from one app.

LibreChat
Open-source, self-hostable ChatGPT-style frontend that brings every major LLM provider under one roof.

LiveBench
Contamination-free LLM benchmark that refreshes its questions monthly to keep frontier models honest.

Ludwig
Declarative, YAML-driven deep learning framework for fine-tuning LLMs and multi-modal models without writing training loops.

Msty
Privacy-first desktop AI workspace that runs local and cloud models side by side.

SGLang
Open-source high-throughput inference engine for LLMs and multimodal models with OpenAI-compatible serving.

Sim
Open-source visual workspace for building, deploying, and monitoring AI agents.

Unsloth
Open-source LLM fine-tuning toolkit with custom kernels that train 2-30x faster and use up to 90% less VRAM.

Vellum
Personal AI assistant with persistent memory across email, calendar, and chat channels.

Writer
Enterprise generative AI platform built around in-house Palmyra LLMs for regulated, brand-consistent content.

Athina AI
Collaborative LLM evaluation and observability platform for teams shipping AI features to production.

Berkeley Function-Calling Leaderboard
Open benchmark from UC Berkeley that ranks LLMs on real-world tool-use and function-calling accuracy.

Cube
Semantic layer that grounds LLM agents in your real business metrics instead of letting them hallucinate SQL.

ElevenLabs Conversational AI
Production-grade voice agent platform layering ElevenLabs TTS, ASR, and LLM orchestration into a single deployable stack.