Skip to main content
📖 The AI Tool Bible

AI tools tagged Powered By Llama

48 tools matching this tag.model

All tags →
LlamaIndex preview image
LlamaIndex logo

LlamaIndex

Featured
RAG · BYO (Claude / GPT / open)
8.7

Data framework for connecting LLMs to your data.

Freemium· Free open-source; LlamaCloud paidRAGdata ingestion
Together AI preview image
Together AI logo

Together AI

Featured
Fine-tuning · Llama / Mistral / Qwen / DeepSeek and others
8.6

Fine-tune & serve open-weight models (Llama, Mistral, DeepSeek).

Paid· Pay-per-token; fine-tuning per-tokenopen modelsfine-tuning
Cline preview image
Cline logo

Cline

Coding · Model-agnostic: Claude (Anthropic), GPT (OpenAI), Gemini (Google), DeepSeek, Grok, Mistral, Cerebras, plus local Ollama/LM Studio
8.7

Open-source agentic coding assistant that plans, edits, and runs code inside your IDE

Freemium· ClinePass: $9.99/monthMulti-file feature scaffoldingLarge-scale refactors
Elasticsearch Vector Search preview image
Elasticsearch Vector Search logo

Elasticsearch Vector Search

RAG · BYO embeddings (OpenAI, Cohere, Hugging Face, Mistral, Bedrock, Vertex, Azure) plus Elastic's built-in ELSER sparse model and E5 dense model
8.7

Hybrid vector + keyword search in the enterprise-grade Elasticsearch engine

Freemium· Resource based pricing: Pay as you go (monthly) or prepaid · Usage based pricing: Pay as you go (monthly) or prepaid · License based pricing: ?RAG chatbot over enterprise docsHybrid semantic + keyword product search
Snowflake Cortex preview image
Snowflake Cortex logo

Snowflake Cortex

RAG · Anthropic Claude, Meta Llama, Mistral Large 2, Snowflake Arctic
8.7

Generative AI and RAG built into the Snowflake data cloud

Enterprise· Standard: Contact sales · Enterprise: Contact sales · Business Critical: Contact sales · Virtual Private Snowflake: Contact salesEnterprise RAG chatbot over governed dataNatural-language SQL for business analysts
AWS Bedrock preview image
AWS Bedrock logo

AWS Bedrock

Agents · Multi-model: Anthropic Claude, Meta Llama, Mistral, Cohere, AI21, Amazon Nova/Titan, DeepSeek, Stability, OpenAI GPT
8.6

Build and scale generative AI applications with foundation models

Paid· Standard: Contact sales · Flex: Contact sales · Priority: Contact sales · Reserved: Contact salesEnterprise RAG chatbot over private documentsMulti-step tool-using agents via AgentCore
DataStax Astra DB preview image
DataStax Astra DB logo

DataStax Astra DB

RAG · Bring-your-own embeddings; integrates with OpenAI, Cohere, Hugging Face, Mistral, NVIDIA NIM, and Vertex AI via server-side vectorize
8.6

Serverless vector and document database for production RAG and AI agents

Freemium· Small On-Demand: Contact sales · Medium (Balanced): Contact sales · Medium (Storage Optimized): Contact sales · Large (Balanced): Contact sales · Large (Storage Optimized): Contact salesRAG chatbot over enterprise documentsAgent long-term memory store
Google Vertex AI preview image
Google Vertex AI logo

Google Vertex AI

Agents · Gemini 2.5 (Pro/Flash/Nano), Imagen, Veo, Chirp, plus Model Garden (Llama, Mistral, Claude via partner)
8.6

Google Cloud's unified platform for building, deploying, and scaling enterprise AI agents and models.

Paid· Image Data - Training (Classification): $3.465 / 1 hour · Image Data - Training (Object Detection): $3.465 / 1 hour · Image Data - Deployment and Online Prediction: $1.375 / 1 hour · Image Data - Batch Prediction: $2.222 / 1 hour · Tabular Data - Training (Classification/Regression): $21.252 / 1 hourEnterprise RAG chatbotMulti-agent customer service
IBM watsonx preview image
IBM watsonx logo

IBM watsonx

Agents · IBM Granite (3.x, Code, Time Series), Meta Llama 3.x, Mistral, plus other curated open models
8.6

Enterprise AI platform for building, deploying, and governing models and agents

Enterprise· watsonx.ai has a free tier on IBM Cloud with limited tokens; paid usage is metered per 1M tokens by model family (Granite, Llama, Mistral, etc.). watsonx.governance and watsonx.data are quoted per environment. Enterprise deals via IBM sales; on-prem/Cloud Pak for Data is separately licensed.Enterprise RAG chatbot over private documentsCustomer service agents with guardrails
MongoDB Atlas Vector Search preview image
MongoDB Atlas Vector Search logo

MongoDB Atlas Vector Search

RAG · Bring-your-own embeddings (OpenAI, Cohere, open models); native Voyage AI embeddings and rerankers
8.6

Vector search built into the operational database you're already using.

Freemium· Free: $0 · Flex: Up to $30 · Dedicated: Starts at $56.94RAG over enterprise documentsProduct and content recommendation engines
Glean preview image
Glean logo

Glean

Agents · Model-agnostic: routes across 35+ LLMs including GPT-4o, Claude 3.5/4 Sonnet, Gemini 1.5/2, Llama 3, Mistral, plus Glean in-house models
8.5

Work AI platform that unifies enterprise knowledge, search, and agents

Enterprise· Enterprise pricing only; commonly reported in the $40-50/user/month range with a floor typically starting around 100 seats. No public self-serve tier. Contact sales for a quote.Enterprise search across Slack, Drive, Confluence and JiraCompany-wide AI assistant with citations
Replicate preview image
Replicate logo

Replicate

Fine-tuning · Thousands of community + first-party models
8.5

One-API platform for running and fine-tuning open-source models.

Paid· Pay-per-second of GPU timemodel hostingfine-tuning
Palantir AIP preview image
Palantir AIP logo

Palantir AIP

Agents · Multi-model (GPT, Claude, Llama, customer-hosted)
8.4

Enterprise AI platform that grounds LLMs in your operational data and runs agents against real business systems.

Enterprise· Contact sales; typically bundled with Foundryenterprise-agentsoperational-ai
Yi (01.AI) preview image
Yi (01.AI) logo

Yi (01.AI)

Agents · Yi-Lightning (MoE), Yi-Large, Yi-1.5 (6B/9B/34B), Yi-VL, Yi-Coder — in-house 01.AI foundation models
8.4

Foundation models from 01.AI — open-weight Yi family plus frontier Yi-Lightning and Yi-Large

Freemium· Open-source Yi models free under permissive license; hosted API via platform.lingyiwanwu.com with pay-per-token pricing (Yi-Lightning positioned as a low-cost frontier tier; Yi-Large priced higher; exact per-token rates on the platform dashboard). Enterprise custom-training and consulting on quote.Self-hosted coding assistantBilingual English-Chinese chatbot
Zed preview image
Zed logo

Zed

Coding · Claude (Sonnet/Opus), GPT-4o/5, Gemini, Ollama-hosted local models, and Zed's own Zeta2 open-weight edit-prediction model
8.4

A high-performance, multiplayer code editor built in Rust with native AI agent workflows.

Freemium· Personal: $0 · Pro: $10 · Business: $30AI pair-programming with Claude or GPT agentsInline code refactoring via natural-language prompts
LangChain preview image
LangChain logo

LangChain

RAG · BYO (any major LLM)
8.3

The broad LLM application framework — chains, agents, retrievers.

Freemium· Free open-source; LangSmith paidgeneral LLM appsRAG
Llama preview image
Llama logo

Llama

Fine-tuning · Llama 4 (Maverick, Scout), Llama 3.3/3.2/3.1
8.3

Meta's open-weight LLM family covering 1B mobile models up to 405B frontier and natively multimodal 10M-context Llama 4 variants.

Freemium· Basic: $15 · Pro: $30 · Enterprise: $100self-hosted-llmfine-tuning
Llama 3 preview image
Llama 3 logo

Llama 3

Writing · Llama 3 / 3.1 (8B, 70B, 405B)
8.3

Meta's open-weights LLM family that put serious frontier-adjacent models in everyone's hands.

Free· Weights free under Meta Llama Community License; inference cost via self-hosting or 3rd-party providerschatlong-context reasoning
LM Studio preview image
LM Studio logo

LM Studio

Agents · Multi-model (gpt-oss, Qwen3, Gemma, DeepSeek-R1, Llama, others)
8.3

Desktop app for discovering, downloading, and running open-weight LLMs locally with an OpenAI-compatible server.

Freemium· Free: $0local-llm-inferenceprivate-chat
n8n preview image
n8n logo

n8n

Agents · Multi-model
8.3

Source-available workflow automation with first-class AI-agent and RAG building blocks.

Freemium· Starter: 20€ · Pro: 50€ · Business: 667€ · Enterprise: Contact Salesai-agentsworkflow-automation
Qwen preview image
Qwen logo

Qwen

Writing · Qwen3 / Qwen-Image / Qwen-MT / Qwen3Guard
8.3

Alibaba's open-weight foundation model family covering chat, vision, image generation, translation, and safety classification.

Freemium· Open weights free; hosted API priced per-token via Alibaba Cloud DashScopechatreasoning
Skyvern preview image
Skyvern logo

Skyvern

Agents · Multi-model (OpenAI, Anthropic, Gemini, Ollama)
8.3

AI browser agent that automates web workflows from natural-language instructions, with CAPTCHA and 2FA handling built in.

Freemium· Free: $0/month · Hobby: $29/month · Pro: $149/month · Enterprise: Custom/monthbrowser-automationdata-extraction
Vanna.ai preview image
Vanna.ai logo

Vanna.ai

RAG · Multi-model (Anthropic, OpenAI, Gemini, Ollama)
8.3

Open-source text-to-SQL agent that learns your schema and writes queries against your real warehouse.

Freemium· Explorer: $50 · Team: $500 · Enterprise: Contact salestext-to-sqlnatural-language-bi
vLLM preview image
vLLM logo

vLLM

Fine-tuning · Multi-model (open-weight LLMs: Llama, Qwen, DeepSeek, Mistral, Gemma, Phi, etc.)
8.3

Open-source high-throughput inference engine for serving LLMs with PagedAttention and continuous batching.

Free· Free and open-source (Apache 2.0); self-hosted infrastructure costs applyllm-servingself-hosted-inference
BentoML preview image
BentoML logo

BentoML

Agents · Multi-model
8.2

Open-source framework and managed platform for serving and scaling AI models in production.

Freemium· OSS free (Apache 2.0); managed Bento cloud has free tier + usage-based pricingmodel-servingllm-inference
Browser Use preview image
Browser Use logo

Browser Use

Agents · Multi-model (BYO LLM; Claude in hosted Box)
8.2

Open-source browser automation harness and cloud platform for LLM agents that drive real websites.

Freemium· PAYG: $0 · Dev: $29 · Business: $299 · Scaleup: $999web-automationscraping
Jan preview image
Jan logo

Jan

Writing · Multi-model (local open-weights + OpenAI/Claude/Gemini via API)
8.2

Open-source desktop ChatGPT alternative that runs local LLMs and routes to cloud providers from one app.

Free· Free and open source; bring-your-own keys for cloud modelslocal-llm-chatprivate-ai-assistant
LibreChat preview image
LibreChat logo

LibreChat

Writing · Multi-model (OpenAI, Anthropic, Google, AWS Bedrock, Azure, Ollama, and others)
8.2

Open-source, self-hostable ChatGPT-style frontend that brings every major LLM provider under one roof.

Free· Free and open source; self-hosted (you pay model providers for API usage)multi-model chatself-hosted chatgpt
LiveBench preview image
LiveBench logo

LiveBench

Evaluation · Multi-model
8.2

Contamination-free LLM benchmark that refreshes its questions monthly to keep frontier models honest.

Free· Free and open source; self-hosted evaluation runnerllm-benchmarkingmodel-selection
Msty preview image
Msty logo

Msty

Writing · Multi-model (Ollama, Claude, GPT, Gemini, others)
8.2

Privacy-first desktop AI workspace that runs local and cloud models side by side.

Freemium· Free: $0 · Aurum: ? · Aurum Lifetime: ? · Enterprise & Teams: Talk to uslocal-llm-chatmulti-model-comparison
OpenPipe preview image
OpenPipe logo

OpenPipe

Fine-tuning · Llama, Mistral, Qwen and other open-weight base models
8.2

Fine-tuning and reinforcement learning platform for turning expensive prompts into cheap, fast, task-specific models.

Freemium· Free tier available; usage-based pricing for training and hosted inference; enterprise plans on requestllm-cost-reductionfine-tuning
SGLang preview image
SGLang logo

SGLang

Fine-tuning · Multi-model (DeepSeek, Qwen, Llama, Mistral, GLM, GPT-OSS)
8.2

Open-source high-throughput inference engine for LLMs and multimodal models with OpenAI-compatible serving.

Free· Free, open-source (Apache 2.0); self-hosted infra cost onlyllm-servingmultimodal-inference
Unsloth preview image
Unsloth logo

Unsloth

Fine-tuning · Llama, Mistral, Gemma, Qwen, GLM (multi-model)
8.2

Open-source LLM fine-tuning toolkit with custom kernels that train 2-30x faster and use up to 90% less VRAM.

Freemium· Free open-source; Pro and Enterprise contact saleslora-finetuningqlora
Databricks Vector Search preview image
Databricks Vector Search logo

Databricks Vector Search

RAG · Multi-model (BYO embeddings or Databricks-hosted)
8.1

Managed hybrid vector search that lives inside the Databricks lakehouse and auto-syncs with your source tables.

Enterprise· Standard: $605 · Storage Optimized: $922rag-retrievalhybrid-search
Langflow preview image
Langflow logo

Langflow

Agents · Multi-model
8.1

Open-source visual builder for LangChain-style AI agents and RAG pipelines.

Freemium· Open-source free; hosted free tier + paid enterprise via DataStaxagent-prototypingrag-pipelines
LocalAI preview image
LocalAI logo

LocalAI

Writing · Multi-model (llama.cpp, diffusers, whisper, etc.)
8.1

Self-hosted OpenAI-compatible API for running LLMs, image, and audio models on your own hardware.

Free· Free and open source (MIT)local-llm-inferenceopenai-api-replacement
PyGPT preview image
PyGPT logo

PyGPT

Writing · Multi-model (GPT-5, Claude, Gemini, Grok, DeepSeek, Mistral, Ollama)
8.1

Open-source desktop AI assistant that wires every major LLM provider into one local app with agents, vision, and a code interpreter.

Free· Free and open-source (MIT); bring your own provider API keysdesktop-ai-assistantchat-with-files
Together AI Fine-tuning preview image
Together AI Fine-tuning logo

Together AI Fine-tuning

Fine-tuning · Multi-model (any Hugging Face open-source model)
8.1

Managed fine-tuning platform for open-source LLMs and vision models with LoRA, full fine-tuning, and RL support.

Paid· Usage-based; cost estimator in-product, no public price listllm-fine-tuningvision-fine-tuning
TruLens preview image
TruLens logo

TruLens

Evaluation · Multi-model (LLM-as-judge)
8.1

Open-source evaluation and tracing framework for LLM apps and agents, built on OpenTelemetry.

Free· Free, open source (Apache-licensed Python package)llm-evaluationrag-evaluation
W&B Weave preview image
W&B Weave logo

W&B Weave

Evaluation · Multi-model
8.1

Production observability, tracing, and evaluation for LLM and agent systems from the Weights & Biases stack.

Freemium· Free tier available; paid and enterprise plans via W&Bllm-tracingagent-observability
NovelCrafter preview image
NovelCrafter logo

NovelCrafter

Writing · Multi-model (BYO key: OpenAI, Anthropic, Gemini, Mistral, OpenRouter, Ollama)
8.0

Bring-your-own-key novel-writing workspace with a wiki-style codex and multi-model AI assistance.

Freemium· Basic: $10 · Pro: $20 · Enterprise: Contact salesnovel-writingworldbuilding
AnythingLLM preview image
AnythingLLM logo

AnythingLLM

RAG · Multi-model
7.9

Open-source desktop and self-hosted app that turns your documents into a private chat-and-agent workspace.

Freemium· Basic: $50/monthly · Pro: $99/monthly · Enterprise: Contact Usdocument-chatprivate-rag
Continue preview image
Continue logo

Continue

Coding · BYO (any OpenAI-compatible API + Ollama for local)
7.9

Open-source, self-hostable VS Code/JetBrains AI assistant.

Free· Free / open-source; you pay model costsself-hostedopen source
Meta AI preview image
Meta AI logo

Meta AI

Writing · Llama (Meta's open-weight family)
7.9

Meta's free consumer AI assistant powered by the Llama family of open-weight models.

Free· Free consumer product; no paid tierchat-assistantwriting-help
TypingMind preview image
TypingMind logo

TypingMind

Writing · Multi-model (GPT, Claude, Gemini, Mistral, DeepSeek, local)
7.7

BYOK chat frontend that puts GPT, Claude, Gemini, and local models behind one polished UI.

Paid· Standard: $39 · Extended: $79 · Premium: $99multi-model-chatprompt-library
oMLX preview image
oMLX logo

oMLX

Coding · Multi-model (Qwen, Llama, Mistral, Gemma, DeepSeek, MiniMax, GLM)
7.5

Native macOS LLM inference server built on MLX, with paged SSD KV caching for Apple Silicon agents.

Free· Free, Apache 2.0 open sourcelocal-llm-inferencecoding-agents
Langchain-Chatchat preview image
Langchain-Chatchat logo

Langchain-Chatchat

RAG · Multi-model (GLM-4, Qwen2, Llama 3, etc. via Xinference/Ollama/LocalAI/FastChat)
7.4

Self-hostable RAG and agent framework that wires LangChain to any local open-source LLM and a knowledge base.

Free· Apache-2.0 open source; self-hosted, infra costs onlyprivate-knowledge-baseoffline-rag
AstrBot preview image
AstrBot logo

AstrBot

Agents · Multi-model (OpenAI, Anthropic, Gemini, DeepSeek, Ollama, Dify, Coze)
7.3

Open-source agentic AI assistant that bridges chat platforms like Telegram, Discord, and QQ with any LLM and a 1000+ plugin ecosystem.

Free· Free, open-source (AGPL-3.0); self-hosted, you pay your own LLM API costs.chatbotgroup-chat-assistant