Skip to main content
📖 The AI Tool Bible

AI prompt management

Editorial picks for "prompt management tool".

34 tools

All Evaluation →
BR

Braintrust

Featured
Evaluation · Platform (any LLM)
8.9

Eval, monitor, and improve AI products end-to-end.

Freemium· Starter: $0 · Pro: $249 · Enterprise: Custom pricingevalsmonitoring
LA

LangSmith

Evaluation · Platform (any LLM)
8.7

LangChain's eval + observability platform.

Freemium· Developer: $0 · Plus: $39 · Enterprise: Custom pricingLLM tracingevals
WB

Weights & Biases

Evaluation · Platform (any LLM)
8.4

The ML experiment tracker, now with LLM eval features.

Freemium· Free: $0/mo · Pro: Starts at $60/month · Enterprise: Custom plans · Personal: $0/mo · Advanced Enterprise: Custom planML experimentsLLM eval
HE

Helicone

Evaluation · Platform (any LLM)
8.3

Open-source LLM observability — one-line proxy install.

Freemium· Free 100k req/mo; Pro from $25/moobservabilitycost tracking
HU

Humanloop

Evaluation · Platform (any LLM)
8.2

Prompt management + evals for collaborative AI teams.

Paid· From $200/mo teamprompt managementteam collab
OP

OpenAI Playground

Writing · Multi-model (GPT-4o, GPT-4.1, o-series, DALL-E, Whisper, TTS)
8.2

OpenAI's official browser sandbox for prompting, tuning, and testing every model on the platform before you ship API code.

Paid· Basic: $10 · Pro: $20 · Enterprise: Contact salesprompt-engineeringmodel-comparison
AA

Athina AI

Evaluation · Multi-model
8.1

Collaborative LLM evaluation and observability platform for teams shipping AI features to production.

Freemium· Starter free (10k logs/mo); Pro & Enterprise customllm-evaluationprompt-management
HO

HoneyHive

Evaluation · Multi-model
8.1

OpenTelemetry-native observability and evaluation platform for LLM agents in production.

Freemium· Free tier available; paid/enterprise tiers via salesagent-observabilityllm-evaluation
ML

MLflow

Evaluation · Multi-model
8.1

Open-source platform for tracking, evaluating, and deploying ML models and LLM applications.

Free· Free and open source (Apache 2.0); managed offering via Databricksllm-evaluationexperiment-tracking
WB

W&B Weave

Evaluation · Multi-model
8.1

Production observability, tracing, and evaluation for LLM and agent systems from the Weights & Biases stack.

Freemium· Free tier available; paid and enterprise plans via W&Bllm-tracingagent-observability
PR

PromptHub

Writing · Multi-model (OpenAI, Anthropic, Google, Meta, Mistral, Bedrock, Azure)
8.0

Git-style prompt management, testing, and deployment platform for teams running multiple LLMs in production.

Freemium· Free signup; paid team plans (contact sales / in-app)prompt-managementprompt-versioning
RI

Rivet

Agents · Multi-model
8.0

Open-source visual IDE for building and debugging LLM agent graphs.

Free· Free and open source (MIT)agent-orchestrationprompt-chaining
PR

PromptLayer

Evaluation · Platform (any LLM)
7.9

Lightweight prompt logging + management for OpenAI/Claude apps.

Freemium· Free: $0 · Pro: $49 · Team: $500 · Enterprise: Customprompt loggingversioning
LA

Langfuse

Evaluation · Model-agnostic
7.3

Open-source LLM observability, prompt management, and evaluation in one platform.

Freemium· Free self-host & Hobby tier; Core $29/mo, Pro $199/mo, Enterprise $2,499/mollm-observabilityprompt-management
PO

Portkey

Agents · Multi-model
7.1

Production LLM gateway with observability, guardrails, and prompt management for teams shipping AI in anger.

Freemium· Developer: Free Forever · Production: $49/month · Enterprise: Custom Pricingllm-gatewayobservability
PF

Prompt Foundry

Evaluation · OpenAI + Anthropic (multi-model)
7.1

Prompt management and side-by-side LLM evaluation for OpenAI and Anthropic models.

Freemium· Free tier (10 prompts, 500 evals/mo); Pro $15/user/mo; Enterprise customprompt-managementmodel-comparison
PA

Puzzlet AI

Agents · Multi-model
7.1

Git-native prompt management and observability platform for teams shipping LLM applications.

Freemium· Basic: $20 · Pro: $50 · Enterprise: Contact salesprompt-managementllm-observability
RF

Respan (formerly Keywords AI)

Evaluation · Multi-model (500+ via gateway)
7.1

LLM engineering platform combining a multi-model gateway with tracing, evals, and prompt management.

Freemium· Free tier; paid plans (pricing not public); enterprise on requestllm-observabilityprompt-management
AI

AICamp

Writing · Multi-model (GPT-5.2, Claude 4, Gemini)
7.0

Team workspace that puts GPT, Claude, and Gemini behind one admin console with shared prompts, agents, and usage controls.

Freemium· Starter free (3 users, $5 credits); Business $10/user/mo; Enterprise customteam-ai-workspacemulti-model-chat
LA

LangFast

Evaluation · Multi-model
7.0

No-signup LLM playground for testing, comparing, and versioning prompts against your own API keys.

Paid· One-time lifetime ~$60-$120; 14-day money-backprompt-testingprompt-versioning
LM

LMQL

Coding · Multi-model (OpenAI, Hugging Face Transformers, llama.cpp)
7.0

A query language for LLMs that bolts types, templates, and constraints onto prompting.

Free· Free and open source (Apache-style); self-host or use with your own model API keysconstrained-decodingstructured-output
PR

PromptBase

Writing · Multi-model
7.0

Marketplace for buying, selling, and running prompts across the major image and text models.

Freemium· PromptBase Select: $14prompt-marketplacemidjourney-prompts
PR

Prompteams

Writing
7.0

Git-style version control and testing for LLM prompts, with auto-generated APIs that ship updates without redeploys.

Freemium· Starter: 100% Free · Enterprise: Customprompt-managementprompt-versioning
SC

Superpower ChatGPT

Agents · GPT (via ChatGPT UI)
7.0

Chrome extension that bolts folders, prompt libraries, and bulk export onto the ChatGPT web UI.

Freemium· Basic: $10 · Pro: $20 · Enterprise: Contact saleschatgpt-organizationprompt-management
AG

Agenta

Evaluation · Multi-model
6.9

Open-source LLMOps platform for prompt engineering, evaluation, and observability in one workspace.

Freemium· Hobby: $0 forever · Pro: $29 /month · Business: $299 /month · Enterprise: Custom · Open source: Free foreverprompt-engineeringllm-evaluation
IZ

Izlo

Agents · Model-agnostic
6.9

Prompt management platform with version control, collaboration, and an API for production deployment.

Paid· Solo: $20 · Pro: $25 per user/month · Enterprise: $39 per user/monthprompt-managementversion-control
TR

TreeScale

Agents · Multi-model
6.9

No-code platform that wraps LLM prompt chains into deployable, integration-ready APIs.

Freemium· Free tier to publish first LLM app; paid tiers on topllm-api-deploymentprompt-chaining
PA

Parea AI

Evaluation · Multi-model
6.8

LLM evaluation, observability, and prompt management platform for teams shipping production AI apps.

Freemium· Free (2 seats, 3k logs/mo); Team $150/mo; Enterprise customllm-evaluationprompt-management
PR

PromptPerfect

Writing · Multi-model (GPT, Claude, Stable Diffusion, Midjourney targets)
6.8

Prompt optimizer from Jina AI that rewrites and stress-tests prompts across major LLMs.

Freemium· Free: Free · Pro: $19.99 · Pro Max: $99.99 · Enterprise: Contact salesprompt-optimizationprompt-engineering
AP

AI Prompt Genius

Writing
6.6

Open-source Chrome extension that turns your browser into a searchable, taggable prompt library for any AI chatbot.

Free· Free and open-sourceprompt-managementprompt-library
AP

AI Prompt Genius

Writing
6.5

Open-source Chrome extension for building a searchable, tagged library of reusable AI prompts.

Free· Free and open-sourceprompt-libraryprompt-templates
MP

Magic Potion

Writing · Multi-model
6.5

Visual drag-and-drop prompt editor for crafting, organizing, and reusing prompts across OpenAI, Anthropic, and Google models.

Freemiumprompt-engineeringprompt-library
FA

Fabric

Agents · Model-agnostic: OpenAI GPT-4o/GPT-4.1, Anthropic Claude (incl. Opus 4.7), Google Gemini, Azure OpenAI, Bedrock, Vertex AI, plus local Ollama and LM Studio models.

An open-source framework for augmenting humans with AI, one composable prompt at a time.

Free· Free and open-source (MIT). Users bring their own API keys and pay each LLM provider directly; local models via Ollama or LM Studio incur no per-token cost.YouTube video summarisation and wisdom extractionLong-form article and PDF summarisation
LA

LangWatch

Evaluation · Model-agnostic; supports OpenAI, Anthropic, AWS Bedrock, Azure OpenAI, Vertex AI, and any OpenTelemetry-instrumented LLM

Simulation-based testing, evaluation, and observability for LLM agents

Freemium· Developer: €0 · Growth: €29/ core-seat / month · Enterprise: CustomLLM agent regression testing in CIRAG answer-quality evaluation