AI cost tracking
Editorial picks for "track llm costs".
31 tools
LangSmith
LangChain's eval + observability platform.
Helicone
Open-source LLM observability — one-line proxy install.
AgentOps
Observability and debugging platform purpose-built for AI agents, with time-travel replay and cost tracking across 400+ LLMs.
Arize AI
Enterprise observability and evaluation platform for LLM agents and generative AI applications.
Athina AI
Collaborative LLM evaluation and observability platform for teams shipping AI features to production.
HoneyHive
OpenTelemetry-native observability and evaluation platform for LLM agents in production.
W&B Weave
Production observability, tracing, and evaluation for LLM and agent systems from the Weights & Biases stack.
Headroom
Open-source context compression layer that strips 70-95% of boilerplate before it hits your LLM.
Bifrost
Open-source AI gateway that unifies 1000+ models behind one OpenAI-compatible endpoint with failover, budgets, and MCP routing.
Kong AI Gateway
Enterprise API gateway extended to route, govern, and observe LLM and agent traffic across providers.
Langfuse
Open-source LLM observability, prompt management, and evaluation in one platform.
Opik
Open-source LLM observability and evaluation platform for debugging and monitoring AI agents in production.
Plano
Envoy-based data plane for AI agents that handles routing, guardrails, and observability outside your app code.
Portkey AI Gateway
Open-source AI gateway that routes a single API call across 1,600+ LLMs with caching, fallbacks, and observability.
Arthur
Open-source toolkit for testing, tracing, and monitoring production AI agents.
Manifest
Open-source LLM router that fans your agent traffic across providers and your existing AI subscriptions.
Portkey
Production LLM gateway with observability, guardrails, and prompt management for teams shipping AI in anger.
Puzzlet AI
Git-native prompt management and observability platform for teams shipping LLM applications.
Respan (formerly Keywords AI)
LLM engineering platform combining a multi-model gateway with tracing, evals, and prompt management.
AICamp
Team workspace that puts GPT, Claude, and Gemini behind one admin console with shared prompts, agents, and usage controls.
SystemPrompt
Self-hosted AI governance gateway that audits, gates, and logs every LLM call before it leaves your network.
TeamoRouter
Unified LLM gateway that brokers Claude, GPT, and Gemini through one API key with usage-based discounts.
Guild AI
Control plane for deploying, governing, and auditing AI agents in production.
Price Per Token
Daily-updated LLM API pricing comparison across 300+ models, with calculators, leaderboards, and a free MCP server.
Artificial Analysis
Independent benchmarking platform comparing AI models and inference providers across intelligence, speed, and cost.
Parea AI
LLM evaluation, observability, and prompt management platform for teams shipping production AI apps.
ClevAgent
Middleware that supervises AI coding agents in real time, catching wasteful actions and enforcing safety guardrails.
AI Meter
Local usage meter that turns AI coding-agent tokens into estimated electricity and water consumption.
ClickHouse
The open-source columnar database powering real-time analytics — and, increasingly, LLM observability and RAG backends.
Hydra
Local-first trust control plane that routes AI tasks to the cheapest model that clears your confidence bar.
LangWatch
Simulation-based testing, evaluation, and observability for LLM agents