Skip to main content
📖 The AI Tool Bible

AI tools tagged Cloud Only

48 tools matching this tag.tech

All tags →
PI

Pinecone

Featured
RAG · Hosted vector DB (not an LLM)
8.8

Managed vector database for production-scale similarity search.

Freemium· Starter: Free · Builder: $20/month flat · Standard: $50/month min. usage · Enterprise: $500/month min. usagemanaged vector DBproduction RAG
TA

Together AI

Featured
Fine-tuning · Llama / Mistral / Qwen / DeepSeek and others
8.6

Fine-tune & serve open-weight models (Llama, Mistral, DeepSeek).

Paid· Pay-per-token; fine-tuning per-tokenopen modelsfine-tuning
MO

Modal

Fine-tuning · Infrastructure (any model you can host)
8.7

Serverless GPUs and infra for training & serving ML.

Freemium· $30/mo free credits; pay-as-you-go GPU ratesserverless GPUfine-tuning
SC

Snowflake Cortex

RAG · Anthropic Claude, Meta Llama, Mistral Large 2, Snowflake Arctic
8.7

Generative AI and RAG built into the Snowflake data cloud

Enterprise· Standard: Contact sales · Enterprise: Contact sales · Business Critical: Contact sales · Virtual Private Snowflake: Contact salesEnterprise RAG chatbot over governed dataNatural-language SQL for business analysts
DE

DALL·E 3

Image Generation · DALL·E 3
8.6

OpenAI's image model — strong on prompt adherence and text-in-image.

Freemium· Basic: $10 · Pro: $20postersinfographics
DA

DataStax Astra DB

RAG · Bring-your-own embeddings; integrates with OpenAI, Cohere, Hugging Face, Mistral, NVIDIA NIM, and Vertex AI via server-side vectorize
8.6

Serverless vector and document database for production RAG and AI agents

Freemium· Small On-Demand: Contact sales · Medium (Balanced): Contact sales · Medium (Storage Optimized): Contact sales · Large (Balanced): Contact sales · Large (Storage Optimized): Contact salesRAG chatbot over enterprise documentsAgent long-term memory store
MA

MongoDB Atlas Vector Search

RAG · Bring-your-own embeddings (OpenAI, Cohere, open models); native Voyage AI embeddings and rerankers
8.6

Vector search built into the operational database you're already using.

Freemium· Free: $0 · Flex: Up to $30 · Dedicated: Starts at $56.94RAG over enterprise documentsProduct and content recommendation engines
RE

Replicate

Fine-tuning · Thousands of community + first-party models
8.5

One-API platform for running and fine-tuning open-source models.

Paid· Pay-per-second of GPU timemodel hostingfine-tuning
SE

Seedream

Image Generation · ByteDance Seedream (4.0 / 4.5 / 5.0 Lite family, in-house)
8.5

ByteDance's unified text-to-image and image-editing model, served via Volcengine.

Paid· Usage-based via Volcengine / third-party gateways. Approx: Seedream 4.0 ~$0.025-$0.069 per image; Seedream 4.5 ~$0.035-$0.045 per image (fal, OpenRouter list ~$0.04 flat); Seedream 5.0 Lite ~$0.035 per image. Cheaper CN gateway rates (¥0.12-0.20/image) available. Enterprise SLA pricing on request.High-volume ad creative generationE-commerce product photography and variants
AI

Aider

Coding · BYO (Claude / GPT-4 / Gemini / DeepSeek)
8.4

Terminal-based AI pair programmer that writes commits.

Free· Free / open-source; you pay the underlying LLM API costsCLIgit workflow
OF

OpenAI Fine-tuning

Fine-tuning · GPT-4o-mini / GPT-3.5
8.4

Fine-tune GPT-4o-mini and friends on your own data.

Paid· Basic: $10 · Pro: $25 · Enterprise: Contact salesstyleformat
PA

Palantir AIP

Agents · Multi-model (GPT, Claude, Llama, customer-hosted)
8.4

Enterprise AI platform that grounds LLMs in your operational data and runs agents against real business systems.

Enterprise· Contact sales; typically bundled with Foundryenterprise-agentsoperational-ai
WE

Weaviate

RAG · Hosted vector DB (not an LLM)
8.4

Open-source vector DB with hybrid search and modules.

Freemium· Free: $0 · Flex: $45 · Premium: $400self-hosted RAGhybrid search
RU

RunPod

Fine-tuning · Bring-your-own (any open-weight or custom model)
8.3

On-demand GPU cloud and serverless inference platform built specifically for AI workloads.

Paid· Pod: $7.39/hr · Pod: $4.39/hr · Pod: $5.89/hr · Pod: $1.99/hr · Pod: $3.19/hrllm-fine-tuninggpu-rental
SK

Skyvern

Agents · Multi-model (OpenAI, Anthropic, Gemini, Ollama)
8.3

AI browser agent that automates web workflows from natural-language instructions, with CAPTCHA and 2FA handling built in.

Freemium· Free: $0/month · Hobby: $29/month · Pro: $149/month · Enterprise: Custom/monthbrowser-automationdata-extraction
BU

Browser Use

Agents · Multi-model (BYO LLM; Claude in hosted Box)
8.2

Open-source browser automation harness and cloud platform for LLM agents that drive real websites.

Freemium· PAYG: $0 · Dev: $29 · Business: $299 · Scaleup: $999web-automationscraping
CO

CoreWeave

Fine-tuning · DeepSeek
8.2

AI-native GPU cloud built for large-scale training, fine-tuning, and inference on NVIDIA hardware.

Enterprise· NVIDIA GB300 NVL72: Contact sales · NVIDIA GB200 NVL72: $42.00 · NVIDIA HGX B300: Contact sales · NVIDIA HGX B200: $68.80 · NVIDIA RTX PRO 6000 Blackwell Server Edition: $20.00model-trainingfine-tuning
DA

Daytona

Agents
8.2

Secure, isolated sandboxes for running AI-generated code with sub-90ms cold starts.

Freemium· Pay-per-second from $0.000014/sec; $200 free creditagent-sandboxescode-interpreters
DV

Databricks Vector Search

RAG · Multi-model (BYO embeddings or Databricks-hosted)
8.1

Managed hybrid vector search that lives inside the Databricks lakehouse and auto-syncs with your source tables.

Enterprise· Standard: $605 · Storage Optimized: $922rag-retrievalhybrid-search
LA

Lambda

Fine-tuning · NVIDIA VR200 NVL72, NVIDIA GB300 NVL72, NVIDIA HGX B200, NVIDIA HGX B300, NVIDIA H100
8.1

On-demand NVIDIA GPU cloud built specifically for training, fine-tuning, and serving large AI models.

Paid· Basic: $10 · Pro: $20 · Enterprise: Contact salesllm-trainingfine-tuning
WB

W&B Weave

Evaluation · Multi-model
8.1

Production observability, tracing, and evaluation for LLM and agent systems from the Weights & Biases stack.

Freemium· Free tier available; paid and enterprise plans via W&Bllm-tracingagent-observability
EX

Exa

RAG · Proprietary neural + keyword search
8.0

Web search API built for AI agents, with structured outputs and token-efficient highlights.

Freemium· Free Tier: Free · Search: $7/1k requests · Agent: $0.012–$1.00/run · Contents: $1/1k pages per content type · Deep Search: $12–15/1k requestsagent-web-searchrag-retrieval
FA

Fal.ai

Image Generation · Multi-model (Flux, Stable Diffusion, video/audio models)
8.0

Serverless GPU inference platform optimized for fast diffusion and generative media APIs.

Paid· Usage-based; serverless from ~$1.89/GPU-hour, per-output pricing on model APIstext-to-imagetext-to-video
AQ

Amazon Q

Coding · Multi-model (Amazon Bedrock)
7.9

AWS's enterprise AI assistant for developers, analysts, and knowledge workers, wired into your company data.

Freemium· Amazon Q Business Lite: $3 · Amazon Q Business Pro: $20 · Amazon Q Developer Free Tier: Free · Amazon Q Developer Pro Tier: ? · Amazon Q in QuickSight Author: $24code-generationenterprise-search
AN

Anyscale

Fine-tuning · Infrastructure (any model)
7.9

Ray-powered platform for training, serving, and scaling LLMs.

Paid· Enterprise / contact salesdistributed trainingRay
DV

DVC

Coding
7.9

Git-style version control for datasets, ML models, and experiment pipelines.

Free· Free and open source; lakeFS Enterprise available for large-scale deploymentsdata-versioningml-experiment-tracking
MA

Manus

Agents · Multi-model
7.9

Generalist agent for research, code, and web tasks.

Paid· Credit-based; tiers from $19/moresearchweb tasks
DE

Devin

Agents · Multi-model (Claude / GPT configurable)
7.8

Cognition Labs' "autonomous software engineer" agent.

Paid· From $500/mo Coreautonomous codingticket resolution
AG

Agentset

RAG · Multi-model (Claude, OpenAI, Google, xAI, Cohere, Mistral, DeepSeek)
7.3

Production-ready RAG infrastructure with agentic search, citations, and model-agnostic plumbing.

Freemium· Free: $0 · Pro: $49 · Enterprise: Customdocument-qaagentic-search
AA

Azure AI Speech (Neural TTS)

Audio · Azure Neural TTS (plus HD and Azure OpenAI voices)
7.3

Microsoft's enterprise-grade neural text-to-speech with 100+ languages, custom brand voices, and SSML control.

Freemium· Free (F0): Free · Pay as You Go: Voice Live Prices: $- · Commitment Tiers – Standard: $- for 2,000 hourstext-to-speechvoice-cloning
EP

Epsilla

RAG · Multi-model
7.3

Agent-as-a-Service platform with managed RAG and a no-code builder for vertical enterprise AI.

Freemium· Free Tier: $0/month · Starter Tier: $29/month · Professional Tier: $249/month · AI Concierge: $2,499/month · Enterprise Tier: Custom/monthenterprise-ragai-agents
FE

FedML

Fine-tuning · Bring-your-own (PyTorch, Hugging Face)
7.3

Distributed training, fine-tuning, and serving platform with federated learning roots.

Freemium· Open-source library free; managed GPU usage pay-as-you-gofine-tuningdistributed-training
BE

Beam

Coding · openai/gpt-oss-20b
7.2

Serverless GPU infrastructure for AI workloads with sub-second cold starts and bring-your-own-cloud support.

Freemium· $30 free credit refreshed monthly; usage-based beyond thatgpu-inferenceagent-sandboxes
ME

Metabase

Agents · Multi-model
7.2

Open-source BI platform with a Metabot AI layer for natural-language querying over your warehouse.

Freemium· Open Source: Free · Starter: $90 · Pro: $517.50 · Enterprise: Custom pricingbusiness-intelligencenatural-language-querying
OL

Ollama

Coding · Multi-model (Llama, Qwen, Gemma, DeepSeek, Mistral, Phi, etc.)
7.2

The de facto runtime for running open-weights LLMs locally, now with a paid cloud tier for bigger models.

Freemium· Free local; Pro $20/mo; Max $100/molocal-llmself-hosted-inference
PG

Paperspace Gradient

Fine-tuning · Bring-your-own (PyTorch, TensorFlow, Hugging Face)
7.2

End-to-end MLOps platform with GPU notebooks, training jobs, and model deployment, now folded into DigitalOcean.

Freemium· Free: $0 · Pro: $8 · Growth: $39 · T0: $0 · T1: $12model-trainingfine-tuning
SV

SAS Viya

Agents · Multi-model
7.2

Enterprise-grade data and AI analytics platform with built-in governance, MCP server, and a Copilot for regulated industries.

Enterprise· Contact sales; 14-day free trialenterprise-analyticsai-governance
FA

Fiddler AI

Evaluation · Fiddler Centor (proprietary evaluators)
7.1

Enterprise AI observability and guardrails platform for monitoring agents, LLMs, and ML models in production.

Enterprise· Free: Free · Developer: $0.002 per trace · Enterprise: Contact salesllm-observabilityagent-monitoring
FP

FutureHouse Platform

RAG · Multi-model
7.1

Multi-agent AI research stack for scientists, with retrieval over 175M+ papers, patents, and trials.

Freemium· Basic: $10 · Pro: $25 · Enterprise: Contact salesscientific-literature-searchautonomous-research-agent
HY

Hyperbrowser

Agents · Model-agnostic (BYO LLM)
7.1

Cloud browser infrastructure built for AI agents that need to scrape, click, and navigate the live web.

Freemium· Free developer tier; paid usage-based plansagent-browser-controlweb-scraping
NR

Neu.ro

Agents · Multi-model
7.1

Infrastructure-agnostic MLOps platform for the full ML/DL lifecycle across public, hybrid, and on-prem clouds.

Enterprise· Contact sales; no public pricingmlopsmodel-training
RI

Rivestack

RAG · OpenAI embeddings (auto-embeddings)
7.1

Managed Postgres with pgvector on dedicated NVMe, pitched as a cheaper RAG backend than Pinecone or Supabase.

Freemium· Free: $0/month · Solo: $29/month · HA Cluster: Starting at $49/node/monthrag-backendvector-search
ST

StarOps

Coding · Multi-model
7.1

AI-native platform engineering engine that provisions and manages cloud infrastructure from natural-language prompts.

Freemium· Free tier; paid from $199/mo; custom enterpriseinfrastructure-automationkubernetes-management
WB

W&B Sweeps

Fine-tuning · Multi-model (Llama, DeepSeek, Qwen, Kimi)
7.1

Hyperparameter optimization from Weights & Biases with Bayesian search and Hyperband early stopping.

Freemium· Free: $0/mo · Pro: $60/month, billed monthly · Enterprise: Custom plans · Personal: $0/mo · Advanced Enterprise: Custom planhyperparameter-tuningbayesian-optimization
AQ

Amazon Q Developer

Coding · Multi-model (Amazon proprietary + partners)
7.0

AWS's in-IDE coding assistant and agent, formerly CodeWhisperer, tuned for cloud-heavy workflows.

Freemium· Free Tier: Free · Pro Tier: $19/mo. per usercode-completionagentic-coding
AS

Amazon SageMaker

Agents · Multi-model
7.0

AWS's end-to-end platform for building, training, and deploying machine learning models and AI agents at enterprise scale.

Paid· Pay-as-you-go; free tier available for new AWS accountsmodel-trainingmodel-deployment
FO

Forefront

Fine-tuning · Multi-model (Mistral-7B, Mixtral, Phi-2)
7.0

Fine-tune and serve open-source LLMs on your own data without managing GPUs.

Paid· Basic: $20 · Pro: $50 · Enterprise: Contact salesfine-tuningopen-source-llms
IS

iSpeech

Audio
7.0

Veteran cloud TTS and speech recognition API with broad SDK and language coverage.

Freemium· Free mobile SDK for non-revenue apps; ~$0.0001-$0.05 per word/transactiontext-to-speechspeech-recognition