AI tools tagged Cloud Only
48 tools matching this tag.tech
Pinecone
FeaturedManaged vector database for production-scale similarity search.
Together AI
FeaturedFine-tune & serve open-weight models (Llama, Mistral, DeepSeek).
Modal
Serverless GPUs and infra for training & serving ML.
Snowflake Cortex
Generative AI and RAG built into the Snowflake data cloud
DALL·E 3
OpenAI's image model — strong on prompt adherence and text-in-image.
DataStax Astra DB
Serverless vector and document database for production RAG and AI agents
MongoDB Atlas Vector Search
Vector search built into the operational database you're already using.
Replicate
One-API platform for running and fine-tuning open-source models.
Seedream
ByteDance's unified text-to-image and image-editing model, served via Volcengine.
Aider
Terminal-based AI pair programmer that writes commits.
OpenAI Fine-tuning
Fine-tune GPT-4o-mini and friends on your own data.
Palantir AIP
Enterprise AI platform that grounds LLMs in your operational data and runs agents against real business systems.
Weaviate
Open-source vector DB with hybrid search and modules.
RunPod
On-demand GPU cloud and serverless inference platform built specifically for AI workloads.
Skyvern
AI browser agent that automates web workflows from natural-language instructions, with CAPTCHA and 2FA handling built in.
Browser Use
Open-source browser automation harness and cloud platform for LLM agents that drive real websites.
CoreWeave
AI-native GPU cloud built for large-scale training, fine-tuning, and inference on NVIDIA hardware.
Daytona
Secure, isolated sandboxes for running AI-generated code with sub-90ms cold starts.
Databricks Vector Search
Managed hybrid vector search that lives inside the Databricks lakehouse and auto-syncs with your source tables.
Lambda
On-demand NVIDIA GPU cloud built specifically for training, fine-tuning, and serving large AI models.
W&B Weave
Production observability, tracing, and evaluation for LLM and agent systems from the Weights & Biases stack.
Exa
Web search API built for AI agents, with structured outputs and token-efficient highlights.
Fal.ai
Serverless GPU inference platform optimized for fast diffusion and generative media APIs.
Amazon Q
AWS's enterprise AI assistant for developers, analysts, and knowledge workers, wired into your company data.
Anyscale
Ray-powered platform for training, serving, and scaling LLMs.
DVC
Git-style version control for datasets, ML models, and experiment pipelines.
Manus
Generalist agent for research, code, and web tasks.
Devin
Cognition Labs' "autonomous software engineer" agent.
Agentset
Production-ready RAG infrastructure with agentic search, citations, and model-agnostic plumbing.
Azure AI Speech (Neural TTS)
Microsoft's enterprise-grade neural text-to-speech with 100+ languages, custom brand voices, and SSML control.
Epsilla
Agent-as-a-Service platform with managed RAG and a no-code builder for vertical enterprise AI.
FedML
Distributed training, fine-tuning, and serving platform with federated learning roots.
Beam
Serverless GPU infrastructure for AI workloads with sub-second cold starts and bring-your-own-cloud support.
Metabase
Open-source BI platform with a Metabot AI layer for natural-language querying over your warehouse.
Ollama
The de facto runtime for running open-weights LLMs locally, now with a paid cloud tier for bigger models.
Paperspace Gradient
End-to-end MLOps platform with GPU notebooks, training jobs, and model deployment, now folded into DigitalOcean.
SAS Viya
Enterprise-grade data and AI analytics platform with built-in governance, MCP server, and a Copilot for regulated industries.
Fiddler AI
Enterprise AI observability and guardrails platform for monitoring agents, LLMs, and ML models in production.
FutureHouse Platform
Multi-agent AI research stack for scientists, with retrieval over 175M+ papers, patents, and trials.
Hyperbrowser
Cloud browser infrastructure built for AI agents that need to scrape, click, and navigate the live web.
Neu.ro
Infrastructure-agnostic MLOps platform for the full ML/DL lifecycle across public, hybrid, and on-prem clouds.
Rivestack
Managed Postgres with pgvector on dedicated NVMe, pitched as a cheaper RAG backend than Pinecone or Supabase.
StarOps
AI-native platform engineering engine that provisions and manages cloud infrastructure from natural-language prompts.
W&B Sweeps
Hyperparameter optimization from Weights & Biases with Bayesian search and Hyperband early stopping.
Amazon Q Developer
AWS's in-IDE coding assistant and agent, formerly CodeWhisperer, tuned for cloud-heavy workflows.
Amazon SageMaker
AWS's end-to-end platform for building, training, and deploying machine learning models and AI agents at enterprise scale.
Forefront
Fine-tune and serve open-source LLMs on your own data without managing GPUs.
iSpeech
Veteran cloud TTS and speech recognition API with broad SDK and language coverage.