Skip to main content
📖 The AI Tool Bible

AI real-time translation

Editorial picks for "real time translation ai".

6 tools

All Audio →
WH

Whisper

Audio · Whisper large-v3
8.6

OpenAI's open-source speech-to-text — the de-facto baseline.

Free· Free open weights; $0.006/min via OpenAI APItranscriptionself-hosted
EC

ElevenLabs Conversational AI

Audio · ElevenLabs Scribe (ASR) + pluggable LLM + ElevenLabs TTS
8.1

Production-grade voice agent platform layering ElevenLabs TTS, ASR, and LLM orchestration into a single deployable stack.

Paid· Free: $0 · Starter: $6 · Creator: $11 · Pro: $99 · Scale: $299voice-agentsivr-replacement
DE

Deepgram

Audio · Nova, Flux, Speak (proprietary)
7.9

Production-grade speech-to-text, text-to-speech, and voice-agent APIs for real-time and batch audio.

Freemium· Free credits on signup; usage-based pricing; enterprise contracts availablespeech-to-texttext-to-speech
AA

Azure AI Speech (Neural TTS)

Audio · Azure Neural TTS (plus HD and Azure OpenAI voices)
7.3

Microsoft's enterprise-grade neural text-to-speech with 100+ languages, custom brand voices, and SSML control.

Freemium· Free (F0): Free · Pay as You Go: Voice Live Prices: $- · Commitment Tiers – Standard: $- for 2,000 hourstext-to-speechvoice-cloning
KM

Kyutai Moshi

Audio · Moshi (7B-class speech-text foundation model) + Mimi neural audio codec, in-house by Kyutai

Open-source, full-duplex speech-to-speech foundation model with sub-200ms latency

Free· Free and open source. Models under CC-BY 4.0, code under MIT (Python) / Apache 2.0 (Rust). Self-hosted only — you pay your own compute (24GB+ GPU for PyTorch, or Apple Silicon via MLX).Real-time voice assistant prototypesResearch on full-duplex spoken dialogue
TR

Transept

Writing · Multi-model (benchmarks Claude, GPT, Gemini, and open models; final routing not disclosed)

AI translation workspace with shared glossaries, styleguides, and decision-context memory

Freemium· Free: €0 · Starter: €29 · Pro: €79literary translationbook localization