Skip to main content
📖 The AI Tool Bible

AudioCraft vs Horch

A side-by-side look at pricing, capabilities, pros, cons, and our editorial scores.

 AudioCraft logo
AudioCraft
Audio
Horch logo
Horch
Audio
TaglineMeta's open-source research toolkit for generating music and sound effects from text via a single autoregressive language model.Privacy-first, on-device meeting assistant for macOS
CategoryAudioAudio
PricingFree· Free and open source; self-hostedPaid· One-time purchase: €49 once
ModelMusicGen, AudioGen, EnCodecWhisper (local) for transcription; optional Ollama / MLX local LLMs for summarization
Editorial score8.2 / 10—
Use cases
text-to-musicsound-effectsaudio-compressionresearchself-hosted-generation
Confidential client meeting transcriptionAutomatic action-item extractionPer-contact relationship historyPre-meeting briefings from prior callsLocal Whisper transcription without cloud uploadFeeding meeting context into Claude or Cursor via MCPPersonal second-brain in MarkdownLegal, medical, or NDA-bound conversation notes
Pros
  • Fully open source with code and weights published by Meta
  • Single-LM architecture is simpler than diffusion pipelines
  • Covers music, sound effects, and neural codec in one repo
  • Strong baseline used widely in audio ML research
  • No usage fees once self-hosted
  • Fully on-device by default — audio and transcripts never leave the Mac unless the user opts in
  • One-time €49 purchase instead of a recurring per-seat SaaS bill
  • Records at OS level so no bot appears in the meeting and any app (Zoom, Meet, Teams, in-person) works
  • Notes stored as plain Markdown in ~/Meetings — portable, greppable, Obsidian-friendly
  • Auto-generates action items, topic summaries, and per-person profiles from spoken content
  • MCP server exposes meeting history to Claude, Cursor, and other agent clients
  • Local Whisper transcription plus optional Ollama/MLX means you pick the summarizer
Cons
  • No hosted product or managed API - you must run it yourself
  • Model weights typically CC-BY-NC, limiting commercial use
  • Requires GPU and ML tooling to operate
  • Output quality trails newer commercial models like Suno v4
  • macOS only — no Windows, Linux, iOS, or web client
  • Local Whisper and summarization need a reasonably modern Apple Silicon Mac to feel fast
  • No cloud sync or team workspace, so sharing across a team requires bring-your-own storage
  • Small independent product without the integrations catalog of Fireflies, Otter, or Fathom
  • OS-level capture depends on macOS screen/audio permissions, which some corporate MDM setups block
Websiteaudiocraft.metademolab.comhorch.app
Pick AudioCraft if
  • ✅ Fully open source with code and weights published by Meta
  • ✅ Single-LM architecture is simpler than diffusion pipelines
  • ✅ Covers music, sound effects, and neural codec in one repo
  • ✅ Strong baseline used widely in audio ML research
Pick Horch if
  • ✅ Fully on-device by default — audio and transcripts never leave the Mac unless the user opts in
  • ✅ One-time €49 purchase instead of a recurring per-seat SaaS bill
  • ✅ Records at OS level so no bot appears in the meeting and any app (Zoom, Meet, Teams, in-person) works
  • ✅ Notes stored as plain Markdown in ~/Meetings — portable, greppable, Obsidian-friendly