AudioCraft vs Horch
A side-by-side look at pricing, capabilities, pros, cons, and our editorial scores.
AudioCraft Audio | Horch Audio | |
|---|---|---|
| Tagline | Meta's open-source research toolkit for generating music and sound effects from text via a single autoregressive language model. | Privacy-first, on-device meeting assistant for macOS |
| Category | Audio | Audio |
| Pricing | Free· Free and open source; self-hosted | Paid· One-time purchase: €49 once |
| Model | MusicGen, AudioGen, EnCodec | Whisper (local) for transcription; optional Ollama / MLX local LLMs for summarization |
| Editorial score | 8.2 / 10 | — |
| Use cases | text-to-musicsound-effectsaudio-compressionresearchself-hosted-generation | Confidential client meeting transcriptionAutomatic action-item extractionPer-contact relationship historyPre-meeting briefings from prior callsLocal Whisper transcription without cloud uploadFeeding meeting context into Claude or Cursor via MCPPersonal second-brain in MarkdownLegal, medical, or NDA-bound conversation notes |
| Pros |
|
|
| Cons |
|
|
| Website | audiocraft.metademolab.com | horch.app |
Pick AudioCraft if
- ✅ Fully open source with code and weights published by Meta
- ✅ Single-LM architecture is simpler than diffusion pipelines
- ✅ Covers music, sound effects, and neural codec in one repo
- ✅ Strong baseline used widely in audio ML research
Pick Horch if
- ✅ Fully on-device by default — audio and transcripts never leave the Mac unless the user opts in
- ✅ One-time €49 purchase instead of a recurring per-seat SaaS bill
- ✅ Records at OS level so no bot appears in the meeting and any app (Zoom, Meet, Teams, in-person) works
- ✅ Notes stored as plain Markdown in ~/Meetings — portable, greppable, Obsidian-friendly