Open Source Voice AI Apps

Explore open source voice AI tools for speech-to-text dictation, text-to-speech, voice cloning and voice assistants, plus the commercial products they replace.

11 apps available

Apps in Voice AI

TypeScript
64%
Other

Epicenter

Developer Tools · Knowledge Management · Note Taking

4,808

A local-first monorepo led by Whispering, an open-source speech-to-text app, built on an MIT toolkit that turns your data into plain Markdown and SQLite files you own instead of a database you rent.

View details
88
Repo Health
90
Technical
64
Dependency
Built with
TypeScript 64%
HTML 13%
Svelte 13%
Updated 6 days ago
JavaScript
48%
MIT

OpenWhispr

AI Assistants · Productivity · Voice AI

8,711

Privacy-first, cross-platform voice-to-text with local AI and cloud options

View details
88
Repo Health
80
Technical
71
Dependency
Built with
JavaScript 48%
TypeScript 46%
Updated 5 days ago
Swift
99%
Other

VoiceInk

AI Assistants · Productivity · Voice AI

6,579

A native macOS dictation app that turns speech into text on-device with Whisper or Parakeet models, then cleans the result up with an AI pass tuned to whichever app you are typing into.

View details
88
Repo Health
71
Technical
0
Dependency
Built with
Swift 99%
Updated 5 days ago
Rust
59%
MIT

Handy

Productivity · Voice AI

32,320

Free, offline, open-source speech-to-text that pastes directly into any app on Windows, macOS, and Linux.

View details
87
Repo Health
78
Technical
72
Dependency
Built with
Rust 59%
TypeScript 30%
Updated 5 days ago
Python
56%
AGPL 3.0

VoiceStudio

Mcp · Music Audio · Voice AI

51,962

Open-source, fully local ElevenLabs alternative for voice cloning, voice design, video dubbing, dictation, transcription and audiobooks, with a local API and MCP server for agents.

View details
85
Repo Health
83
Technical
70
Dependency
Built with
Python 56%
JavaScript 23%
TypeScript 19%
Updated today
TypeScript
89%
MIT

Amical

AI Assistants · Note Taking · Voice AI

1,540

Local-first AI dictation that understands your active app — private, offline, and built for speed.

View details
82
Repo Health
82
Technical
68
Dependency
Built with
TypeScript 89%
Updated 1 weeks ago
TypeScript
53%
MIT

Voicebox

AI Development · Productivity · Voice AI

55,864

Clone voices, dictate anywhere, and give AI agents your voice — all locally.

View details
79
Repo Health
76
Technical
68
Dependency
Built with
TypeScript 53%
Python 35%
Updated 1 months ago
Python
82%
Other

fish-speech

AI Development · Developer Tools · Music Audio

32,864

SOTA open-source dual-autoregressive text-to-speech model with rapid voice cloning, inline emotion tags, and real-time streaming inference across 80+ languages.

View details
69
Repo Health
71
Technical
73
Dependency
Built with
Python 82%
TypeScript 14%
Updated 2 weeks ago
Swift
99%
Other

Ghost Pepper

AI Assistants · Voice AI

3,192

A 100% private, on-device macOS dictation and meeting-transcription app — hold Control to talk, transcribes and pastes locally via Apple Silicon with no cloud APIs or data leaving your Mac.

View details
67
Repo Health
68
Technical
0
Dependency
Built with
Swift 99%
Updated 2 months ago
Python
99%
Apache 2.0

Rasa Open Source

AI Assistants · AI Development · Voice AI

21,332

Rasa Open Source is a Python machine learning framework for building contextual, multi-turn chatbots and voice assistants that understand natural language and maintain conversation state.

View details
64
Repo Health
78
Technical
63
Dependency
Built with
Python 99%
Updated 2 months ago
Swift
78%
MIT

Frog

AI Assistants · Productivity · Voice AI

47

An open-source, on-device Mac desktop pet that types what you say, takes meeting notes with speaker labels, and talks back — a private alternative to Wispr Flow and Granola.

View details
52
Repo Health
65
Technical
91
Dependency
Built with
Swift 78%
JavaScript 20%
Updated 3 weeks ago

About Voice AI

Voice as an interface

Speech models got good and small at the same time. Whisper, Parakeet and a wave of open text-to-speech models now run on a laptop, which means voice software no longer has to send your audio to someone else’s server. This category collects the open source projects built on that shift, alongside the commercial products they compete with.

What’s in this category

  • Dictation and speech-to-text — system-wide apps that let you talk into any text field and get clean, punctuated writing back. Many add an AI pass that removes filler words and adapts formatting to the app you’re typing in.
  • Text-to-speech and voice cloning — models and studios that generate natural speech from text, clone a voice from a short sample, and stream audio in real time for narration, voiceovers or agents.
  • Voice assistants — frameworks for building assistants that understand spoken requests, keep track of a conversation and act on it.

How to choose

Local or cloud. On-device tools keep audio private and cost nothing per use, but quality and speed depend on your hardware and the model you pick. Cloud services usually win on polish and cross-device sync, at the cost of sending recordings to a third party.

Model choice. Check which speech models a tool supports and whether you can swap them. Being able to move from a small, fast model to a larger, more accurate one matters more over time than any single feature.

Platforms. Many dictation apps are macOS-only. If you work across Windows, Linux or mobile, filter for that first.

Latency. For dictation and live assistants, the delay between speaking and seeing or hearing a result decides whether a tool feels usable. Streaming support is worth looking for.

Licensing for voices. For text-to-speech and cloning, read the model license as well as the code license — some open weights restrict commercial use.

Join founders buildingwith open source

Opinionated takes, migration guides, cost-saving tips, and insights from the open source ecosystem.

Subscribe on Substack
Join 750+ subscribers