All 26 Dependencies
Every package Speakr depends on, ranked by repo health score.
Speakr is a self-hosted web application that transforms audio and video recordings into organized, searchable, AI-powered notes. Built for privacy-conscious individuals and teams, it runs entirely on your own infrastructure so sensitive conversations never leave your control. From one-click Docker deployment to GPU-accelerated WhisperX diarization, Speakr handles the full pipeline from capture to insight.
The platform supports a connector-based transcription architecture that auto-detects your preferred engine — whether that is self-hosted WhisperX for best-in-class diarization and voice profiles, OpenAI's gpt-4o-transcribe-diarize for cloud simplicity, Mistral Voxtral, VibeVoice, or Azure OpenAI. Smart tags carry their own AI prompts that stack and layer, transforming raw transcripts into recipes, action item lists, study notes, or any other structured format you define. Groups with granular sharing permissions, OIDC SSO, retention policies, signed webhooks, and a full REST API make Speakr a serious platform rather than a weekend project.
Version 0.9.0 elevated Speakr's mobile experience to first-class status with a proper bottom-nav detail view, a drag-to-dismiss upload sheet, and a redesigned recording surface that works edge-to-edge on phones. A new Stats tab shows per-speaker breakdowns of speaking time, turn count, and words per minute. Server-side recording sessions enable long multi-hour captures that survive page reloads, and a Phase 1-3 webhook system with HMAC signing and exponential-backoff retry connects Speakr to automation tools like n8n, Zapier, and Make.
Speakr ships as an installable Progressive Web App with a web share target so your OS can push audio files directly from the share sheet. Seven languages are fully localized, including English, French, German, Spanish, Russian, Simplified Chinese, and Brazilian Portuguese.