#transcription (29 Repositories)
Ranked open-source repositories tagged with #transcription, scored by pull request acceptance likelihood and maintainer engagement velocity.
61.2%
58.6h
29 repositories tagged #transcription
QuintinShaw/openasr
Local-first speech-to-text: no cloud, no telemetry, fail-closed by design. One CLI, seven model families, signed model catalog, OpenAI-compatible local API.
veedstudio/open-edit
Open-source, agent-driven editing pipeline: create subtitles, motion graphics, slides, edit and render videos.
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Picovoice/cheetah
On-device streaming speech-to-text engine powered by deep learning
silverstein/minutes
Every meeting, every idea, every voice note, searchable by your AI. Open-source, privacy-first conversation memory layer.
TypeWhisper/typewhisper-win
TypeWhisper for Windows - Local speech-to-text with translation
drakulavich/kesha-voice-kit
Give your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
moona3k/macparakeet
Fast, private, local-first voice app for Apple Silicon Macs — dictation, file/media transcription, meeting recording, Transforms, and a public automation CLI. Free and open-source.
island-io/mila
Mila — native macOS local transcription app (whisper.cpp) with optional speaker diarization. Apache-2.0.
BasedHardware/omi
AI that sees your screen, listens to your conversations and tells you what to do
TypeWhisper/typewhisper-mac
Local speech-to-text for macOS on-device AI, fully private, optional cloud
karolswdev/HoldSpeak
Cross-platform local voice typing and meeting transcription for macOS and Linux.
debpalash/VoiceStudio
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
yusufipk/dikte
Linux first voice-to-text dictation app.
bitwize-ai/Logue
Privacy-first AI meeting notes & writing assistant for Mac — on-device transcription, smart minutes, action items & 60+ writing modes, running entirely on Apple Silicon
CrispStrobe/CrispASR
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
richiejp/VoxInput
🎤 A voice assistant that let's you control any Linux desktop and transcribe any audio.
Muesli-HQ/muesli
Muesli - local meeting transcription + dictation for macOS (Granola + WisprFlow alternative)
Aayush9029/petal
Petal is a native macOS menu bar app for fast, local-first audio transcription.
kar2phi/video-lens
video-lens is a coding agent skill that fetches a YouTube transcript and generates a structured HTML report: executive summary, key points, analysis, takeaway, timestamped topic outline, and an embedded in-page player. No API keys, no external services beyond the coding agent itself.
franklioxygen/v2s
Live bilingual subtitles for any app on macOS. Captures audio, transcribes speech, and translates — all from your menu bar.
DevEmperor/DictateKeyboard
A powerful Whisper AI keyboard for reliable speech transcription
Zackriya-Solutions/meetily
Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust. 100% local processing. no cloud required. Meetily (Meetly Ai - https://meetily.ai) is the #1 Self-hosted, Open-source Ai meeting note taker for macOS & Windows. Understand How to write meeting minutes
JimLiu/baocut
Open-source Agent Skill that drives the BaoCut macOS app CLI (transcribe · subtitle · translate · cut) from Claude Code, Codex, and other agents
Picovoice/leopard
On-device speech-to-text engine powered by deep learning
AbhishekBarali/SpeakoFlow
Free, open-source offline voice dictation for Windows, macOS, and Linux. A Wispr Flow alternative with an AI assistant that can read your screen on request and answer questions.
chubbyguan/chubbyskills
把中文全渠道内容(抖音 / B站 / 小红书 / 公众号 / X / 播客)采集进个人知识库的 13 个 AI Skill:图文存图、视频转文字稿、字幕优先免 GPU,附带知识库 MCP server。 | Ingest Chinese content into your personal knowledge base — image/video routing, subtitle-first transcription, and a KB MCP server.
kristoferlund/ostt
Open source voice-to-text for the terminal. Record from a hotkey, transcribe with any provider, pipe to AI or shell commands.
LibraryOfCongress/concordia
Crowdsourcing platform for full text transcription and tagging. https://crowd.loc.gov