Cutting edge AI technology for automated audio transcription. A nice GUI for OpenAIs Whisper and pyannote (speaker identification)
-
Updated
Sep 27, 2026 - Python
Cutting edge AI technology for automated audio transcription. A nice GUI for OpenAIs Whisper and pyannote (speaker identification)
Pybind11 bindings for Whisper.cpp
The main repo for Stage Whisper — a free, secure, and easy-to-use transcription app for journalists, powered by OpenAI's Whisper automatic speech recognition (ASR) machine learning models.
Record audio or transcribe files using ctranslate2 and whisper!
A static site demonstrating real-time audio transcription via Amazon Transcribe over a WebSocket.
🎙️ Offline audio transcription with Whisper.
AI-powered transcription for audio & video with Whisper — self-hosted, fast, and open-source.
WhisperClip simplifies your life by automatically transcribing audio recordings and saving the text directly to your clipboard. With just a click of a button, you can effortlessly convert spoken words into written text, ready to be pasted wherever you need it. This application harnesses the power of OpenAI’s Whisper for free.
Free speech to text
GUI for Faster‑Whisper‑XXL transcription tool: download YouTube audio, transcribe local files, manage models, and export multiple formats with themes and auto yt‑dlp updates.
Efficient LLM inference on Slurm clusters.
Transcribe any audio or video file. Edit and view your transcripts in a standalone HTML editor.
Uses the powerful WhisperS2T and Ctranslate2 libraries to batch transcribe multiple files
CLI for audio, video, and text transcription with ASR providers and LLM-powered summarization via local or cloud backends.
Open source transcription, live subtitling and meeting summarization. Web app, API and SDKs of the LinTO platform.
Record and transcribe Teams, Zoom, and Google Meet calls locally with AI-powered speaker identification. Open-source alternative to Evaer, Otter.ai, and Fireflies. Offline speech-to-text using Whisper — no cloud, no subscriptions.
Speaker identification powered by pyannote and resemblyzer
Streamlit Audio Transcription with OPENAI's Whisper Ai: An interactive Streamlit app demonstrating real-time audio transcription using OPENAI's Whisper Ai.
Podcast/ YouTube video → Transcript!
本地优先的 macOS 与 Windows 语音输入工具,支持端侧 SenseVoice 识别、全局语音输入、文件转写及可选 AI 校对。Local-first voice dictation for macOS and Windows with on-device SenseVoice, typing into any app, file transcription, and optional AI proofreading.
To associate your repository with the audio-transcription topic, visit your repo's landing page and select "manage topics."