ggml-org/whisper.cpp
Port of OpenAI's Whisper model in C/C++
MCP server that turns any video — YouTube, Instagram, TikTok, Loom, X, Vimeo, direct URLs, local files — into transcripts, key frames, OCR text, and metadata for AI agents.
Appears on
Quick read
Latest capture 2026-08-08 03:06
6 paths
Agent instructions and tool configuration found in this repository.
Agent instructions
Claude Code 5
7 observed captures since 2026-05-22. Observed captures are shown by default.
Stars from first capture +33
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Port of OpenAI's Whisper model in C/C++
MCP server for OpenRouter — chat with 300+ LLMs (Claude, Gemini, GPT), analyze images / audio / video, generate images / speech / music / video (Veo 3.1, Sora, Seedance, Wan) from Claude Desktop, Cursor, Kiro, VS Code.
Faster Whisper transcription with CTranslate2
A nearly-live implementation of OpenAI's Whisper.
Local speech-to-text for macOS on-device AI, fully private, optional cloud
Robust Speech Recognition via Large-Scale Weak Supervision