Skip to content
#

whisper-alternative

Here are 26 public repositories matching this topic...

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

  • Updated Jul 30, 2026
  • Python

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

  • Updated Jul 27, 2026
  • C

🎬 AI subtitle generator: convert video to SRT subtitles locally with NVIDIA NeMo Parakeet-TDT speech-to-text. GPU-accelerated, word-level timestamps, VAD, LLM correction — a fast offline Whisper alternative.

  • Updated Jul 27, 2026
  • Python

Fully-local speech-to-text dictation. Hold a hotkey, talk, and the transcript lands in the field you're already in — an NVIDIA Parakeet streaming server plus native macOS and Windows clients. Your voice never leaves your LAN.

  • Updated Jul 29, 2026
  • Python

Fast native C inference engine for speech recognition & translation, llama.cpp-style: streaming + offline ASR, word timestamps, int8/int4 quantization, CPU/Metal/CUDA. Today it runs the best open models (Parakeet, Canary, Nemotron) — built to host more engines tomorrow.

  • Updated Jul 30, 2026
  • C

Improve this page

Add a description, image, and links to the whisper-alternative topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the whisper-alternative topic, visit your repo's landing page and select "manage topics."

Learn more