
Vibe
thewh1teagle · Audio
GitHub stars
7,673Stats updated
What is Vibe?
Vibe turns local recordings, downloads, microphone input, and system audio into transcripts across desktop platforms. It combines several speech-model families with GPU acceleration, batch queues, diarization, stable subtitle timing, wide export support, local Ollama analysis, optional Claude summaries, and programmatic access through its CLI and HTTP API.
Vibe is a desktop app for offline audio and video transcription on macOS, Windows, and Linux. It supports batch transcription, speaker diarization, and exports including SRT, VTT, TXT, HTML, and PDF.
Core capabilities
- Transcribe audio, video, microphone, system audio, or supported web media
- Run Whisper, Nemotron, or Parakeet models with local hardware acceleration
- Export subtitles and documents or automate work through CLI and HTTP API
Local model choice across desktops
Vibe supports Whisper, Nemotron, and Parakeet model paths across macOS, Windows, and Linux, with CoreML or Vulkan-oriented acceleration depending on the system. Users can customize model arguments, transcribe from files or live sources, and keep the default recognition step on their own machine.
- Whisper
- Parakeet
- Nemotron
- offline
Transcripts for many destinations
Batch queues, diarization, subtitle-focused stable timestamps, translation to English, printing, and exports ranging from TXT and JSON to SRT, PDF, and DOCX turn one recognition run into several deliverables. CLI commands, an HTTP API with Swagger documentation, Ollama, and optional Claude summaries extend the same output programmatically.
- subtitles
- diarization
- exports
- API
Where it fits
Use cases
- 01
Produce subtitles for a recording
Open a video or audio file, choose language and timing settings, apply speaker diarization when needed, and export SRT, VTT, or another supported format for editing or delivery.
- 02
Process a local interview archive
Queue multiple recordings, run transcription on the workstation, preview text while processing, and export searchable document formats without uploading the source media to a required transcription service.
- 03
Connect transcripts to analysis
Send completed text to a local Ollama workflow or an optional Claude-powered summary, or use the CLI and HTTP API to incorporate transcription and output handling into another tool.