RepoCatalog

Voice Transcription

skill

Transcribe audio to text — speech-to-text using Whisper, deepgram, and other ASR models

Tags:speechtranscriptionaudiowhisper

The voice transcription skill converts spoken audio to text using automatic speech recognition (ASR). Supports OpenAI Whisper, Deepgram, and self-hosted models. Features include speaker diarization, timestamping, and multi-language support. Ideal for meeting notes, voice memos, and dictation.