Voice Transcription
skillTranscribe audio to text — speech-to-text using Whisper, deepgram, and other ASR models
The voice transcription skill converts spoken audio to text using automatic speech recognition (ASR). Supports OpenAI Whisper, Deepgram, and self-hosted models. Features include speaker diarization, timestamping, and multi-language support. Ideal for meeting notes, voice memos, and dictation.