home/categories/media/benchflow-ai-skillsbench-tasks-speaker-diarization-subtitles-environment-skills-automatic-speech-recognition-skill-md
mediacontent-media
automatic-speech-recognition-asr
Transcribe audio segments to text using Whisper models. Use larger models (small, base, medium, large-v3) for better accuracy, or faster-whisper for optimized performance. Always align transcription timestamps with diarization segments for accurate speaker-labeled subtitles.
maintainer
benchflow-ai
अपडेट किया गया 1/23/2026
स्टार
946
फोर्क
244
quick start
Installation and usage
Transcribe audio segments to text using Whisper models. Use larger models (small, base, medium, large-v3) for better accuracy, or faster-whisper for optimized performance. Always align transcription timestamps with diarization segments for accurate speaker-labeled subtitles.
इंस्टॉलेशन
$ install --globalskills.sh
उपयोग
इंस्टॉल करने के बाद, आप टर्मिनल में यह कमांड चलाकर इस स्किल का उपयोग कर सकते हैं:
skills use automatic-speech-recognition-asr