home/categories/media/benchflow-ai-skillsbench-tasks-speaker-diarization-subtitles-environment-skills-automatic-speech-recognition-skill-md
mediacontent-media

automatic-speech-recognition-asr

Transcribe audio segments to text using Whisper models. Use larger models (small, base, medium, large-v3) for better accuracy, or faster-whisper for optimized performance. Always align transcription timestamps with diarization segments for accurate speaker-labeled subtitles.

benchflow-ai
maintainer
benchflow-ai
আপডেট হয়েছে 1/23/2026
স্টার
946
ফর্ক
244
quick start

Installation and usage

Transcribe audio segments to text using Whisper models. Use larger models (small, base, medium, large-v3) for better accuracy, or faster-whisper for optimized performance. Always align transcription timestamps with diarization segments for accurate speaker-labeled subtitles.

ইনস্টলেশন
$ install --globalskills.sh
ব্যবহার

ইনস্টল করার পর, টার্মিনালে নিচের কমান্ড চালিয়ে আপনি এই স্কিল ব্যবহার করতে পারবেন:

skills use automatic-speech-recognition-asr