whisper-at

0.5
46.22k

Joint speech recognition and audio tagging model.