multichannel-asr

(★ 11)

A novel approach for multi-channel call transcription using mono-channel ASR models. Leverages VAD-based channel merging with silence insertion strategy to enable speaker diarization without complex separation models. Supports Whisper and other end-to-end ASR models.

multichannel-asr 최신버젼 다운로드

최종 버전 다운로드 (.zip)
// repository documentation