SHAS: Approaching optimal Segmentation for End-to-End Speech Translation
Explore Similar Repositories
Vocal-Melody-Extraction:Source code for "Vocal melody extraction with semantic segmentation and audio-symbolic domain transfer learning".
avis:[CVPR 2025] ๐ฅ Official impl. of "Audio-Visual Instance Segmentation".
COMBO-AVS:[CVPR 2024 Highlight] Official implementation of the paper: Cooperation Does Matter: Exploring Multi-Order Bilateral Relations for Audio-Visual Segmentation
Wnet:Wnet: Audio-Guided Video Object Segmentation via Wavelet-Based Cross-Modal Denoising Networks
audio-seg-data-synth:Artificially synthesising data for audio segmentation to improve music-speech detection
// repository documentation
Was this content helpful?
โ 0(0 ratings)
Recent Feedback
Download README
Do you want to download the README.md file for SHAS?