Skip to content
@Audio-WestlakeU

Audio-WestlakeU

Audio Signal and Information Processing Lab at Westlake University

Pinned Loading

  1. FullSubNet FullSubNet Public

    PyTorch implementation of "FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement."

    Python 613 158

  2. NBSS NBSS Public

    The official repo of NBC & SpatialNet for multichannel speech separation, denoising, and dereverberation

    Python 373 46

  3. McNet McNet Public

    The official repo: "McNet: Fuse Multiple Cues for Multichannel Speech Enhancement", ICASSP 2023

    Python 133 17

  4. audiossl audiossl Public

    A library built for easier audio self-supervised training, downstream tasks evaluation

    Python 141 11

  5. FN-SSL FN-SSL Public

    The Official PyTorch Implementation of FN-SSL & IPDnet for Sound Source Localization [INTERSPEECH2023 & TASLP2024]

    Python 165 19

  6. ATST-SED ATST-SED Public

    This repo includes the official implementations of "Fine-tune the pretrained ATST model for sound event detection".

    Jupyter Notebook 174 17

Repositories

Showing 10 of 34 repositories
  • CleanMelPlus Public

    Pytorch implementation of "CleanMel+: Narrow-Band and Full-Band Sequence Modeling for Mel-Spectrogram Enhancement".

    Audio-WestlakeU/CleanMelPlus's past year of commit activity
    4 0 0 0 Updated Aug 11, 2026
  • Rec-RIR Public

    Official PyTorch implementation of 'Blind Room Impulse Response Identification via Reverberant Speech Spectrum Reconstruction' [Interspeech 2026]

    Audio-WestlakeU/Rec-RIR's past year of commit activity
    Python 37 MIT 2 0 0 Updated Aug 10, 2026
  • ATST-SED Public

    This repo includes the official implementations of "Fine-tune the pretrained ATST model for sound event detection".

    Audio-WestlakeU/ATST-SED's past year of commit activity
    Jupyter Notebook 174 MIT 17 3 0 Updated Jun 8, 2026
  • Mel-McNet Public

    The Official PyTorch Implementation of "Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement" [Interspeech 2025]

    Audio-WestlakeU/Mel-McNet's past year of commit activity
    Python 26 1 3 0 Updated May 14, 2026
  • FS-EEND Public

    The official Pytorch implementation of "Frame-wise streaming end-to-end speaker diarization with non-autoregressive self-attention-based attractors". [ICASSP 2024] and "LS-EEND: long-form streaming end-to-end neural diarization with online attractor extraction". [TASLP 2025]

    Audio-WestlakeU/FS-EEND's past year of commit activity
    Python 190 18 8 0 Updated May 7, 2026
  • VING Public

    Official implementation of VING: Variational Bayesian Inference with Multi-Aspect Neural Guidance for Speech Dereverberation

    Audio-WestlakeU/VING's past year of commit activity
    4 MIT 1 0 0 Updated Apr 17, 2026
  • FN-SSL Public

    The Official PyTorch Implementation of FN-SSL & IPDnet for Sound Source Localization [INTERSPEECH2023 & TASLP2024]

    Audio-WestlakeU/FN-SSL's past year of commit activity
    Python 165 19 7 0 Updated Mar 10, 2026
  • VINP Public

    Official PyTorch implementation of 'VINP: Variational Bayesian Inference with Neural Speech Prior for Joint ASR-Effective Speech Dereverberation and Blind RIR Identification' [IEEE TASLP]

    Audio-WestlakeU/VINP's past year of commit activity
    Python 38 MIT 7 1 0 Updated Feb 23, 2026
  • CleanMel Public

    Pytorch implementation of "CleanMel: Mel-Spectrogram Enhancement for Improving Both Speech Quality and ASR".

    Audio-WestlakeU/CleanMel's past year of commit activity
    Python 97 Apache-2.0 12 3 0 Updated Feb 2, 2026
  • audiossl Public

    A library built for easier audio self-supervised training, downstream tasks evaluation

    Audio-WestlakeU/audiossl's past year of commit activity
    Python 141 11 6 1 Updated Sep 25, 2025