Skip to content
#

specaugment

Here are 17 public repositories matching this topic...

PyTorch implementation of Transformer-based Automatic Speech Recognition with attention mechanisms, SpecAugment, CTC loss, and mixed precision training. Achieves competitive WER/CER on LibriSpeech.

  • Updated Feb 21, 2025
  • Python

REST API based on PyTorch (ResNet18) for classifying 50 categories of natural and household sounds (rain, chainsaw, glass breaking, etc.) from audio files. Mel spectrograms + FastAPI. Val accuracy 86%. Trained in Google Colab on ESC-50.

  • Updated Aug 13, 2026
  • Jupyter Notebook

Improve this page

Add a description, image, and links to the specaugment topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the specaugment topic, visit your repo's landing page and select "manage topics."

Learn more