FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

specaugment · GitHub Topics · GitHub

#

specaugment

Here are 18 public repositories matching this topic...

A Implementation of SpecAugment with Tensorflow & Pytorch, introduced by Google Brain

  • Updated Apr 5, 2022
  • Python

SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition

  • Updated Sep 5, 2020
  • Python

Tensor2tensor experiment with SpecAugment

  • Updated May 13, 2019
  • Python

End-to-end speech recognition on AISHELL dataset.

  • Updated Nov 9, 2021
  • Python

fast SpecAugmentation code with numpy and scipy

  • Updated Jul 5, 2019
  • Python

tf 2.0 implementation of Listen, attend and spell

  • Updated Jan 19, 2021
  • Python

A minimalistic Tensorflow 2.x Keras layer which applies SpecAugment to its input

  • Updated Aug 11, 2021
  • Jupyter Notebook

Emotion recognition with IEMOCAP datasets. We compare the results with SpecAugmentation and CodecAugmentation. For audio codec implementation, we have selected opus.

  • Updated Apr 21, 2021
  • Jupyter Notebook

XSpeech: A Novel Deep Learning Approach to Classifying Stutters

  • Updated Aug 3, 2025
  • Jupyter Notebook

End-to-end English speech recognition in PyTorch from scratch: CNN + BiLSTM + CTC trained on 100h LibriSpeech. 22.6% WER greedy, 12.2% with beam search + 4-gram LM under 7 GPU-hours on a single laptop GPU.

  • Updated Jun 12, 2026
  • Python

Radio signal spectrogram classification with PyTorch, SpecAugment-based data augmentation, and a pretrained EfficientNet-B0 model for four-class signal recognition.

  • Updated May 8, 2026
  • Jupyter Notebook

Speech recognition toolkit featuring SpecAugment, Whisper fine-tuning, LAS ASR, synthetic data generation, benchmarking, and a desktop GUI. Built with PyTorch, Transformers, Librosa, and CustomTkinter.

  • Updated Jul 1, 2026
  • Python

PyTorch implementation of Transformer-based Automatic Speech Recognition with attention mechanisms, SpecAugment, CTC loss, and mixed precision training. Achieves competitive WER/CER on LibriSpeech.

  • Updated Feb 21, 2025
  • Python

Simple numpy-based implementation of SpecAugment

  • Updated Sep 25, 2023
  • Python

Performs data augmentation as according to the SpecAugment paper. Modified from Lingvo (TensorFlow > 1.10.0).

  • Updated Jan 26, 2022
  • Python

An Audio Classification task with two types of inputs to the CNN models for intended work using Tensorflow.

  • Updated Jul 15, 2025
  • Jupyter Notebook

REST API based on PyTorch (ResNet18) for classifying 50 categories of natural and household sounds (rain, chainsaw, glass breaking, etc.) from audio files. Mel spectrograms + FastAPI. Val accuracy 86%. Trained in Google Colab on ESC-50.

  • Updated Aug 26, 2026
  • Jupyter Notebook

FastAPI service for music genre classification (GTZAN, 10 classes) using a CNN with Mel spectrograms. 78% test accuracy, SpecAugment data augmentation, fully offline inference.

  • Updated Aug 26, 2026
  • Jupyter Notebook

Improve this page

Add a description, image, and links to the specaugment topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the specaugment topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL