Source-Denoising-Pix2Pix-cGAN

Basic Information

Author: Gregory Hunkins

Organization: University of Rochester

License: MIT

Abstract: An Conditional Generative Adverserial Network (cGAN) was adapted for the task of source de-noising of noise voice auditory images. The base architecture is adapted from Pix2Pix. The cGAN inputs fixed-length short-time Fourier Transform (STFT) magnitude features of the noisy speech and returns a denoised image that can be converted to the time-domain via IFFT. The dataset is created by randomly combined TIMIT speaker samples and non-stationary noise. For the non-stationary noise, Duan et al.'s dataset is used: http://www2.ece.rochester.edu/~zduan/data/noise/.

Running The Code

Reference: https://cs.rochester.edu/~cxu22/t/577F17/bluehive_tutorial.html

For the most recent architecture, navigate into the Architecture_v4 folder. Submit job.sh to train the architecture and save results.

sbatch src/model/job.sh

Subjective Evaluation

The full validation set can be downloaded in two different ways: via validation example or seperated into dB noise classes. The first link contains all validation examples with both the appropriate WAV and PNG files. The second link only contains the relevant WAV files. Both are ~2GB.

Full Validation (WAV and PNG) and dB Separated Validation Set (WAV)

Data

The data is available in HDF5 format.

NoisyAudImg10K

Name		Name	Last commit message	Last commit date
Latest commit History 84 Commits
Architecture_v2		Architecture_v2
Architecture_v3		Architecture_v3
Architecture_v4		Architecture_v4
SEGAN_test		SEGAN_test
.DS_Store		.DS_Store
README.md		README.md

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Source-Denoising-Pix2Pix-cGAN

Basic Information

Running The Code

Subjective Evaluation

Data

About

Releases

Packages

Languages

ghunkins/Voice-Denoising-AN

Folders and files

Latest commit

History

Repository files navigation

Source-Denoising-Pix2Pix-cGAN

Basic Information

Running The Code

Subjective Evaluation

Data

About

Topics

Resources

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages