Repositories
PennyZhang9/homo
一个高性能,易于扩展且完全开源的自然交互系统
PennyZhang9/DataMining-Forest_Cover_Type
Data Mining Homework
PennyZhang9/kaldi-onnx
Kaldi model converter to ONNX
PennyZhang9/the-gan-zoo
A list of all named GANs!
PennyZhang9/RecorderManager
Android仿微信录制音视频的管理工具,支持自定义
PennyZhang9/SincNet
SincNet is a neural architecture for efficiently processing raw audio samples.
PennyZhang9/spec_augment
🔦 A Pytorch implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition
PennyZhang9/ENAS-pytorch
PyTorch implementation of "Efficient Neural Architecture Search via Parameters Sharing"
PennyZhang9/attention-is-all-you-need-pytorch
A PyTorch implementation of the Transformer model in "Attention is All You Need".
PennyZhang9/Attention-Augmented-Conv2d
Implementing Attention Augmented Convolutional Networks using Pytorch
PennyZhang9/kaldi-io-for-python
Python functions for reading kaldi data formats. Useful for rapid prototyping with python.
PennyZhang9/pytorch-kaldi
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label computation, and decoding are performed with the kaldi toolkit.
PennyZhang9/OpenSeq2Seq
Toolkit for efficient experimentation with Speech Recognition, Text2Speech and NLP
PennyZhang9/g2p
g2p: English Grapheme To Phoneme Conversion
PennyZhang9/phoneme_ctc
Bidirectional dynamic RNN + CTC for phoneme recognition
PennyZhang9/CTC-speech-recognition
This is a working example of using CTC for phone recognition on TIMIT
PennyZhang9/gans-in-action
Companion repository to GANs in Action: Deep learning with Generative Adversarial Networks
PennyZhang9/LSTM_PIT_Speech_Separation
Multi-talker Speech Separation with LSTM/BLSTM by Permutation Invariant Training method.
PennyZhang9/pyfasst
Python implementation of the Flexible Audio Source Separation Toolbox (FASST)
PennyZhang9/pytorch-vqvae
Vector Quantized VAEs - PyTorch Implementation
PennyZhang9/RIR-Generator
Generating room impulse responses
PennyZhang9/SpeakerVerification_AMSoftmax_pytorch
SE-Resnet+AMSoftmax for Speaker Verification
PennyZhang9/DeepLearningTutorials
Deep Learning与PyTorch入门实战视频教程
PennyZhang9/UnsupervisedDeepLearning-Pytorch
This repository tries to provide unsupervised deep learning models with Pytorch
PennyZhang9/Adversarial-Autoencoder
An adversarial autoencoder implementation in pytorch
PennyZhang9/awesome-python-scientific-audio
Curated list of python software and packages related to scientific research in audio
PennyZhang9/Multi-channel-speech-extraction-using-DNN
A tensorflow implementation of my paper Combining beamforming and deep neural networks for multi-channel speech extraction
PennyZhang9/RtmpRecoder
Record camera and push stream to rtmp server.
PennyZhang9/DeepDenoisingAutoencoder
Tensorflow implementation for Speech Enhancement (DDAE)