Paper Reviews

VALL-E Paper Review

less than 1 minute read

Published:

πŸ“ Paper: Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

YourTTS Paper Review

less than 1 minute read

Published:

πŸ“ Paper: YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone

PortaSpeech Paper Review

less than 1 minute read

Published:

πŸ“ Paper: PortaSpeech: Portable and High-Quality Generative Text-to-Speech

wav2vec 2.0 Paper Review

less than 1 minute read

Published:

πŸ“ Paper: wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations

Glow-TTS Paper Review

less than 1 minute read

Published:

πŸ“ Paper: Glow-TTS: A Generative Flow for Text-to-Speech Synthesis
πŸ” Summary: This paper introduces a flow-based model for TTS, improving robustness compared to Tacotron.

VQ-wav2vec Paper Review

less than 1 minute read

Published:

πŸ“ Paper: vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

wav2vec Paper Review

less than 1 minute read

Published:

πŸ“ Paper: wav2vec: Unsupervised Pre-training for Speech Recognition

AutoVC Paper Review

less than 1 minute read

Published:

πŸ“ Paper: AUTOVC: Zero-Shot Voice Style Transfer with Only Autoencoder Loss

Glow Paper Review

less than 1 minute read

Published:

πŸ“ Paper: Glow: Generative Flow with Invertible 1x1 Convolutions

Conv-TasNet Paper Review

less than 1 minute read

Published:

πŸ“ Paper: Conv-TasNet: Surpassing Ideal Time-Frequency Magnitude Masking for Speech Separation