RP

Projects

Explore my key research implementations, neural speech models, and engineering projects. This page lists various tools and frameworks I have developed, focusing on phonetic-preserving Text-to-Speech (TTS), fast-converging GAN vocoders (like FC-HiFiGAN, Swar, and Vaachika), speaker diarization pipelines, and multilingual audio deepfake detection (ADD) datasets. Most of these projects are open-source and available on GitHub for collaboration.

Research implementations, speech synthesis models, and engineering projects.

Speech & Audio Processing
Machine Learning & Computer Vision
Web Applications & Distributed Systems