Official repository for RawNet, RawNet2, and RawNet3
-
Updated
Mar 21, 2024 - Python
Official repository for RawNet, RawNet2, and RawNet3
Deep speaker embeddings in PyTorch, including x-vectors. Code used in this work: https://arxiv.org/abs/2007.16196
Implementation of Neural Voice Cloning with Few Samples Research Paper by Baidu
Speaker identification/verification models for Machine Learning for Computer Vision class at UNIBO
[ICASSP'23] Online speaker clustering
The project is related to the development of labs for the ITMO Speaker Recognition Course.
Angular triplet center loss implementation in Pytorch.
This repository contain the code of the main part of my master thesis degree at Politecnico di Torino in Data science & Engineering
Speaker diarization on VoxData — benchmarked state-of-the-art speaker embedding models against each other, built a custom diarization model, and compared clustering strategies.
I. Thoidis, C. Gaultier, and T. Goehring, "Perceptual Analysis of Speaker Embeddings for Voice Discrimination between Machine And Human Listening," ICASSP 2023 - 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Rhodes Island, Greece, 2023, pp. 1-5
ReDimNet2 inference w/ ONNX
Deterministic checks for speaker-verification studies: detect F0-tracking contamination that can invert your statistics. Zero dependencies, no audio required.
Persistent speaker identity across audio files — a local voice registry that resolves diarization clusters to named speakers
Multi-resolution speaker embeddings with ReDimNet architecture and Matryoshka Representation Learning (MRL)
Research about how to design a robot sound from voice conversion and speaker embeddings
Generation of Novel Human Voices for Voice Anonymization. MSc Project, 2026 | UniStuttgart
To associate your repository with the speaker-embeddings topic, visit your repo's landing page and select "manage topics."