[Interspeech 2025] DualCodec: A Low-Frame-Rate, Semantically-Enhanced Neural Audio Codec
-
Updated
Sep 9, 2026 - Jupyter Notebook
[Interspeech 2025] DualCodec: A Low-Frame-Rate, Semantically-Enhanced Neural Audio Codec
[ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates
Neural Audio Codecs implemented in C# - DAC, SNAC, Encodec, Dia
Unofficial fairseq-free PyTorch implementation of UTMOS (v1, 2022), matching the original system.
Causal zero-shot voice conversion at 44.1kHz.
Implementation of the Descript Audio Codec in MLX
Author's code of "Speaker anonymization using neural audio codec language models" (ICASSP 2024).
Unofficial PyTorch implementation of Higgs Audio V2 Tokenizer with HuBERT semantic features. Complete training pipeline for semantic-acoustic audio tokenization with 960x downsampling and 8-layer RVQ.
A minimal neural audio codec 🔊
Official inference code and checkpoints for CodecSlime.
A lightweight neural audio codec.
Unofficial fairseq-free PyTorch implementation of SCOREQ (2024), matching the original system.
Neural Audio Codec: What Determines the Rate–Distortion Ceiling of Scalar-Quantized Speech Codecs?
Visual catalog of neural audio and speech codec architectures, audio VAEs and continuous autoencoders, with code, checkpoints, figures and license evidence.
Transfer sonic characteristics from reference sounds to a source audio using Music2Latent embeddings
Zero-shot text-to-speech over a neural audio codec: discrete acoustic tokens, a token language model, a neural vocoder, and a voice-cloning CLI
Neural audio codec and tokenizer for audio language models — SEANet encoder with residual vector quantization in PyTorch
Open research toward end-to-end spoken dialogue systems. Ships SoviaMate-Codec — a neural audio codec for LLM integration with ASR-constrained encoding, enhancement training, and zero-shot speaker adaptation.
To associate your repository with the neural-audio-codec topic, visit your repo's landing page and select "manage topics."