Unofficial PyTorch implementation of Higgs Audio V2 Tokenizer with HuBERT semantic features. Complete training pipeline for semantic-acoustic audio tokenization with 960x downsampling and 8-layer RVQ.
pytorchaudio-synthesisspeech-processingaudio-processingvector-quantizationdacsemantic-featureshubertaudio-generationneural-audio-codecrvqaudio-tokenizerneural-codechiggs-audiospeech-tokenization
-
Updated
Oct 8, 2025 - Python