Ultra-low bitrate speech codec (0.27-1 kbps) with cross-modal alignment and real-time capabilities
-
Updated
Aug 27, 2025 - Python
Ultra-low bitrate speech codec (0.27-1 kbps) with cross-modal alignment and real-time capabilities
Ultra-low-bitrate Speech Codec for Speech Language Modeling Applications
ICASSP 2024 - Generative De-Quantization for Neural Speech Codec via Latent Diffusion.
microphone array speech generator (MASG) in room acoustic
AudioCodec-Hub is a Python library for encoding and decoding audio data, supporting various neural audio codec models
3GPP Codec for Enhanced Voice Services (EVS)
Official repo of ICASSP 2021 paper Source-Aware Neural Speech Coding for Noisy Speech Compression (SANAC)
Official inference code and checkpoints for CodecSlime.
test scripts for test lyra voice codec
Hide digital data inside speech-shaped audio that survives Zoom, Discord, WhatsApp, and cellular voice. Reproducible Pareto curve of six trained codecs spanning 76 bps (cellular) to 3196 bps (Zoom-class) with listenable demos.
Pure-Rust Opus audio codec (RFC 6716) with Ogg encapsulation (RFC 7845). Encoder and decoder for SILK, CELT and hybrid modes. No C, no FFI, no dependencies.
A few seconds of speech inside an ordinary QR code. Open format, CVQR1: payloads, decoded entirely offline.
To associate your repository with the speech-codec topic, visit your repo's landing page and select "manage topics."