view article Article Hugging Face and Cerebras bring Gemma 4 to real-time voice AI +2 A-Mahla, andito, lvwerra, vyassaurabh • 25 days ago • 85
A Computational Analysis of Real-World DJ Mixes using Mix-To-Track Subsequence Alignment Paper • 2008.10267 • Published Aug 24, 2020 • 1
Moshi: a speech-text foundation model for real-time dialogue Paper • 2410.00037 • Published Sep 17, 2024 • 18
Moshi v0.1 Release Collection MLX, Candle & PyTorch model checkpoints released as part of the Moshi release from Kyutai. Run inference via: https://github.com/kyutai-labs/moshi • 16 items • Updated 11 days ago • 245
NeuTTS Nano Multilingual Collection Collection NeuTTS Nano is a TTS model, 3x smaller than NeuTTS Air, that runs on CPU in real-time - now in English, Spanish, French, and German versions! • 13 items • Updated 5 days ago • 19
🎵 The MusicBox Collection A collection full of musical tasks demos, for musicians & music enthusiasts • 42 items • Updated 13 days ago • 34
Music Mixing Style Transfer: A Contrastive Learning Approach to Disentangle Audio Effects Paper • 2211.02247 • Published Nov 4, 2022 • 5