r/LocalLLaMA · · 1 min read

🟩 NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

🟩 NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp

🐦‍⬛ Magpie-TTS Multilingual

🦜 Nemotron Speech Streaming EN 0.6B

🦜 Nemotron-3.5 ASR Streaming

🦜 Parakeet CTC 1.1B

🦜 Parakeet TDT 0.6B v3

🥦 NanoCodec Merged PR

https://huggingface.co/nvidia/magpie_tts_multilingual_357m#run-magpietts-locally-with-nemo-speechcpp

I run open source llm models via AI Desktop XP but how am I supposed to run these models on my phone? Any recc?

submitted by /u/ImaginaryRea1ity
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA