Any current Voice2Voice AI model that runs locally that’s good?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
You guys remember sesame AI? With their really good AI voice model? Obviously ChatGPT has their voice model that’s also really good.
Is there any smaller local variant that runs on like consumer grade gpu‘s (12,16 24gb?)
I think NVidia released something but I didn’t really remember much or look for it I think?
Also something new, not something from like 2 years ago, thanks
[link] [comments]
More from r/LocalLLaMA
-
an unscientific qwen 3.8 flash next and glm 5.3 flash comparison
Aug 30
-
Ran Qwen3.8-Flash-Next (79 GB, 2-bit) at 350K ctx for 3.5 hours on a 128 GB M5 Max — speed vs context depth, 100 turns, one graph
Aug 30
-
Nemotron-3.5-Lightning at 11.77 GiB, a 16 GB option for a model that didn't have one
Aug 29
-
Humaneval benchmark for Deepseek V4 Flash 0731 vs GLM5.3 Flash on 2x DGX Spark setup
Aug 29
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.