I have tried open web UI and it's great when it works but for some reason it's not super reliable or always fast enough for me.

I'm just wondering if there's any simpler tools or projects to integrate voice mode with Ollama models that people have found it and like. Thanks!

you are viewing a single comment's thread
view the rest of the comments
[–] 1 point 3 weeks ago (2 children)

That seems to heavy resources for the hardware I have. I did run across these today. Kokoro-82M behind an OpenAI-compatible server

https://localaimaster.com/blog/best-local-tts-models

Kyutai Unmute

https://github.com/kyutai-labs/unmute

I haven't has a chance to start playing with them though

  • source
  • parent
  • hideshow 4 child comments
  • [–] [S] 2 points 3 weeks ago (1 child)

    Oh cool, yeah I'm actually pivoting to kokoro now

  • source
  • parent
  • hideshow 2 child comments