If you’re using the Home Assistant voice assistant mechanism (not Alexa/Google/etc.) how’s it working for you? Given there’s a number of knobs that you can use, what do you use and what works well? - Wake word model. There’s the default models and custom - Conservation agent and model - Speech to text models (e.g. speech-to-phrase or whisper) - Text to speech models

  • AvocadoSandwich@eviltoast.org
    link
    fedilink
    English
    arrow-up
    2
    ·
    16 天前

    It works reasonably well. Not any better or worse then Google Assistant did. Setup:

    • Raspberry Pi 5 to run home assistant and Piper
    • voice assistant preview
    • standard “Mycroft” wake word
    • Whisper small on proxmox server running 2 cores
    • Using English
    • LLM running granite4.1:3b on a laptop with gtx1050 mobile (a bit slow, but works) Intent first, LLM as fallback

    I use it to control the TV, play music, control vacuum and control lights. I have a script for the TV functions and a script that controls different scenes in my house which are both exposed. This works really well as I have commands setup to trigger automations or I can just ask it to do something via the LLM. This last one takes a bit longer (up to 45 seconds or so) but it is workable as things rarely have to happen immediately. I’m upgrading the laptop to a desktop with gtx1660S soon though which will be a huge improvement.

    I also have automations setup that automatically lowers music when the wake word triggers and goes back to volume when it is done listening which helps tremendously as the Voice assistant preview does have trouble understanding in a noisy room. The LLM also partially solved this though

  • ItsPronouncedZed@lemmy.ca
    link
    fedilink
    English
    arrow-up
    0
    ·
    7 个月前

    I’d like a PCB designed as a drop-in replacement for labotomized Echo or Nest devices so we can reuse their existing hardware and recycle millions of older units.