AI Voice Assistant — talk to your Homey
Now certified and live in the Homey App Store!
This app turns small, inexpensive ESPHome voice devices into a natural voice interface for Homey. Say the wake word (e.g. “Okay Nabu”), speak normally, and the assistant controls your devices, answers questions, plays music, sets timers and speaks back — in your language. It keeps listening after it answers, so you can ask follow-up questions without repeating the wake word.
What you can do by voice
- Control your smart home — lights, plugs, dimming, thermostats, in any room/zone
- Lock & unlock doors — locking always works; unlocking is off by default for safety (opt-in in settings, and never more than one lock per command)
- Weather — current conditions and forecasts for your location
- Timers & alarms — “set a 10 minute timer” — the device counts down on its LED ring and chimes when time’s up
- Search the web — “what’s playing at the cinema tonight?” — via OpenAI web search or the Brave Search API
- Manage your Bring! shopping list — “add milk”, “what’s on the list?” (opt-in)
- Play music — “play Abbey Road by the Beatles”, “next song”, “play some Queen in the kitchen” — through a Music Assistant server on your network; the audio streams from the server straight to the speaker
- Ask anything — general questions, time & date, and “what can you do?” for help
The assistant understands and replies in English, Dutch, German, French, Italian, Swedish, Norwegian, Spanish, Danish, Russian, Polish and Korean.
You choose the AI engine
- OpenAI Realtime — low-latency cloud speech-to-speech (full or mini model), with advanced tuning for speech detection sensitivity and pause length
- Google Gemini Live — the same real-time pipeline, powered by Gemini
- Mistral (Voxtral) — the European alternative on one API key: streaming Voxtral speech recognition that transcribes while you talk, a Mistral chat model, and Voxtral voices
- Custom / self-hosted pipeline — Whisper, Ollama, LM Studio, Piper, Wyoming, and any OpenAI-compatible server (Groq, OpenRouter, llama.cpp, vLLM, …). Every stage (speech-to-text, language model, text-to-speech) is independently pluggable, with test buttons in the settings — and with a fully local setup, no audio ever leaves your network
Supported hardware
- Home Assistant Voice Preview Edition — stock ESPHome firmware, no Home Assistant needed
- ThirdReality Voice & Music Assistant — works out of the box, no flashing at all, and doubles as a Music Assistant multi-room speaker
- Seeed ReSpeaker XVF3800 (XIAO ESP32S3) — for the tinkerers: a DIY satellite with a 4-mic XMOS array doing echo cancellation and beamforming on-chip; you compile and flash the community ESPHome firmware yourself
- XiaoZhi AI devices — the cheap ESP32-S3 gadgets in many shapes, running RealDeco’s ESPHome firmware
The app talks to the devices directly over your LAN using the ESPHome native API — no Home Assistant installation is needed. Devices with an ESPHome API encryption key (e.g. previously adopted by Home Assistant) are fully supported with encrypted connections.
Pairing is easy, even for brand-new devices: the pairing wizard can put an out-of-the-box Voice PE or ThirdReality onto your Wi-Fi via Bluetooth — no vendor apps or cables. Devices already on the network are found automatically, and there’s manual IP entry as a fallback for networks where mDNS doesn’t reach the Homey.
Flow support
Flow cards include Ask the assistant (answer as a text tag, or spoken on the device), Say (text-to-speech), Play an audio URL, timer start/cancel plus timer triggers, and Heard something / Thinking triggers that show what the assistant heard, which tools it used and what it replied — great for following a whole conversation on the timeline.
Requirements
- A compatible ESPHome voice device (see above)
- An API key for your chosen cloud engine (OpenAI, Google Gemini or Mistral), or your own local AI services
- Optional: a Music Assistant server for music playback
Links
App Store: AI Voice Assistant | Homey
GitHub (setup guides, docs, source): GitHub - arvebjoe/no.arvebjoe.ai-voice-assistant · GitHub
Issues & feature requests: Issues · arvebjoe/no.arvebjoe.ai-voice-assistant · GitHub
Feedback, questions and device reports are very welcome in this topic!





