Building A Smart Speaker Outside The Corporate Cloud

November 17, 2025

If you’re not worried about corporate surveillance bots scraping your shopping list and manipulating you through marketing, you can buy any number of off-the-shelf smart speakers for your home. Alternatively, you can roll your own like [arpy8] did, and keep your life a little more private.

The build is based around an ESP32 microcontroller. It connects to the ‘net via its inbuilt Wi-Fi connection, and listens out for your voice with an INMP441 omnidirectional microphone module. The audio data is trucked off to a backend server running a Whisper speech-to-text model. The text is then passed to Google’s Gemini 2.5 Flash large language model. The response generated is passed to the Piper Neural Voice text-to-speech engine, sent back to the ESP32, and spat out via the device’s DAC output and a speaker attached to an LM386 amplifier. Basically, anything you could ask Gemini, you can do with this device.

By virtue of using a commercial large language model, it’s not perfectly private by any means. Still, it’s at least a little farther removed than using a smart speaker that’s directly logged in to your Amazon/Google/Hulu/Beanstikk account. Files are on Github for those eager to dive into the code. We’ve seen some other fun builds along these lines before, too. Video after the break.

6 thoughts on “Building A Smart Speaker Outside The Corporate Cloud”

Pat says:

November 17, 2025 at 1:23 pm

This: “worried about corporate surveillance bots scraping your shopping list”

Followed by: “Basically, anything you could ask Gemini”

…. does… not… compute

Report comment

Reply
Nikolai says:

November 17, 2025 at 1:39 pm

The main problem is not about corporate surveillance, but any IOT on the local network is the potential backdoor for hackers regardless how good is your firewall.

Report comment

Reply
Clara Hobbs says:

November 17, 2025 at 1:53 pm

Misleading title. This isn’t “outside the corporate cloud,” since it uses a corporate-cloud-based LLM.

Report comment

Reply
shadester says:

November 17, 2025 at 2:10 pm

Should be able to use a local LLM, but it would be slower.

Report comment

Reply
Make says:

November 17, 2025 at 2:35 pm

Perhaps https://www.home-assistant.io/integrations/ollama

Report comment

Reply
Leonardo says:

November 17, 2025 at 3:02 pm

Please stop with the clickbait.

Report comment

Reply

Hackaday

Building A Smart Speaker Outside The Corporate Cloud

6 thoughts on “Building A Smart Speaker Outside The Corporate Cloud”

Leave a ReplyCancel reply

Search

Never miss a hack

If you missed it

Tech In Plain Sight: Pneumatic Tubes

If IRobot Falls, Hackers Are Ready To Wrangle Roombas

Moving From Windows To FreeBSD As The Linux Chaos Alternative

“AI, Make Me A Degree Certificate”

Japan’s Forgotten Analog HDTV Standard Was Well Ahead Of Its Time

Our Columns

Keebin’ With Kristina: The One With The Cipher-Capable Typewriter

Hackaday Links: November 16, 2025

The Value Of A Worked Example

Hackaday Podcast Episode 345: A Stunning Lightsaber, Two Extreme Cameras, And Wrangling Roombas

This Week In Security: Landfall, Imunify AV, And Sudo Rust

6 thoughts on “Building A Smart Speaker Outside The Corporate Cloud”

Leave a ReplyCancel reply

Search

Never miss a hack

Subscribe

If you missed it

Our Columns