.
@aconchillo has been hacking on an experimental macOS voice assistant. Audio transcription and wake word monitoring run locally. LLM and screen vision use cloud models by default, but of course you can swap in a local model/endpoint.
"Peekaboo, tell me when the build finishes."
"Peekaboo, somebody invited me to a meetup about inference optimization tonight, but I can't find the message. Was that in Slack, email, WhatsApp, or iMessage?"
"Peekaboo, did we do the evals last week for the new Nemotron model on boule, pancake, flatbread, or the dgx spark? I think we forgot to push to the repo."