Diego.

2026

Voice Assistant

100% local voice AI

  • GGML
  • Vulkan
  • whisper.cpp
  • llama.cpp
The voice assistant interface

Full-duplex voice assistant with screen vision for Windows, with interruptible conversation and screen-aware answers, fully offline, no cloud and no API keys.

A voice assistant for Windows with real full-duplex audio: you can interrupt it mid-sentence, and the interruption becomes the next question. The moment you speak, it captures the screen and uses the image as context for its answer.

The whole pipeline runs locally on a single GPU stack (GGML + Vulkan) shared by speech recognition and the LLM, which keeps it portable across AMD, NVIDIA and Intel GPUs. The target is under 2 seconds from the end of your speech to the first syllable of the reply.

Full documentation on GitHub ↗
Next projectToolHaven