Real-time local voice AI speak and it talks back, running VAD/STT/TTS on your own GPU with a cloud LLM brain. Sub-second latency, works on 4GB VRAM.
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).