Reddit r/LocalLLaMASeptember 11, 2026
GPT Live clone on an RTX 3060
Excerpt
I wanted to see how my fully local home voice assistant compared to the latest GPT Live, so I tested it using the same conversation used in their "Improved Intelligence" demo . In this video they ask the AI to see if a flight route is feasible and while it is figuring that out they continue to ask it questions about what they can eat at each destination. The models I ran are (all squeezed into 12 GB VRAM): Speech recognition: Qwen3 1.7B ASR PyTorch LLM: Qwen3.5-9B-UD-Q4_K_XL GGUF with 12K contex