Latest AI/ML News

770 articles · Reddit r/LocalLLaMA

Reddit r/LocalLLaMASep 10, 2026

so i saw that openui.com released OUI-1, a model fine-tuned on DiffusionGemma. the training dataset uses OpenUI-Lang, a custom DSL (domain-specific la…

Reddit r/LocalLLaMASep 10, 2026

I'm currently building my system around 3060s, but I might be able to get a 4060 for a nice deal. At first it seemed like a no brainer, but turns out…

Reddit r/LocalLLaMASep 10, 2026

We replaced the monolithic native modules with TypeScript pipelines you can inspect. 🔧 It runs across all major silicon backends and makes it easier…

Reddit r/LocalLLaMASep 10, 2026

hey guys, I've been working on a reading app for several months now and had problems with getting good quality TTS, the options were kokoro, kitten, p…

Reddit r/LocalLLaMASep 10, 2026

submitted by /u/paf1138 [link] [comments]

Reddit r/LocalLLaMASep 10, 2026

submitted by /u/SirReal14 [link] [comments]

Reddit r/LocalLLaMASep 10, 2026

I have been trying out these models (mostlt GLM 5.3 flash) using different harnesses, but I'm trying to review what these models are exceptionally goo…

Reddit r/LocalLLaMASep 10, 2026

Hi builders, What would be the the best small local models for coding? Are Gemma 4 and Qwen3.8 27B Gemma 4 26B / 31B enough for local development? And…

Reddit r/LocalLLaMASep 10, 2026

submitted by /u/incarnadine72 [link] [comments]

Reddit r/LocalLLaMASep 9, 2026

Surveillance plagiarism - Hosted AI company pumps their stock price by training upon researchers' AI sessions, so that their internal model can solve…

Reddit r/LocalLLaMASep 10, 2026

I initially saw CED as just an efficiency improvement, but the more I read about it, the more it feels like an inference architecture leap. The encode…

Reddit r/LocalLLaMASep 10, 2026

Long time lurker, but I'm finally upgrading to 128GB VRAM, and I'm trying to figure out what to run. I had been leaning towards Qwen3.8 Flash-Next at…

Reddit r/LocalLLaMASep 10, 2026

I am just sharing my config for Qwen 3.8 27b that fits on a 5060TI, what is cool about this is that you can even get vision! and a 85K context (I have…

Reddit r/LocalLLaMASep 10, 2026

If the ratio is the same, Maybe 1.6T -3.1T params plus .56T-1.06T engrams and fable 5.0 level performance? Maybe v4.2 or 4.5 will have engram gradient…

Reddit r/LocalLLaMASep 10, 2026

Trying to get Hermes a local, efficient, tts voice. submitted by /u/JLeonsarmiento [link] [comments]

Reddit r/LocalLLaMASep 10, 2026

I was not aware that the harness makes such a big difference. DeepSeek V4.1 Flash submitted by /u/Specific-Rub-7250 [link] [comments]

Reddit r/LocalLLaMASep 10, 2026

I like to benchmark new models that come out on motion videos. So here's a test I did for deepseek v4.1 flash. And I have to say flash has probably gr…

Reddit r/LocalLLaMASep 9, 2026

It seems to use a 96-bit LPDDR5X memory bus, instead of the previous 64-bit wide busses. Considering it's on 2nm, that's expensive silicon. That shoul…

Reddit r/LocalLLaMASep 10, 2026

Original Source from DeepSeek WeChat Official Account: https://mp.weixin.qq.com/s/qg0NU3NNUbp1co2PdkAPAg Today we're officially releasing the DeepSeek…

Reddit r/LocalLLaMASep 10, 2026

Hey y'all! We've released a new model in our lineup: GigaChat-3.5 Reasoning. It's a 432B-A28B MoE with Gated DeltaNet for long-context efficiency. We…