Latest AI/ML News
770 articles · Reddit r/LocalLLaMA
so i saw that openui.com released OUI-1, a model fine-tuned on DiffusionGemma. the training dataset uses OpenUI-Lang, a custom DSL (domain-specific la…
I'm currently building my system around 3060s, but I might be able to get a 4060 for a nice deal. At first it seemed like a no brainer, but turns out…
We replaced the monolithic native modules with TypeScript pipelines you can inspect. 🔧 It runs across all major silicon backends and makes it easier…
hey guys, I've been working on a reading app for several months now and had problems with getting good quality TTS, the options were kokoro, kitten, p…
submitted by /u/paf1138 [link] [comments]
submitted by /u/SirReal14 [link] [comments]
I have been trying out these models (mostlt GLM 5.3 flash) using different harnesses, but I'm trying to review what these models are exceptionally goo…
Hi builders, What would be the the best small local models for coding? Are Gemma 4 and Qwen3.8 27B Gemma 4 26B / 31B enough for local development? And…
submitted by /u/incarnadine72 [link] [comments]
Surveillance plagiarism - Hosted AI company pumps their stock price by training upon researchers' AI sessions, so that their internal model can solve…
I initially saw CED as just an efficiency improvement, but the more I read about it, the more it feels like an inference architecture leap. The encode…
Long time lurker, but I'm finally upgrading to 128GB VRAM, and I'm trying to figure out what to run. I had been leaning towards Qwen3.8 Flash-Next at…
I am just sharing my config for Qwen 3.8 27b that fits on a 5060TI, what is cool about this is that you can even get vision! and a 85K context (I have…
If the ratio is the same, Maybe 1.6T -3.1T params plus .56T-1.06T engrams and fable 5.0 level performance? Maybe v4.2 or 4.5 will have engram gradient…
Trying to get Hermes a local, efficient, tts voice. submitted by /u/JLeonsarmiento [link] [comments]
I was not aware that the harness makes such a big difference. DeepSeek V4.1 Flash submitted by /u/Specific-Rub-7250 [link] [comments]
I like to benchmark new models that come out on motion videos. So here's a test I did for deepseek v4.1 flash. And I have to say flash has probably gr…
It seems to use a 96-bit LPDDR5X memory bus, instead of the previous 64-bit wide busses. Considering it's on 2nm, that's expensive silicon. That shoul…
Original Source from DeepSeek WeChat Official Account: https://mp.weixin.qq.com/s/qg0NU3NNUbp1co2PdkAPAg Today we're officially releasing the DeepSeek…
Hey y'all! We've released a new model in our lineup: GigaChat-3.5 Reasoning. It's a 432B-A28B MoE with Gated DeltaNet for long-context efficiency. We…