Latest AI/ML News

770 articles · Reddit r/LocalLLaMA

Reddit r/LocalLLaMASep 7, 2026

The missing control is visible in sudoingX’s Ling-3.0-flash benchmark graphics. The earlier table leaves Ling’s no-speculation baseline as “not measur…

Reddit r/LocalLLaMASep 7, 2026

I feel like this year has been insane , the speed of AI race is something that normal human can't catch up anymore submitted by /u/fugogugo [link] [co…

Reddit r/LocalLLaMASep 7, 2026

I was recently trying to load a larger model on my Mac and was going through the drill of checking RAM usage, closing apps like Telegram, Messages, Sp…

Reddit r/LocalLLaMASep 7, 2026

I've hit again a point where me as a developer have to take a break from all this slop shit. Im a Developer for 13+ years and i loved it. But i fell f…

Reddit r/LocalLLaMASep 7, 2026

https://preview.redd.it/qwmh23nfr3oh1.png?width=1062&format=png&auto=webp&s=4741852dc3f5e7d86ae85281b089b2a325a1804a I mean they aren't that low but s…

Reddit r/LocalLLaMASep 7, 2026

This weekend I posted about the gap closing between frontier models and open source models. Well, now I'm coming with receipts. I've been running loca…

Reddit r/LocalLLaMASep 7, 2026

How it works. The model being tested is given a server capable of running it's weights and full context. That server is placed in a median priced apar…

Reddit r/LocalLLaMASep 7, 2026

Quick setup on linux: STEP 0: install/download llama.cpp, Freecad, your favourite gguf model - possibly with multimedia image reading capabilities (I'…

Reddit r/LocalLLaMASep 7, 2026

Model: DeepSeek-V4-Flash-Vision-Exp (local and API when impatient) Time: about one weekend (2 days) of QA and small improvements Full game is here Aft…

Reddit r/LocalLLaMASep 7, 2026

Curious about what people are preferring, if you have the hardware. I have m3 Max 96gb and both run, and largely feel identical, but prefill on qwen 2…

Reddit r/LocalLLaMASep 7, 2026

Hi all! I just wanna say that I am tired lol. Yes, it's another harness, but I spent a lot of time and effort and have forsaken my hobbies to build th…

Reddit r/LocalLLaMASep 7, 2026

OpenBMB's MiniCPM5-2B scores 15 on the Artificial Analysis Intelligence Index v4.2, the highest of any open weights model at 4B parameters or below Hu…

Reddit r/LocalLLaMASep 7, 2026

It just seems every local 30b class model is just trying so hard to be the next Qwen that they all just kinda blend into a mass of code focused models…

Reddit r/LocalLLaMASep 7, 2026

submitted by /u/rm-rf-rm [link] [comments]

Reddit r/LocalLLaMASep 4, 2026

So I used to use Artificial Analysis to compare models. The recent 61 score for Astra had me look into the results more granularly and I was very disa…

Reddit r/LocalLLaMASep 4, 2026

I've got an refurb Dell R740 running Proxmox that I put a Tesla T4 in, mainly to run some CTC local transcription work, but thought it would be fun to…

Reddit r/LocalLLaMASep 4, 2026

I'm one of the developers. We said in August it would go open source in September and it did last night. MIT or Apache-2.0, pick one. The repo you see…

Reddit r/LocalLLaMASep 4, 2026

Here is the news you may have missed: Nvidia already demonstrated 100% on ARC AGI-3 , using their novel harness AVO: https://developer.nvidia.com/blog…

Reddit r/LocalLLaMASep 4, 2026

How good/bad is it against comparable MoEs the same size? How does it compare against Qwen 3.6 35BA3B? Since we don't have 3.8 35B this seems like an…

Reddit r/LocalLLaMASep 3, 2026

Quality: 81.1 → 87.7 (+8%) Speed: 35 → 29 tok/s (−16%) Runtime: 8m51s → 44m39s (5x longer) Output tokens: 18K → 78K (🤯) Noticeably better quality, bu…