Latest AI/ML News
770 articles · Reddit r/LocalLLaMA
The missing control is visible in sudoingX’s Ling-3.0-flash benchmark graphics. The earlier table leaves Ling’s no-speculation baseline as “not measur…
I feel like this year has been insane , the speed of AI race is something that normal human can't catch up anymore submitted by /u/fugogugo [link] [co…
I was recently trying to load a larger model on my Mac and was going through the drill of checking RAM usage, closing apps like Telegram, Messages, Sp…
I've hit again a point where me as a developer have to take a break from all this slop shit. Im a Developer for 13+ years and i loved it. But i fell f…
https://preview.redd.it/qwmh23nfr3oh1.png?width=1062&format=png&auto=webp&s=4741852dc3f5e7d86ae85281b089b2a325a1804a I mean they aren't that low but s…
This weekend I posted about the gap closing between frontier models and open source models. Well, now I'm coming with receipts. I've been running loca…
How it works. The model being tested is given a server capable of running it's weights and full context. That server is placed in a median priced apar…
Quick setup on linux: STEP 0: install/download llama.cpp, Freecad, your favourite gguf model - possibly with multimedia image reading capabilities (I'…
Model: DeepSeek-V4-Flash-Vision-Exp (local and API when impatient) Time: about one weekend (2 days) of QA and small improvements Full game is here Aft…
Curious about what people are preferring, if you have the hardware. I have m3 Max 96gb and both run, and largely feel identical, but prefill on qwen 2…
Hi all! I just wanna say that I am tired lol. Yes, it's another harness, but I spent a lot of time and effort and have forsaken my hobbies to build th…
OpenBMB's MiniCPM5-2B scores 15 on the Artificial Analysis Intelligence Index v4.2, the highest of any open weights model at 4B parameters or below Hu…
It just seems every local 30b class model is just trying so hard to be the next Qwen that they all just kinda blend into a mass of code focused models…
submitted by /u/rm-rf-rm [link] [comments]
So I used to use Artificial Analysis to compare models. The recent 61 score for Astra had me look into the results more granularly and I was very disa…
I've got an refurb Dell R740 running Proxmox that I put a Tesla T4 in, mainly to run some CTC local transcription work, but thought it would be fun to…
I'm one of the developers. We said in August it would go open source in September and it did last night. MIT or Apache-2.0, pick one. The repo you see…
Here is the news you may have missed: Nvidia already demonstrated 100% on ARC AGI-3 , using their novel harness AVO: https://developer.nvidia.com/blog…
How good/bad is it against comparable MoEs the same size? How does it compare against Qwen 3.6 35BA3B? Since we don't have 3.8 35B this seems like an…
Quality: 81.1 → 87.7 (+8%) Speed: 35 → 29 tok/s (−16%) Runtime: 8m51s → 44m39s (5x longer) Output tokens: 18K → 78K (🤯) Noticeably better quality, bu…