Latest AI/ML News

2649 articles · 👥 Community Buzz

Reddit r/LocalLLaMASep 20, 2026

I’ve spent basically the last 8 hours testing different models on the exact same web-development prompt, and I finally finished. The whole point of th…

Reddit r/LocalLLaMASep 21, 2026

Is it worth switching to from Ornith 1.5 9B? Or is there another similarly sized model that's beating both of them? (I head K2 is good but the KV cach…

Reddit r/LocalLLaMASep 20, 2026

​ ​I wanted to share a quick update and performance video running the full Moonshot AI Kimi K3 (moonshotai/Kimi-K3) model across my 16x GB10 cluster.…

Reddit r/LocalLLaMASep 20, 2026

I am seeing this JEV everywhere since yesterday in Localllama and it is passing past my head on what it is? So like what is it? Some new LLM? Or is it…

Reddit r/LocalLLaMASep 20, 2026

So this jev thingy is getting kind of big... tbh it seems overhyped by a large margin, but here we are. Not that its bad, just feels like we usually i…

Reddit r/LocalLLaMASep 20, 2026

TL;DR: Local agent loop, ~21 days, one RTX 3090. Task was pretty much "build a CUDA inference engine for optimized for yourself on this GPU arch." Got…

Reddit r/LocalLLaMASep 21, 2026

https://x.com/QwenDevs/status/2101917379785838660 submitted by /u/Bestlife73 [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

You can simply run any GGUF with llama.cpp with n_predict=1 and n_probs=10, disable reasoning, and prompt it such as "If the following email is spam,…

Reddit r/LocalLLaMASep 20, 2026

To test what it can do. Qwen3.8-Flash-Next Intel Autoround W4A16 running locally on 4xV620 ~2k prefill and 70ts decode.. Were running around 3 hours.…

Reddit r/LocalLLaMASep 20, 2026

submitted by /u/fallingdowndizzyvr [link] [comments]

Reddit r/LocalLLaMASep 20, 2026

So my brother and I both use LLMs for coding. I've started using a local GLM 5.3 Flash instance - q4 qat. My brother uses GPT-6-Astra as his daily dri…

Reddit r/LocalLLaMASep 21, 2026

Sorry for the pretentious name, I know, I know.. It just contains all the pieces I would like to see a AGI model to have, and I can't stand the tempta…

Reddit r/LocalLLaMASep 20, 2026

Just find it interesting, since if Taalas tried to get into the consumer market, it could be interesting. Kind of like how we have game discs on CDs.…

Reddit r/LocalLLaMASep 20, 2026

Meet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨 A unified model for bo…

Reddit r/LocalLLaMASep 21, 2026

ZCode is now open source , and the reported security issues have been addressed. Source code: https://github.com/zai-org/ZCode The repo includes its d…

Reddit r/MachineLearningSep 21, 2026

The ICLR review policy says if your name appears on 3 or more papers, you will need to serve as a reviewer, and it did not say anything about qualific…

← PreviousPage 36 of 133Next →