Latest AI/ML News

770 articles · Reddit r/LocalLLaMA

Reddit r/LocalLLaMASep 21, 2026

Quote: DeepSeek is training a 2T-parameter model and plans to eventually build an 8T-parameter model. https://x.com/wallstengine/status/21019828436563…

Reddit r/LocalLLaMASep 21, 2026

submitted by /u/Bestlife73 [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

I’ve been working on making small models more capable at agentic coding and work, because most people in the world don’t have the sort of hardware nee…

Reddit r/LocalLLaMASep 21, 2026

submitted by /u/themixtergames [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

https://huggingface.co/yandex/AliceAI-Foundation-80B-A3B-Base It's not a Qwen3 finetune, it's actually its own fully custom architecture. No Llama.cpp…

Reddit r/LocalLLaMASep 21, 2026

Hey everyone! It has been quite a while since the last SupraLabs model - but today we've something special for y'all: Supra2-IMG It's a 100M parameter…

Reddit r/LocalLLaMASep 21, 2026

This sub is, needless to say very niche and skewed towards the high end. There are tons of extremely high end setups here with multiple gpu's etc. Eve…

Reddit r/LocalLLaMASep 21, 2026

submitted by /u/Bestlife73 [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

submitted by /u/Hyacin75 [link] [comments]

Reddit r/LocalLLaMASep 20, 2026

submitted by /u/johnnyApplePRNG [link] [comments]

Reddit r/LocalLLaMASep 20, 2026

After seeing a few posts on here about it, I finally tried exl3 3bpw and exllamav3 for running flash next - with amazing results. On 3x3090s, 128GB DD…

Reddit r/LocalLLaMASep 21, 2026

I have been reasonably satisfied with my single R9700 (32GB) as I can run practical quants of Qwen 3.8-27B at good speeds, as well as other similar mo…

Reddit r/LocalLLaMASep 20, 2026

I know cloud subscription plans are not exactly the core focus of r/LocalLLaMA , so I want to be clear about why I’m posting this here. I’ve been buil…

Reddit r/LocalLLaMASep 20, 2026

Another day, another Qwen Flash Next speedup submitted by /u/jacek2023 [link] [comments]

Reddit r/LocalLLaMASep 20, 2026

After seeing u/Nandakishor_ml’s post introducing Laya , I wanted to see how fast it could run in a standalone C++ implementation. Credit to u/Nandakis…

Reddit r/LocalLLaMASep 20, 2026

submitted by /u/johnnyApplePRNG [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

Laya is an open-weight (Apache 2.0) "System 1" decision model from Convai Innovations, built by Nandakishor M as an open alternative to TypeSafe's clo…

Reddit r/LocalLLaMASep 21, 2026

so I was doing some stuffs in Deepseek harness with Qwen 3.8 27b and just glanced over to see the progress and saw this, make me chuckle 😄 submitted…

Reddit r/LocalLLaMASep 20, 2026

I’ve spent basically the last 8 hours testing different models on the exact same web-development prompt, and I finally finished. The whole point of th…