Latest AI/ML News

770 articles · Reddit r/LocalLLaMA

Reddit r/LocalLLaMASep 22, 2026

This post is written by a human and I'd appreciate it if you treated it as such. Thanks. So, I've been noticing a pretty clear interest in developing…

Reddit r/LocalLLaMASep 21, 2026

Although it makes mistakes and use more tokens, but after some corrections and steering , It(max) gives really good outputs like on par with 5.6 sol a…

Reddit r/LocalLLaMASep 22, 2026

Recently, the new DeepSeek-V4.1-Flash architecture showed how a causal encoder-decoder can work, but it was trained from scratch. Model Grafting does…

Reddit r/LocalLLaMASep 22, 2026

Imagine Git for model fine-tunes that also saves you storage. DeltaTensors compresses fine-tuned model checkpoints by storing the weight difference fr…

Reddit r/LocalLLaMASep 22, 2026

I decided to see what effects recent PRs have had on the performance of the two models I care about, DSv4 Flash 0731 and Qwen 3.8 Flash Next, on my ha…

Reddit r/LocalLLaMASep 21, 2026

This experiment is live, you can inspect all the internal reasoning, memories, attempts here: https://artificium-covering-experiment.gr.bio/ The probl…

Reddit r/LocalLLaMASep 22, 2026

Been playing with judge models in my eval pipeline, so this landed at the right time. Kev is a small family of Jev-architecture decision models (0.8B,…

Reddit r/LocalLLaMASep 21, 2026

https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B That's serious. submitted by /u/Beamsters [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

I shipped something I've been building for the last few weeks : phantom-kv , a refusal-removal system for large language models that doesn't touch a s…

Reddit r/LocalLLaMASep 22, 2026

submitted by /u/tengo_harambe [link] [comments]

Reddit r/LocalLLaMASep 22, 2026

https://preview.redd.it/bpbc9i6hizqh1.png?width=1270&format=png&auto=webp&s=e8aa8301895735a05c3c61a5e793018a23d1cac5 I wanted to share a quick update:…

Reddit r/LocalLLaMASep 21, 2026

Small, free finding. I use local Qwen models for typed decisions: a state plus a question with fixed allowed answers, and I read the probability of ea…

Reddit r/LocalLLaMASep 21, 2026

# the What An engine to run Gemma 4 31B on blackwell under massive concurrency and rather specific workload patterns. I've been waiting for someone to…

Reddit r/LocalLLaMASep 21, 2026

https://preview.redd.it/nulsv53o8vqh1.png?width=4500&format=png&auto=webp&s=74765dbd409f4c221640f9f6000a685f6fdbb242 Spent weekend benchmarking the Sp…

Reddit r/LocalLLaMASep 21, 2026

Since there’s no comparison chart on the model page, I asked Perplexity to compare it against some relatively small open-weight models in a similar si…

Reddit r/LocalLLaMASep 21, 2026

Hey all! I am Aritra from Hugging Face. I wanted to share an update on the `tokenizers` library that we have at Hugging Face. It has gone under major…

Reddit r/LocalLLaMASep 21, 2026

submitted by /u/Bitter-College8786 [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

we're so back?!? submitted by /u/VoiceApprehensive893 [link] [comments]

Reddit r/LocalLLaMASep 21, 2026

Isn't this what simple neural networks have been able to do for years? Doesn't seem anything special to me. submitted by /u/Manerfish [link] [comments…