Latest AI/ML News

770 articles · Reddit r/LocalLLaMA

Reddit r/LocalLLaMAAug 19, 2026

I hacked this together so there's probably more on the table in terms of performance. Measured with the Club-3090 canonical bench suite (bench.sh, 3 w…

Reddit r/LocalLLaMAAug 19, 2026

Community manager mentioned this in the Qwen Ambassador Discord, put an X reaction on someone asking for 35B... and said We'll have a new midsize open…

Reddit r/LocalLLaMAAug 19, 2026

Thoughts About Scaling Law Scaling, but not only of parameters. Every model release now ends with the same question: how many parameters? It isn't a q…

Reddit r/LocalLLaMAAug 10, 2026

Wowee!! Just when you thought it couldn't get better for open weight models, we probably have had our best period yet!?!?! Models that rival the close…

Reddit r/LocalLLaMAAug 15, 2026

Megathread to help with the influx of duplicate / similar posts around the release of the Qwen 3.8 27B release. Quants Fine-Tunes & Abliterations Chat…

Reddit r/LocalLLaMAJun 28, 2026

I have been running an open evaluation setup where N models answer the same prompt, then blind-grade each other in an N x N matrix with self-judgments…

Reddit r/LocalLLaMAJun 28, 2026

Would it not be possible to create crowd sourced, truly open sourced distilled LLMs with a simple wrapper around command line based AI services that e…

Reddit r/LocalLLaMAJun 27, 2026

If you use Claude Code, every session is already sitting on disk as a .jsonl file under ~/.claude/projects/ . It has real coding conversations: multi-…

Reddit r/LocalLLaMAJun 27, 2026

https://techcrunch.com/2026/06/26/openai-limits-gpt-5-6-rollout-after-government-request-says-restrictions-shouldnt-be-the-norm/ Either a hype before…

Reddit r/LocalLLaMAJun 28, 2026

I'm building a local LLM workstation and would appreciate some advice from people already running 2×3090s. Current hardware: ASUS Crosshair VIII Hero…

Reddit r/LocalLLaMAJun 28, 2026

A Sana 1.6B te

Reddit r/LocalLLaMAJun 28, 2026

hey, I’m currently getting enough VRAM to run something in the GLM-5.2 range, but I’m wondering: do we actually have a solid ranking that compares clo…

Reddit r/LocalLLaMAJun 27, 2026

Highlights: 4 x 48GB modded 4090s - 192GB VRAM

Reddit r/LocalLLaMAJun 27, 2026

Too many times I hear people whine about not being ble to run SOTA models or claim it would require $50k, or $100k. https://www.ebay.com/itm/398079051…

Reddit r/LocalLLaMAJun 28, 2026

My goal has always been to be productive with commodity hard

Reddit r/LocalLLaMAJun 27, 2026

Hi, I created a repo/site for publishing/sharing .torrent files for popular open models, added web seed support and a few scripts to automate it. Repo…

Reddit r/LocalLLaMAJun 27, 2026

I run a small gpu lab in the USA and work closely with two factories in china designing/producing 48gb 4090 PCB's. The only recent card weve gotten wa…

Reddit r/LocalLLaMAJun 28, 2026

Got all the parts before the crazy price increase except for the rtx pro 5k! Was saving up to order rtx pro 6000 in US and i

Reddit r/LocalLLaMAJun 28, 2026

<img alt="Ornith-1.0-35B GGUF update: native MTP speculative-decode graft + full serving/TTFT/long-context numbers (llama.cpp, tp=1)" src="https://pre…

Reddit r/LocalLLaMAJun 28, 2026

&#32; submitted by &#32; /u/Fcking_Chuck <a href="https://git