Latest AI/ML News
165 articles Β· π₯ Community Buzz
A Sana 1.6B te
hey, Iβm currently getting enough VRAM to run something in the GLM-5.2 range, but Iβm wondering: do we actually have a solid ranking that compares cloβ¦
Highlights: 4 x 48GB modded 4090s - 192GB VRAM
Too many times I hear people whine about not being ble to run SOTA models or claim it would require $50k, or $100k. https://www.ebay.com/itm/398079051β¦
My goal has always been to be productive with commodity hard
Hi, I created a repo/site for publishing/sharing .torrent files for popular open models, added web seed support and a few scripts to automate it. Repoβ¦
I run a small gpu lab in the USA and work closely with two factories in china designing/producing 48gb 4090 PCB's. The only recent card weve gotten waβ¦
Got all the parts before the crazy price increase except for the rtx pro 5k! Was saving up to order rtx pro 6000 in US and i
<img alt="Ornith-1.0-35B GGUF update: native MTP speculative-decode graft + full serving/TTFT/long-context numbers (llama.cpp, tp=1)" src="https://preβ¦
  submitted by   /u/Fcking_Chuck <a href="https://git
I've been meaning to post about this. The community has
TL;DR: The (very messy) code and writeups can be found at https://github.com/jakint0sh/qwen3-engine Read the README for instructions on how to get staβ¦
Sharing popular(also recent) models for reference: 151-250B : DeepSeek-V4-Flash Step-3.X-Flash Command-a-plus-05-2026 Laguna-M.1 MiniMax-M2.X Qwen3-23β¦
Just
DeepSpec <a href="https://github.com/deepseek-ai/DeepSpec#de
  submitted by   /u/sammcj <span
  submitted by   /u/pscoutou [link]   [comments]
From: Vladik on π: https
Dario's args: "Opens
A megathread that is overdue! Let's discuss and debate on what the best local agents available today are Prologue First a note on terminology: While mβ¦