Latest AI/ML News
770 articles · Reddit r/LocalLLaMA
So when Qwen3.8 27b dropped they were hinting for another model which is Qwen3.8-next-flash , i was hoping for something more light like Qwen 3.6 35b…
Unsloth GGUFs and Vision support https://huggingface.co/unsloth/DeepSeek-V4-Flash-Vision-Exp-GGUF submitted by /u/fmillar [link] [comments]
saw the post the other day where people said Minecraft clones aren't impressive anymore, because at this point the whole thing might as well be in the…
Language-Native Control: Composes character and camera actions into textual instructions and injects them through MiniMax-H3’s pretrained text pathway…
I'm planning to upgrade my workstation (linux with 5700X/64GB DDR4) for local inference and pytorch training. I'm trying to decide between: 2× AMD Rad…
I see a lot of people buy DGX Sparks, and turn them in to clusters to run large models. Wouldn't it be better to invest $16k into an AMD Epyc server w…
Can't wait to test! This should significantly boost TPS! Now we just need more llama cpp optimizations to be merged in! Edit: For anyone who wants to…
Hello all you smarter people. I recently retired and have taken on a task that is going to stretch me a bit. TL;DR My aging aunt is going blind and wa…
Based on https://huggingface.co/hardware , the RTX 3090 is the second most used GPU by LLM enthusiasts. Because RTX 3090 has native INT8 tensors cores…
Anybody else struggling with deepseek after the initial prompt? Somehow it is getting mixed up very easily, even button functionality has been PITA wh…
It is crazy how fast prices are increasing. I'm pulling my hair out to keep ahead of this for students. Servers aren't even an option any more. submit…
2 Months ago I had made a post how I was working on my dual R9700's. It's wild to look back at where we were then and where things now stand. Since th…
Asus Ascent GX10 is now priced at $5999 (1TB), $6999 (2TB), and $7999 (4TB). Buy ASUS Ascent GX10 | Desktop-AI-supercomputer | Networking-IoT-Servers…
I was browsing HF for small LLMs and run into this model. It does not seem to be a fine tune - the model has its own architecture. https://huggingface…
https://preview.redd.it/6e9xb57a42nh1.png?width=787&format=png&auto=webp&s=ffae7996bbf8ab00498cc62c733e7597dc550f24 I'm not sure how many people care…
https://preview.redd.it/via5e88evvmh1.png?width=566&format=png&auto=webp&s=669459ca93ff292f4e1574d098e3e2a0b2c12de4 Gemma 5 or something else? submitt…
A comment really doesn't need to be made, does it? I looked away from the 5090 for a week to other options like the DGX Spark and the M5 Ultra. Both o…
Kaitchup just posted results of his benchmarks for Qwen3.8 27B for quants from different labs, Q4 to Q1, . All the details are hidden behind the paywa…
Feels like maybe we have one more present left, for Christmas. submitted by /u/Miserable-Dare5090 [link] [comments]
submitted by /u/yogthos [link] [comments]