Latest AI/ML News
770 articles · Reddit r/LocalLLaMA
A 421M-parameter model just played Flappy Bird on my desktop CPU (OpenVINO int8) Running on my Intel Core i7 12th gen CPU Converted laya system one mo…
• Ming-Image-0.1-Design, 6B • Ming-Image-0.1-Design-Layer, 6B • Two open-source Agent Skills: the Ling UI Design Skill and the Image-to-Editable-PPT S…
Dynamic Quantiser TLDR - Makes dynamic and/or custom quants of ggufs. Doesn't need data, just gguf file + llama.cpp. Minimises cosine deviation - high…
I asked Qwen3.8 how this is done, because it works so well honestly it's like having a local claude. I have done a few shot prompts, like a multiplane…
https://github.com/lennartschoch/opencode-cache-compact The default context compacting mechanism in OpenCode strips a bunch of tokens from the beginni…
Someone recommended that I try the ByteShape Qwen 3.8 27B IQ3-XXS GGUF after seeing my previous testing of the GSQ quant. So I did. And the result was…
I asked this back in 2025, but the AI landscape has changed a lot since then. Not looking for the usual ChatGPT, Claude, Gemini, Midjourney, etc. I'm…
I am trying out a few fine-tunes of Qwen 3.5 9B @ IQ4_XS @ 131K context and trying to go mostly local (free beats cheap, after all). However, it is mu…
I want to use it for tasks like researching something, reading and modifying files on my PC, summarising and finding things from PDFs, solving math an…
Box arrived only a couple hours ago, mostly was just transferring data from old mbp, but prepped a raw mlx-vlm patch to run Mimo flash (loading origin…
I've been running FreeToken on my 2x3090 box for a while and ended up maintaining a fork of it. Posting it in case it's useful to anyone else here. Qu…
I was checking nee Mimo 2.6 architecture on huggingface page and it looks very simple. I dont mean in a bad way but when we compare recent open models…
That's the whole question really. Which uncensored / abliterated versions of Qwen Next Flash have you used, and what's your experience? Pros / cons? s…
Github: https://github.com/OpenEuroLLM/ComplexKDA HuggingFace: https://huggingface.co/collections/openeurollm/complexkda Arxiv: https://arxiv.org/abs/…
A note: this entire thing is 100% human written, not even AI drafted or edited. So, enjoy. Or not. edit for a tldr that completed just after posting:…
I didn't know my set up was outperforming nearly everyone until reading another discussion where people were struggling getting half of that speed wit…
https://preview.redd.it/ybwqqbrst0rh1.png?width=811&format=png&auto=webp&s=1e40c5b53e304caf2a10efa2baa6f998b9a9a0bb MiMo-V2.6-Distill-Qwen-9B is a 9B…
submitted by /u/Aggravating-Push-207 [link] [comments]
Basically the title.We did not get a new moe model with qwen 3.8 and Alibaba did not announce any small moe models on apsara.I know we might get an an…
submitted by /u/johnnyApplePRNG [link] [comments]