Latest AI/ML News
383 articles · Reddit r/MachineLearning
I tested nine vision models on the same 2,000 spider photos. The highest exact-species accuracy was 49.85%. The tasks, predictions, scoring code, and…
Hi! Neurips Sydney is sold out so I joined the waitlist. Does anyone have any concrete understanding or insights of how the tickets get released and t…
Hi everyone, Our arXiv paper was originally scheduled for announcement on Sep 15, but it was put on hold for moderation before publication. The hold w…
OpenTrainDNN is an open-source, client-side web application designed to render the step-by-step training mechanics of deep neural networks in real-tim…
I’m working on a student machine learning/computer vision project and recently realized that my validation set was not completely independent from my…
Hi, I'm a freshly graduated (undergrad) guy currently working on 3–4 projects in parallel: - An AI research project: working on a research problem sta…
Can you attach the supplementary material along with my paper when uploading it? Or would the paper be desk-rejected if I do so? I don’t want to split…
Just some hours or a day for NeurIPS decisions. How you guys are feeling ? Anybody got any decisions ? Would be great to know ballpark submission # an…
I’m 28 and trying to make a pretty major career decision. I currently have the option of finishing an MD. I have about 2 years left, but I genuinely d…
With NeurIPS author notifications coming up, I’m realizing that I’m way more stressed than I expected to be. I know the usual advice: reviews are nois…
The rebuttal is still ongoing so this could also change, but my reviewers have not responded back to my responses yet. I know they have the right not…
$17.2 million investment, spanning three to five years from Simons Foundation International , XTX Markets , and Siegel Family Endowment #Philanthropy)…
play here Play games like poker, risk, diplomacy with friends or alone against AI models, guess what, you can talk to them and change strategies and o…
LinearSolverBench measures the ability of a model or harness to write fast, accurate, and general numerical solvers for large sparse linear systems in…
Our most recent work at Templar explores fault tolerance in Crucible, our distributed pre-training platform. The goal is to keep healthy workers train…
Retrieval benchmarks sometimes feel benchmaxxed by models, so we wanted to find a way to tie it as close as possible to my objective: finding the arti…
TLDR: We demonstrate and explain the difference in expressivity of Gated Deltanet (GDN) and Kimi Delta Attention (KDA). We show how the full diagonal…
Hi, I have a solo paper on arxiv from my masters studies, which was my part-time work and a few months back I got back to it and tried to finally make…
The total training cost was just $3.5M. The model comes with a live benchmaxxing dashboard. https://preview.redd.it/89uurv5r21rh1.png?width=1518&forma…
Source: Jev Benchmarks Its training method is literally called "Reinforcement Learning for Calibrated Decisions." Calibration gap vs human labels (low…