Latest AI/ML News
389 articles · Reddit r/MachineLearning
Hi everyone, The NeurIPS results aren’t out yet, but I’m expecting my paper to be rejected, so I’m considering submitting it to a journal instead. My…
I'm at a crossroads with two grad school options that would take me in somewhat different research directions, and I wanted some general advice on the…
I'm trying to find out whether back-and-forth between two different LLMs improves task success beyond strong alternatives under a controlled resource…
AAAI-27 Phase 1 results are expected on September 24. Anyone else waiting for the decision? Would be useful to keep this thread for updates when peopl…
What is the generally thought of as the upper limit of the predictive power of XGBoost vs an aggregate of humans? Right now I feed the model the same…
Can we edit our title, abstract and other details till the main paper submission deadline? submitted by /u/CplusplusSupremacy [link] [comments]
I am preparing an ICLR 2027 submission using the official LaTeX style. Some main-text tables use \footnotesize, and two wide tables also use \resizebo…
Hey, I’ve been contributing some of my leftover tokens to this open project called solveathome.org. The idea is pretty straightforward: people can sen…
You can read more here: https://medium.com/@TmlrOrg/asking-authors-about-their-own-papers-3d2e04e5dee0 The results are (imho) concerning. Taken from t…
I have a submission under review in TMLR. Less than a month after submission, I have already received 2 reviews. However, a month has passed since tho…
When models from different families are given the same underspecified task, they often fail in the same way rather than in independent ways. My questi…
I have received the email for the NeurIPS E&D track reference review checker mentioning the 2 hallucinated references. Where do we respond to this ema…
GoBench evaluates LLMs on 9x9 Go games against a ladder of KataGo opponents, from random to superhuman. It measures general reasoning ability, strongl…
GitHub: https://github.com/pfekin/LARA I've been working on LARA (Lightweight Additive Residual Adaptation), a research project on making post-trainin…
Suppose I create two machine learning models suppose tree and neural network for a task let's suppose regression problem, now suppose I am sending bot…
TLDR: I made "poor man’s" DSSM (Deep Structured Semantic Model) — the count-based translation table that can enrich the inverted index for full-text s…
I have created a VAE model using PyTorch on White Wine dataset. Basically, the main goal is to discover a brand-new white wine recipe. It puts all the…
Hi, For single-GPU training, I’m using Hugging Face SFTTrainer with auto_find_batch_size=True, which automatically reduces the batch size after a CUDA…
Hi all, I'm trying to reproduce a paper where the reported dataset statistics in Table 1 don't match what I get from the public raw data, even after i…
Hi, let's suppose I am working on an algorithm that uses principles x to solve problems A and B. I already implemented a very basic algorithm that use…