Latest AI/ML News
389 articles · Reddit r/MachineLearning
LLMs describe physics well but don't "understand" it in any grounded sense — they've learned statistical relationships between tokens like "falls" and…
Hi everyone, I just quickly wanted to share a paper I was working on for around a year now. I created this summary website with key results: https://f…
Has anyone else received an AAAI-27 desk rejection related to modifications to the title or abstract between the abstract-registration deadline and th…
Three weeks from decisions even. I wonder what percentage is industry and VC funded AI labs looking to mingle and recruit. submitted by /u/alrojo [lin…
submitted by /u/Winter_Mistake_3185 [link] [comments]
Benchmark scores: https://preview.redd.it/dgumcg67ggnh1.png?width=1378&format=png&auto=webp&s=fae8fb006ef46fcdebb0876717fc977a905baa89 https://openai.…
From what I've seen online so far, the description of these systems is roughly: They asked the model (often Aster) to generate statements in LEAN and…
Just received an email about the automatic reference/citation checker. Did anyone receive a follow up email about whether the checker was included in…
Abstract Language models spend most of their attention on a small fraction of context, yet they read the entire KV cache to find the few tokens that m…
building a missing data infrastructure and started benchmarking long multi-session conversations (LoCoMo). I know the data looks like: people, facts,…
Hi All, Was reading AIStats' website and it seems like abstract submission is due in 3 weeks. Does anyone know where to find the LaTex template for 20…
I ran a side-by-side ML text-processing and model-training workflow using Fable 5.1 vs. Astra (both on xhigh), and the results could not have been mor…
A researcher has reported a jailbreak of GPT-6 Astra within a day after release. The attack is described as combination of TIP (Task-in-Prompt) attack…
How to get rejected by IEEE T-PAMI with 'Excellent' scores?[D] submitted by /u/cussealin [link] [comments]
AACL-IJCNLP 2026 acceptance results will be released in a few hours. Feel free to share your thoughts and feelings! How did you do? submitted by /u/St…
One thing that has bothered me about LLM benchmarks for a while is that most of them are essentially snapshots. A model is evaluated, a score is publi…
When I first started working in scientific machine learning, I understood the physics much better than the coding. Every time I wanted to try a new ph…
Hello all, I'm a radar signal processing engineer and i trained a 5-class classifier (car, large_vehicle, two_wheeler, pedestrian, pedestrian_group) o…
Is LfD and BC research being effected by recent advances in (so-called) Frontier LLMs? Or is research in LfD and BC sort of going along in an independ…
I used an LLM to iteratively evolve an optimization algorithm rather than solve the packing directly. Starting from a simple seed solver, the LLM prop…