Latest AI/ML News
8160 articles · arXiv cs.AI
arXiv:2610.04339v1 Announce Type: new Abstract: ShadowMiner v1 is a system that automatically discovers research problems and generates hypotheses fro…
arXiv:2610.04329v1 Announce Type: new Abstract: Reliable language models should resist unsupported user pressure while effectively using objective con…
arXiv:2610.04328v1 Announce Type: new Abstract: Correction-based offline preference pipelines commonly treat model failures only as rejected responses…
arXiv:2610.04316v1 Announce Type: new Abstract: Output-only safety monitoring sees only the end of a model's computation, yet the model computes its a…
arXiv:2610.04313v1 Announce Type: new Abstract: Long-context decoding is limited by memory bandwidth, because every output token reads the KV cache of…
arXiv:2610.04301v1 Announce Type: new Abstract: Large datasets and high capacity models have accelerated progress in vision and language. This work in…
arXiv:2610.04295v1 Announce Type: new Abstract: Large language models incur language-dependent representation and inference costs, but existing compar…
arXiv:2610.04292v1 Announce Type: new Abstract: LLM-based agents are increasingly capable of generating complex 3D structures, with the potential to r…
arXiv:2610.04287v1 Announce Type: new Abstract: Enterprise agents should improve from delayed feedback without allowing every correction to rewrite sy…
arXiv:2610.04280v1 Announce Type: new Abstract: Geometry reasoning is naturally stateful: solving a problem repeatedly alternates between structural p…
arXiv:2610.04262v1 Announce Type: new Abstract: Reconstructing editable parametric CAD models from a single-view image is of great practical value for…
arXiv:2610.04253v1 Announce Type: new Abstract: Generating an executable program does not necessarily mean that it correctly implements the behavioral…
arXiv:2610.04245v1 Announce Type: new Abstract: Existing activation steering methods often assume that a high-level concept can be mediated by a singl…
arXiv:2610.04215v1 Announce Type: new Abstract: Large language models (LLMs) can generate clinical narratives that are insufficiently grounded in pati…
arXiv:2610.04206v2 Announce Type: new Abstract: Spatial intelligence is a fundamental skill in multiple domains, such as Science, Technology, Engineer…
arXiv:2610.04198v1 Announce Type: new Abstract: Diffusion language models (DLMs) enable fast generation by predicting multiple tokens in parallel, but…
arXiv:2610.04196v1 Announce Type: new Abstract: LLM-based long-horizon agentic post-training is often bottlenecked by rollout generation: trajectories…
arXiv:2610.04195v1 Announce Type: new Abstract: Personal AI agents in enterprise multi-tenant deployments share a common vector store for long-term me…
arXiv:2610.04188v2 Announce Type: new Abstract: Recent studies show that artificial intelligence (AI) with language and vision capabilities still expe…
arXiv:2610.04184v1 Announce Type: new Abstract: Recursive self-improvement (RSI) relies on evaluation feedback to assess progress and guide further re…