Latest AI/ML News
9894 articles · arXiv cs.LG
arXiv:2605.19462v2 Announce Type: replace Abstract: Self-supervised learning (SSL) assumes that solving pretext tasks on unlabeled data yields represe…
arXiv:2605.16339v2 Announce Type: replace Abstract: Preference learning in large language models relies on reward models as proxies for human judgment…
arXiv:2605.15692v2 Announce Type: replace Abstract: We study episodic reinforcement learning with fixed reward and transition functions, but with epis…
arXiv:2605.14698v2 Announce Type: replace Abstract: Foundation models (FMs) promise to extract unified representations that generalize across downstre…
arXiv:2605.14220v2 Announce Type: replace Abstract: Modern LLM RL systems separate rollout generation from policy optimization. These two stages are e…
arXiv:2605.13684v2 Announce Type: replace Abstract: We study the optimal scale at which real-valued function classes exhibit uniform convergence and l…
arXiv:2605.12904v2 Announce Type: replace Abstract: Tabular foundation models (TFMs) have emerged as a powerful paradigm for in-context learning on st…
arXiv:2605.07938v2 Announce Type: replace Abstract: Single-cell representation learning (SCRL) from gene expression data offers a way to uncover the c…
arXiv:2605.06814v2 Announce Type: replace Abstract: Graph neural networks (GNNs) increasingly rely on sophisticated architectures and training procedu…
arXiv:2605.06462v2 Announce Type: replace Abstract: Progress in graph learning is hindered by benchmark practices that conflate the contributions of n…
arXiv:2605.06272v2 Announce Type: replace Abstract: While generative modeling has achieved remarkable success on tasks like natural language-condition…
arXiv:2604.19146v2 Announce Type: replace Abstract: Particle accelerator beamline optimization is a high-dimensional control problem traditionally req…
arXiv:2604.16778v3 Announce Type: replace Abstract: Modern agents specialize in varying domains while there is no clear approach combining different d…
arXiv:2604.08454v2 Announce Type: replace Abstract: Large language models are increasingly deployed in high-stakes domains, where confident yet incorr…
arXiv:2603.26164v2 Announce Type: replace Abstract: Data-centric training has emerged as a promising direction for improving large language models (LL…
arXiv:2603.14474v2 Announce Type: replace Abstract: Sketch techniques have been extensively studied in recent years and are especially well-suited to…
arXiv:2603.08459v2 Announce Type: replace Abstract: Safe predictions are a crucial requirement for integrating predictive models into clinical decisio…
arXiv:2603.00326v2 Announce Type: replace Abstract: Sparse oblique (SPO), part of the top-ranked configuration of Google's Yggdrasil Decision Forests…
arXiv:2602.14049v2 Announce Type: replace Abstract: Spatio-temporal traffic forecasting is a core component of intelligent transportation systems, sup…
arXiv:2602.10727v3 Announce Type: replace Abstract: Rising Multi-Armed Bandits (RMABs) model sequential decision problems where each arm's expected re…