Latest AI/ML News
9894 articles · arXiv cs.LG
arXiv:2610.07283v1 Announce Type: new Abstract: Analog hardware platforms offer the potential to reduce energy consumption over digital architectures,…
arXiv:2610.07271v1 Announce Type: new Abstract: Linkage algorithms for hierarchical clustering (HC) are a powerful and efficient framework for constru…
arXiv:2610.07255v1 Announce Type: new Abstract: Neural algorithmic reasoning, or aligning a neural network with an algorithmic paradigm, has emerged a…
arXiv:2610.07253v1 Announce Type: new Abstract: Neural fields are usually evaluated by how well they reconstruct an observation. We show that this mis…
arXiv:2610.07247v1 Announce Type: new Abstract: Large language models have shown strong reasoning capabilities, but their high inference costs make kn…
arXiv:2610.07232v1 Announce Type: new Abstract: Accurate short-term load forecasting (STLF) is essential for the reliable and efficient operation of m…
arXiv:2610.07229v1 Announce Type: new Abstract: Motivated by sequence-to-sequence transport in the context time-series domain adaptation, we study the…
arXiv:2610.07226v1 Announce Type: new Abstract: ``What are the irreducible conditions that are sufficient to produce an outcome?'' is one of the most…
arXiv:2610.07220v1 Announce Type: new Abstract: We present three practical tutorials on numerical computation and machine learning for mathematical re…
arXiv:2610.07218v1 Announce Type: new Abstract: Recent advances in representation learning have highlighted the utility of constant-curvature models,…
arXiv:2610.07212v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) trains a language model on problems that may the…
arXiv:2610.07208v1 Announce Type: new Abstract: Predicting migration flows remains a significant challenge for traditional gravity-based forecasting m…
arXiv:2610.07207v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) transformers scale capacity by activating only a few experts per token, but t…
arXiv:2610.07197v1 Announce Type: new Abstract: Exact unlearning requires a deployed predictor to match one rebuilt without the information named by a…
arXiv:2610.07184v1 Announce Type: new Abstract: A key challenge in building AI systems for scientific research is enabling $\textit{scientific explora…
arXiv:2610.07177v1 Announce Type: new Abstract: An open contrastive decision model is near chance as a judge on the hard public benchmarks: Contrastiv…
arXiv:2610.07168v1 Announce Type: new Abstract: Representations of translated sentences are similar in the inner layers of multilingual language model…
arXiv:2610.07162v1 Announce Type: new Abstract: Deep hedging learns trading policies from historical or simulated market trajectories, yet under nonst…
arXiv:2610.07131v1 Announce Type: new Abstract: We study the implicit bias of Riemannian gradient flow for hyperbolic multiclass classification with f…
arXiv:2610.07121v1 Announce Type: new Abstract: Diffusion large language models dLLMs) have emerged as a promising alternative to autoregressive langu…