Latest AI/ML News
9894 articles · arXiv cs.LG
arXiv:2610.07973v1 Announce Type: new Abstract: We study the problem of recovering the ranking of a fixed set of items according to their unknown nume…
arXiv:2610.07967v1 Announce Type: new Abstract: As large language model (LLM) agents become increasingly autonomous, they may pursue task performance…
arXiv:2610.07938v1 Announce Type: new Abstract: Implicit-process priors specify distributions over functions through sample-forward mechanisms such as…
arXiv:2610.07910v1 Announce Type: new Abstract: Deep Reinforcement Learning policies can produce nonsmooth action oscillations that hinder deployment…
arXiv:2610.07904v1 Announce Type: new Abstract: We introduce ApexQuant, a calibration-free quantization method that recursively re-quantizes the resid…
arXiv:2610.07899v1 Announce Type: new Abstract: Generative actors are transforming offline reinforcement learning (RL) by enabling expressive policy c…
arXiv:2610.07898v1 Announce Type: new Abstract: Repository-level software engineering (SWE) is a challenging long-horizon setting: agents must reason…
arXiv:2610.07885v1 Announce Type: new Abstract: Electrocardiogram (ECG) delineation, the identification of waveform boundaries, is a foundational step…
arXiv:2610.07874v1 Announce Type: new Abstract: On-policy distillation (OPD) has been widely studied as a post-training method in which a student mode…
arXiv:2610.07859v1 Announce Type: new Abstract: Conventional decentralized federated learning (DFL) often focuses on clients, with each client maintai…
arXiv:2610.07857v1 Announce Type: new Abstract: This study formulates individual route reproduction as a shortest-path problem over learned driver-spe…
arXiv:2610.07853v1 Announce Type: new Abstract: Ternary language models such as BitNet b1.58, Falcon-E and BitCPM are fine-tuned with higher-precision…
arXiv:2610.07842v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a language model to match a copy of itself conditioned on pr…
arXiv:2610.07834v1 Announce Type: new Abstract: Retrieval-augmented time-series forecasting uses the continuations of historical segments similar to t…
arXiv:2610.07824v1 Announce Type: new Abstract: Electrophysiological source imaging (ESI) aims to estimate cortical source activity from noninvasive e…
arXiv:2610.07823v1 Announce Type: new Abstract: The AI CUP 2025 Precise Analysis of Table Tennis Smart Racket Data Competition introduced smart table…
arXiv:2610.07819v1 Announce Type: new Abstract: Model merging offers a promising solution for combining multiple fine-tuned checkpoints into a single…
arXiv:2610.07810v1 Announce Type: new Abstract: Search filters help guests navigate vast catalogs in two-sided marketplaces like Airbnb, and recommend…
arXiv:2610.07809v1 Announce Type: new Abstract: Sparsely activated Mixture-of-Experts (MoE) models increase model capacity without a proportional incr…
arXiv:2610.07804v1 Announce Type: new Abstract: Prior Fitted Networks (PFNs) such as TabPFN now rival established statistical procedures across predic…