Latest AI/ML News

9894 articles · arXiv cs.LG

arXiv cs.LGOct 7, 2026

arXiv:2610.07825v1 Announce Type: cross Abstract: Time series forecasting models are typically compared on pointwise error, which scores a prediction…

arXiv cs.LGOct 7, 2026

arXiv:2610.07816v1 Announce Type: cross Abstract: Small language models (SLMs) are attractive as local agent controllers because they reduce remote in…

arXiv cs.LGOct 7, 2026

arXiv:2610.07814v1 Announce Type: cross Abstract: How far can stochastic gradient descent ascent (SGDA) go by tuning its timescale ratio and step size…

arXiv cs.LGOct 7, 2026

arXiv:2610.07808v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) offer sparse and event-driven computation, making them attractive for…

arXiv cs.LGOct 7, 2026

arXiv:2610.07782v1 Announce Type: cross Abstract: Decomposing long-context inference across cooperating agents bounds the active KV cache per call rat…

arXiv cs.LGOct 7, 2026

arXiv:2610.07781v1 Announce Type: cross Abstract: Post-training quantization reduces the cost of deploying language-model agents, but its effect on re…

arXiv cs.LGOct 7, 2026

arXiv:2610.07780v1 Announce Type: cross Abstract: Speculative decoding reduces large language model inference latency by drafting multiple tokens befo…

arXiv cs.LGOct 7, 2026

arXiv:2610.07755v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as judges for automated AI evaluation. A common p…

arXiv cs.LGOct 7, 2026

arXiv:2610.07740v1 Announce Type: cross Abstract: We study the online calibration of multidimensional forecasts over an arbitrary convex set $Y\subset…

arXiv cs.LGOct 7, 2026

arXiv:2610.07737v1 Announce Type: cross Abstract: We study fair multi-armed bandits under the Nash Social Welfare (NSW) objective, which measures perf…

arXiv cs.LGOct 7, 2026

arXiv:2610.07731v1 Announce Type: cross Abstract: Dense retrieval models are typically trained with contrastive objectives that learn effective repres…

arXiv cs.LGOct 7, 2026

arXiv:2610.07730v1 Announce Type: cross Abstract: Typed decision models answer a declared question without generating text: a decision head returns a…

arXiv cs.LGOct 7, 2026

arXiv:2610.07723v1 Announce Type: cross Abstract: Safety alignment in Large Language Models (LLMs) remains vulnerable to backdoor attacks. Existing LL…

arXiv cs.LGOct 7, 2026

arXiv:2610.07721v1 Announce Type: cross Abstract: We study ridge regression from exactly $s$ distinct rows of a fixed design. Responses are fixed, and…

arXiv cs.LGOct 7, 2026

arXiv:2610.07720v1 Announce Type: cross Abstract: Multi-reference image generation requires preserving the appearance of multiple subjects while compo…

arXiv cs.LGOct 7, 2026

arXiv:2610.07717v1 Announce Type: cross Abstract: Transformers have exhibited impressive empirical success across various domains, but their theoretic…

arXiv cs.LGOct 7, 2026

arXiv:2610.07712v1 Announce Type: cross Abstract: Existing molecular and materials learning approaches often rely on a limited set of structural repre…

arXiv cs.LGOct 7, 2026

arXiv:2610.07704v1 Announce Type: cross Abstract: Fully decentralized multi-agent reinforcement learning (MARL), also referred to as independent learn…

arXiv cs.LGOct 7, 2026

arXiv:2610.07668v1 Announce Type: cross Abstract: Modern cache replacement designs saturate because they operate within fixed representational structu…

arXiv cs.LGOct 7, 2026

arXiv:2610.07663v1 Announce Type: cross Abstract: User behavior simulation is the computational modeling of user interactions within information syste…

← PreviousPage 10 of 495Next →