Latest AI/ML News

8160 articles · arXiv cs.AI

arXiv cs.AIOct 7, 2026

arXiv:2610.04319v1 Announce Type: cross Abstract: Transformer models such as BERT and Vision Transformer~(ViT) achieve strong performance via densely…

arXiv cs.AIOct 7, 2026

arXiv:2610.04303v1 Announce Type: cross Abstract: Recursive computation repeatedly compresses or reuses intermediate states, creating a simple tension…

arXiv cs.AIOct 7, 2026

arXiv:2610.04299v1 Announce Type: cross Abstract: Self-evolving reasoning models learn from their own generated questions, yet repeated self-training…

arXiv cs.AIOct 7, 2026

arXiv:2610.04296v1 Announce Type: cross Abstract: Learning-based models for radio resource management (RRM) are typically built for a single function…

arXiv cs.AIOct 7, 2026

arXiv:2610.04289v1 Announce Type: cross Abstract: Wireless foundation models learn representations from unlabeled radio signals for reuse across downs…

arXiv cs.AIOct 7, 2026

arXiv:2610.04286v1 Announce Type: cross Abstract: Distributed Energy Resource (DER) environments rely on network communication protocols to coordinate…

arXiv cs.AIOct 7, 2026

arXiv:2610.04283v1 Announce Type: cross Abstract: Activation steering exploits interpretable directions in the residual stream to enable inference-tim…

arXiv cs.AIOct 7, 2026

arXiv:2610.04272v1 Announce Type: cross Abstract: Combining capabilities of multiple expert models trained starting from the same base checkpoint has…

arXiv cs.AIOct 7, 2026

arXiv:2610.04267v1 Announce Type: cross Abstract: Generating multiple-choice questions is increasingly scalable, but establishing their assessment qua…

arXiv cs.AIOct 7, 2026

arXiv:2610.04261v1 Announce Type: cross Abstract: Reinforcement fine-tuning (RFT) is increasingly used in applications where large language models (LL…

arXiv cs.AIOct 7, 2026

arXiv:2610.04246v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed to simulate human behavior, acting as computa…

arXiv cs.AIOct 7, 2026

arXiv:2610.04240v1 Announce Type: cross Abstract: For ground robots operating in outdoor environments, understanding the properties of the underlying…

arXiv cs.AIOct 7, 2026

arXiv:2610.04219v1 Announce Type: cross Abstract: The electric autonomous dial-a-ride problem (EADARP) extends the classical dial-a-ride problem by in…

arXiv cs.AIOct 7, 2026

arXiv:2610.04212v1 Announce Type: cross Abstract: Graph Neural Operators (GNOs) provide flexible surrogate models for learning solution operators of p…

arXiv cs.AIOct 7, 2026

arXiv:2610.04211v1 Announce Type: cross Abstract: Skeletal motion is stored as every joint's transform at every frame, yet most of it is implied by th…

arXiv cs.AIOct 7, 2026

arXiv:2610.04204v1 Announce Type: cross Abstract: Training large language models (LLMs) entails a fundamental trade-off: memory-efficient optimizers s…

arXiv cs.AIOct 7, 2026

arXiv:2610.04171v1 Announce Type: cross Abstract: Inference-time steering combines pretrained diffusion experts or rewards without retraining by chang…

arXiv cs.AIOct 7, 2026

arXiv:2610.04169v1 Announce Type: cross Abstract: LLM fingerprinting via watermark distillation embeds a statistical watermark signal into model weigh…

arXiv cs.AIOct 7, 2026

arXiv:2610.04158v1 Announce Type: cross Abstract: Recent studies on reinforcement learning (RL) report seemingly conflicting evidence about large lang…

arXiv cs.AIOct 7, 2026

arXiv:2610.04156v1 Announce Type: cross Abstract: LLM agents for clinical text-to-SQL applications reason autonomously over multiple steps but cannot…

← PreviousPage 25 of 408Next →