Latest AI/ML News

268 articles · arXiv cs.AI

arXiv cs.AIAug 17, 2026

arXiv:2608.11625v2 Announce Type: replace Abstract: Feedback processes strongly influence student learning, yet their educational value depends on add…

arXiv cs.AIAug 17, 2026

arXiv:2608.11195v3 Announce Type: replace Abstract: AI agents are increasingly used in mathematics research, but it is often unclear how to use them e…

arXiv cs.AIAug 17, 2026

arXiv:2608.10538v2 Announce Type: replace Abstract: Agent skills represent a standardized format for packaging procedural knowledge and domain experti…

arXiv cs.AIAug 17, 2026

arXiv:2608.10492v2 Announce Type: replace Abstract: Large Language Model (LLM)-based simulators often reproduce observable actions but fail to capture…

arXiv cs.AIAug 17, 2026

arXiv:2608.08802v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) makes Multimodal Large Language Models more…

arXiv cs.AIAug 17, 2026

arXiv:2608.05246v2 Announce Type: replace Abstract: Existing personalized LLM benchmarks primarily rely on textual personas or isolated behavioral sig…

arXiv cs.AIAug 17, 2026

arXiv:2608.03682v3 Announce Type: replace Abstract: Physical AI policies require inference throughout their lifecycle, including model evaluation, clo…

arXiv cs.AIAug 17, 2026

arXiv:2608.02606v2 Announce Type: replace Abstract: Fault tolerance in classical computing has traditionally relied on static strategies like hardware…

arXiv cs.AIAug 17, 2026

arXiv:2608.01856v2 Announce Type: replace Abstract: Bi-temporal remote-sensing disaster change captioning often needs to identify sparse and spatially…

arXiv cs.AIAug 17, 2026

arXiv:2607.28336v3 Announce Type: replace Abstract: On-policy distillation provides dense supervision for multimodal reasoners, but its trajectory-lev…

arXiv cs.AIAug 17, 2026

arXiv:2607.18785v3 Announce Type: replace Abstract: As large language model agents gain access to increasingly large skill libraries, retrieving the r…

arXiv cs.AIAug 17, 2026

arXiv:2607.14975v2 Announce Type: replace Abstract: Channel foundation models (CFMs) are commonly evaluated in model-specific pipelines that differ in…

arXiv cs.AIAug 17, 2026

arXiv:2607.14616v4 Announce Type: replace Abstract: Vision-language models (VLMs) can describe a scene, but can they act well within one? We study whe…

arXiv cs.AIAug 17, 2026

arXiv:2606.08296v2 Announce Type: replace Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfu…

arXiv cs.AIAug 17, 2026

arXiv:2605.29668v2 Announce Type: replace Abstract: LLM agents acting in structured environments fail in operational rather than conversational ways,…

arXiv cs.AIAug 17, 2026

arXiv:2605.28642v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated significant potential for speech-to-tex…

arXiv cs.AIAug 17, 2026

arXiv:2605.01189v3 Announce Type: replace Abstract: Clinical AI adoption is hindered by the black-box/grey-box nature of high-performing models, which…

arXiv cs.AIAug 17, 2026

arXiv:2604.08525v2 Announce Type: replace Abstract: Large language models (LLMs) are trained to align with user preferences through methods like reinf…

arXiv cs.AIAug 17, 2026

arXiv:2603.28026v2 Announce Type: replace Abstract: Multimodal multiple-choice question answering (MCQA) provides a standardized and objectively measu…

arXiv cs.AIAug 17, 2026

arXiv:2603.18871v2 Announce Type: replace Abstract: Urban Vehicular Ad-Hoc Networks (VANETs) can become fragmented because buildings obstruct wireless…