Latest AI/ML News

8160 articles · arXiv cs.AI

arXiv cs.AIOct 7, 2026

arXiv:2610.03966v1 Announce Type: new Abstract: Each run of an AI-driven research system (ADRS) is an expensive search over a vast solution space, and…

arXiv cs.AIOct 7, 2026

arXiv:2610.03959v1 Announce Type: new Abstract: Recent world action models (WAMs) reuse pretrained video VAEs whose encoder latents directly condition…

arXiv cs.AIOct 7, 2026

arXiv:2610.03948v1 Announce Type: new Abstract: Autonomous driving decision systems must balance safety, efficiency, and social norms in complex traff…

arXiv cs.AIOct 7, 2026

arXiv:2610.03938v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) have recently pushed the frontier of visual reasoning…

arXiv cs.AIOct 7, 2026

arXiv:2610.03934v1 Announce Type: new Abstract: Existing analog integrated circuit design benchmarks make two questions hard to answer: whether a mode…

arXiv cs.AIOct 7, 2026

arXiv:2610.03894v1 Announce Type: new Abstract: A deployed LLM agent emits tool calls, queries, and code that can be silently wrong -- by the time the…

arXiv cs.AIOct 7, 2026

arXiv:2610.03888v1 Announce Type: new Abstract: Reconstructing global sea surface pH from sparse observations is critical for monitoring ocean acidifi…

arXiv cs.AIOct 7, 2026

arXiv:2610.03872v1 Announce Type: new Abstract: AI agents are becoming increasingly capable of generating scientific code, but generating code is not…

arXiv cs.AIOct 2, 2026

arXiv:2605.20072v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly proposed as cognitive components for robotic systems…

arXiv cs.AIOct 2, 2026

arXiv:2605.13737v2 Announce Type: replace Abstract: When an omnimodal large language model accepts a question whose textual premise contradicts what i…

arXiv cs.AIOct 2, 2026

arXiv:2605.08386v2 Announce Type: replace Abstract: Skill libraries have become a practical way for LLM agents to reuse procedural experience across t…

arXiv cs.AIOct 2, 2026

arXiv:2604.21549v2 Announce Type: replace Abstract: Estimating the prevalence of a category in a population using imperfect measurement devices (diagn…

arXiv cs.AIOct 2, 2026

arXiv:2604.20779v2 Announce Type: replace Abstract: AI coding agents are being adopted at scale, yet we lack empirical evidence on how people actually…

arXiv cs.AIOct 2, 2026

arXiv:2604.06820v3 Announce Type: replace Abstract: Evaluating misinformation requires distinguishing whether readers believe content from whether the…

arXiv cs.AIOct 2, 2026

arXiv:2604.01997v2 Announce Type: replace Abstract: Gait analysis provides an objective characterization of locomotor function and is widely used to s…

arXiv cs.AIOct 2, 2026

arXiv:2604.01375v3 Announce Type: replace Abstract: Rubrics distill notions of expert quality and measure agent performance. However, the quality of r…

arXiv cs.AIOct 2, 2026

arXiv:2603.19042v5 Announce Type: replace Abstract: The integration of artificial intelligence (AI) into judicial decision making -- particularly in p…

arXiv cs.AIOct 2, 2026

arXiv:2602.13691v2 Announce Type: replace Abstract: Recent advancements in Large Language Model (LLM) agents have demonstrated strong capabilities in…

arXiv cs.AIOct 2, 2026

arXiv:2602.08889v2 Announce Type: replace Abstract: Quantitative risk assessment relies on structured expert elicitation to estimate unobservable prop…

arXiv cs.AIOct 2, 2026

arXiv:2602.03006v3 Announce Type: replace Abstract: Deploying Large Language Models (LLMs) for discriminative workloads is often limited by inference…

← PreviousPage 39 of 408Next →