Latest AI/ML News

8160 articles · arXiv cs.AI

arXiv cs.AIOct 7, 2026

arXiv:2610.05461v1 Announce Type: new Abstract: Reinforcement learning from human feedback (RLHF) has played a central role in making large language m…

arXiv cs.AIOct 7, 2026

arXiv:2610.05437v2 Announce Type: new Abstract: Computer-use agents need to capture procedural knowledge of how people use software. User telemetry of…

arXiv cs.AIOct 7, 2026

arXiv:2610.05431v1 Announce Type: new Abstract: Molecular optimization must improve target activity and satisfy developability constraints within limi…

arXiv cs.AIOct 7, 2026

arXiv:2610.05400v1 Announce Type: new Abstract: Long-video agents can actively gather question-relevant evidence, but they typically leave a central d…

arXiv cs.AIOct 7, 2026

arXiv:2610.05398v1 Announce Type: new Abstract: Autonomous research seeks sustained model improvements through iterative experimentation and feedback.…

arXiv cs.AIOct 7, 2026

arXiv:2610.05383v1 Announce Type: new Abstract: Small language models (SLMs) offer a promising foundation for on-device agents through low-latency, re…

arXiv cs.AIOct 7, 2026

arXiv:2610.05375v1 Announce Type: new Abstract: Reconstructing physical fields from training samples that are always incomplete requires learning spat…

arXiv cs.AIOct 7, 2026

arXiv:2610.05370v2 Announce Type: new Abstract: Generative reward models (GRMs) are important for LLM optimization. Unlike scalar reward models, GRMs…

arXiv cs.AIOct 7, 2026

arXiv:2610.05334v1 Announce Type: new Abstract: Frameworks that use large language models for scientific discovery typically rely on a fixed, human-de…

arXiv cs.AIOct 7, 2026

arXiv:2610.05300v1 Announce Type: new Abstract: An agent harness is the code that organizes context, maintains state, and coordinates tool calls for a…

arXiv cs.AIOct 7, 2026

arXiv:2610.05295v1 Announce Type: new Abstract: Indirect prompt injection causes LLM agents to follow commands embedded in external data. A probe may…

arXiv cs.AIOct 7, 2026

arXiv:2610.05284v1 Announce Type: new Abstract: Automated agentic workflow optimization relies on costly evaluations, making it essential to allocate…

arXiv cs.AIOct 7, 2026

arXiv:2610.05281v2 Announce Type: new Abstract: Modern agents increasingly ground their reasoning in observations returned by tools, such as file cont…

arXiv cs.AIOct 7, 2026

arXiv:2610.05256v1 Announce Type: new Abstract: Knowledge distillation (KD) improves low-resource acoustic learning by enriching one-hot supervision w…

arXiv cs.AIOct 7, 2026

arXiv:2610.05225v1 Announce Type: new Abstract: Explainable artificial intelligence (XAI) encompasses methods that draw on different sources of inform…

arXiv cs.AIOct 7, 2026

arXiv:2610.05223v1 Announce Type: new Abstract: The architecture of modern LLMs consists of a profound cognitive polarization. LLMs possess implicit i…

arXiv cs.AIOct 7, 2026

arXiv:2610.05219v1 Announce Type: new Abstract: Most Large Language Models exhibit a fundamental tension between two sequential tasks, such as logical…

arXiv cs.AIOct 7, 2026

arXiv:2610.05211v1 Announce Type: new Abstract: Semantic-driven time-series generation offers a promising way to improve downstream learning in few-sh…

arXiv cs.AIOct 7, 2026

arXiv:2610.05197v1 Announce Type: new Abstract: Data-driven mechanistic hypotheses are essential to scientific discovery because they explain how unde…

arXiv cs.AIOct 7, 2026

arXiv:2610.05176v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on memory to transform executio…

← PreviousPage 32 of 408Next →