Latest AI/ML News
8160 articles · arXiv cs.AI
arXiv:2610.05461v1 Announce Type: new Abstract: Reinforcement learning from human feedback (RLHF) has played a central role in making large language m…
arXiv:2610.05437v2 Announce Type: new Abstract: Computer-use agents need to capture procedural knowledge of how people use software. User telemetry of…
arXiv:2610.05431v1 Announce Type: new Abstract: Molecular optimization must improve target activity and satisfy developability constraints within limi…
arXiv:2610.05400v1 Announce Type: new Abstract: Long-video agents can actively gather question-relevant evidence, but they typically leave a central d…
arXiv:2610.05398v1 Announce Type: new Abstract: Autonomous research seeks sustained model improvements through iterative experimentation and feedback.…
arXiv:2610.05383v1 Announce Type: new Abstract: Small language models (SLMs) offer a promising foundation for on-device agents through low-latency, re…
arXiv:2610.05375v1 Announce Type: new Abstract: Reconstructing physical fields from training samples that are always incomplete requires learning spat…
arXiv:2610.05370v2 Announce Type: new Abstract: Generative reward models (GRMs) are important for LLM optimization. Unlike scalar reward models, GRMs…
arXiv:2610.05334v1 Announce Type: new Abstract: Frameworks that use large language models for scientific discovery typically rely on a fixed, human-de…
arXiv:2610.05300v1 Announce Type: new Abstract: An agent harness is the code that organizes context, maintains state, and coordinates tool calls for a…
arXiv:2610.05295v1 Announce Type: new Abstract: Indirect prompt injection causes LLM agents to follow commands embedded in external data. A probe may…
arXiv:2610.05284v1 Announce Type: new Abstract: Automated agentic workflow optimization relies on costly evaluations, making it essential to allocate…
arXiv:2610.05281v2 Announce Type: new Abstract: Modern agents increasingly ground their reasoning in observations returned by tools, such as file cont…
arXiv:2610.05256v1 Announce Type: new Abstract: Knowledge distillation (KD) improves low-resource acoustic learning by enriching one-hot supervision w…
arXiv:2610.05225v1 Announce Type: new Abstract: Explainable artificial intelligence (XAI) encompasses methods that draw on different sources of inform…
arXiv:2610.05223v1 Announce Type: new Abstract: The architecture of modern LLMs consists of a profound cognitive polarization. LLMs possess implicit i…
arXiv:2610.05219v1 Announce Type: new Abstract: Most Large Language Models exhibit a fundamental tension between two sequential tasks, such as logical…
arXiv:2610.05211v1 Announce Type: new Abstract: Semantic-driven time-series generation offers a promising way to improve downstream learning in few-sh…
arXiv:2610.05197v1 Announce Type: new Abstract: Data-driven mechanistic hypotheses are essential to scientific discovery because they explain how unde…
arXiv:2610.05176v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on memory to transform executio…