Latest AI/ML News
8160 articles · arXiv cs.AI
arXiv:2610.04914v1 Announce Type: new Abstract: Monitoring the chain-of-thought (CoT) of large reasoning models (LRMs) is a common way to detect misbe…
arXiv:2610.04911v1 Announce Type: new Abstract: Existing deep research agents are designed primarily for text- and image-based web sources, while vide…
arXiv:2610.04902v1 Announce Type: new Abstract: Recursive language-model agents decompose tasks and delegate subtasks to child instances of the same p…
arXiv:2610.04889v1 Announce Type: new Abstract: Interactive language-model agents increasingly solve complex tasks through long-horizon, multi-call re…
arXiv:2610.04868v1 Announce Type: new Abstract: Memory-augmented agents typically integrate procedural knowledge by injecting retrieved skills directl…
arXiv:2610.04862v1 Announce Type: new Abstract: Long-horizon problem solving and scientific research require computation to accumulate across successi…
arXiv:2610.04838v1 Announce Type: new Abstract: As coding agents take on long-horizon software evolution tasks spanning multiple files and stages, lon…
arXiv:2610.04825v1 Announce Type: new Abstract: Fluent guidance is not the same as useful intervention. LLM tutors are typically trained to generate t…
arXiv:2610.04824v1 Announce Type: new Abstract: AI agents based on foundation models (FMs) have demonstrated strong capabilities to perform complex op…
arXiv:2610.04800v1 Announce Type: new Abstract: Graphical consoles offer a practical interface for medical acquisition assistance, allowing agents to…
arXiv:2610.04794v1 Announce Type: new Abstract: An agent with long-term memory can answer from a record it should no longer use, such as a plan the us…
arXiv:2610.04793v1 Announce Type: new Abstract: Recent incidents show that AI agents sometimes reach measured goals through unsanctioned means. This s…
arXiv:2610.04792v1 Announce Type: new Abstract: Missing modality remains a longstanding challenge in multimodal learning. Existing methods typically a…
arXiv:2610.04791v1 Announce Type: new Abstract: At the core of modern prompting techniques is contextual sensitivity, the ability of large language mo…
arXiv:2610.04756v1 Announce Type: new Abstract: Self-evolving large language model (LLM) systems repeatedly propose, evaluate, and incorporate updates…
arXiv:2610.04751v1 Announce Type: new Abstract: Beyond scaling their parameters and data, large language models currently gain versatility on new prob…
arXiv:2610.04749v1 Announce Type: new Abstract: Routine complete blood counts (CBCs) could yield new biomarkers, but the private records needed to eva…
arXiv:2610.04747v1 Announce Type: new Abstract: In life-or-death situations, a benevolent lie may appear more moral than telling the truth. Yet such l…
arXiv:2610.04740v1 Announce Type: new Abstract: Early-stage computational drug discovery requires coordinating heterogeneous scientific tools across m…
arXiv:2610.04737v1 Announce Type: new Abstract: Reasoning models can remain capable of solving a task while still defaulting to cheaper but misleading…