Latest AI/ML News

8160 articles · arXiv cs.AI

arXiv cs.AIOct 7, 2026

arXiv:2605.23508v2 Announce Type: replace-cross Abstract: Long video generation requires high-fidelity visual synthesis, coherent narrative organizati…

arXiv cs.AIOct 7, 2026

arXiv:2605.23045v4 Announce Type: replace-cross Abstract: Video representation learning has seen tremendous progress in recent years. This has been dr…

arXiv cs.AIOct 7, 2026

arXiv:2605.21318v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are highly sensitive to the prompts used to specify task object…

arXiv cs.AIOct 7, 2026

arXiv:2605.20149v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used for open-ended tasks, but underspecified prompt…

arXiv cs.AIOct 7, 2026

arXiv:2605.18697v2 Announce Type: replace-cross Abstract: Agentic workflows, which compose calls to ML models using a general-purpose programming lang…

arXiv cs.AIOct 7, 2026

arXiv:2605.18648v2 Announce Type: replace-cross Abstract: Central to human-aligned AI is understanding the benefits of human-elicited labels over synt…

arXiv cs.AIOct 7, 2026

arXiv:2605.16668v2 Announce Type: replace-cross Abstract: We introduce GraViti, a transformer-based graph-level variational autoencoder that encodes e…

arXiv cs.AIOct 7, 2026

arXiv:2605.15168v2 Announce Type: replace-cross Abstract: Clinical language models increasingly operate over electronic health records (EHRs), yet pat…

arXiv cs.AIOct 7, 2026

arXiv:2605.14331v2 Announce Type: replace-cross Abstract: Modern edge devices increasingly rely on neural networks for intelligent applications. Howev…

arXiv cs.AIOct 7, 2026

arXiv:2605.12925v4 Announce Type: replace-cross Abstract: Evaluation of software engineering (SWE) agents is dominated by a binary signal: whether the…

arXiv cs.AIOct 7, 2026

arXiv:2605.12729v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used in network operations (NetOps) and…

arXiv cs.AIOct 7, 2026

arXiv:2605.12160v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) policies are typically evaluated under the assumption that the…

arXiv cs.AIOct 7, 2026

arXiv:2605.10659v2 Announce Type: replace-cross Abstract: Digital personas powered by Large Language Models (LLMs) are increasingly proposed as substi…

arXiv cs.AIOct 7, 2026

arXiv:2605.10065v2 Announce Type: replace-cross Abstract: Controlling Large Language Models (LLMs) to prevent the generation of undesirable content, s…

arXiv cs.AIOct 7, 2026

arXiv:2605.09742v2 Announce Type: replace-cross Abstract: Selective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by…

arXiv cs.AIOct 7, 2026

arXiv:2605.07280v2 Announce Type: replace-cross Abstract: One approach to discovering causal relationships in multivariate time series is to ask wheth…

arXiv cs.AIOct 7, 2026

arXiv:2605.06979v2 Announce Type: replace-cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning…

arXiv cs.AIOct 7, 2026

arXiv:2605.06350v2 Announce Type: replace-cross Abstract: LLM cascades, in which a cheap model defers to an expensive one on low-confidence queries, a…

arXiv cs.AIOct 7, 2026

arXiv:2605.05703v3 Announce Type: replace-cross Abstract: Optimizing the communication structure of large language model based multi-agent systems (LL…

arXiv cs.AIOct 7, 2026

arXiv:2605.02035v4 Announce Type: replace-cross Abstract: Ambiguity resolution is a key challenge in multimodal machine translation (MMT), where model…