Latest AI/ML News
2784 articles · arXiv cs.CL
arXiv:2610.08037v1 Announce Type: new Abstract: Language models frequently generate outputs in unintended languages or scripts, a phenomenon known as…
arXiv:2610.08026v1 Announce Type: new Abstract: In recent years, several methods for detecting when large language models (LLMs) hallucinate have been…
arXiv:2610.08018v1 Announce Type: new Abstract: Reliable tool use requires more than triggering a mechanism or matching a query to an API description.…
arXiv:2610.07940v1 Announce Type: new Abstract: Looped language models apply the same stack of layers T times to each token, which deepens the model w…
arXiv:2610.07937v1 Announce Type: new Abstract: Redpine Science gives models and agents a single access point to a wide range of peer-reviewed literat…
arXiv:2610.07936v1 Announce Type: new Abstract: Systematicity, the probabilistic mapping of form to meaning, permeates language at all levels, and sub…
arXiv:2610.07902v1 Announce Type: new Abstract: Cantonese lyric writing requires close alignment between lexical tones and melodic pitch. Existing mel…
arXiv:2610.07887v1 Announce Type: new Abstract: Unified multimodal models (UMMs) integrate understanding and generation, yet their generative behavior…
arXiv:2610.07848v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) has become a standard approach for adapting large language mode…
arXiv:2610.07847v1 Announce Type: new Abstract: As LLMs increasingly assist in moral reasoning, omission bias, the tendency to prefer inaction even wh…
arXiv:2610.07822v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by using a lightweight draft model to propo…
arXiv:2610.07817v1 Announce Type: new Abstract: Organizations automating operational processes need more than a correct outcome: they need to predict…
arXiv:2610.07774v1 Announce Type: new Abstract: Vision-language models (VLMs) face compositional safety risks where harmful intent emerges from the in…
arXiv:2610.07764v1 Announce Type: new Abstract: Natural language processing (NLP) models can detect depression-related language in text written near t…
arXiv:2610.07753v1 Announce Type: new Abstract: Tool-using agents make consequential changes to external state, yet correct outcomes do not guarantee…
arXiv:2610.07722v1 Announce Type: new Abstract: Activation steering provides a lightweight and flexible way to control large language model (LLM) beha…
arXiv:2610.07716v1 Announce Type: new Abstract: Prefill-only decision models inspired by the Jev model score every candidate in a menu during a single…
arXiv:2610.07700v1 Announce Type: new Abstract: We study the robustness of keystroke dynamics for detecting large language model (LLM)-assisted writin…
arXiv:2610.07659v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation in large language models. In each drafting…
arXiv:2610.07643v1 Announce Type: new Abstract: Most KV-cache eviction methods ask, in effect, which memory appeared important while reading the promp…