Latest AI/ML News
2784 articles · arXiv cs.CL
arXiv:2609.20800v1 Announce Type: new Abstract: World modeling enables intelligence to anticipate consequences, guide interventions, and learn from in…
arXiv:2609.20734v1 Announce Type: new Abstract: Reasoning and agentic workloads increasingly demand efficient long-context inference. Yet full-attenti…
arXiv:2609.20712v1 Announce Type: new Abstract: This paper introduces and operationalizes summarization bias: a proposed systematic tendency of large…
arXiv:2609.20684v1 Announce Type: new Abstract: Large language models are increasingly used in healthcare communication, yet most evaluations emphasiz…
arXiv:2609.20630v1 Announce Type: new Abstract: Search advertising connects user intent with commercial content and plays a critical role in platform…
arXiv:2609.20612v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) lets a language model learn from a frozen copy of itself that sees…
arXiv:2609.20593v1 Announce Type: new Abstract: Word-in-Context (WiC) remains challenging for language models, despite recent progress on lexical-sema…
arXiv:2609.20584v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly considered for safety-critical engineering, yet their re…
arXiv:2609.20565v1 Announce Type: new Abstract: Recent advancements in large language models have revolutionized the field of psychological counseling…
arXiv:2609.20541v1 Announce Type: new Abstract: Large language models can report a numerical confidence together with generated content, but it is unc…
arXiv:2609.20398v1 Announce Type: new Abstract: Semantic parsing (SP)-based knowledge base question answering aims to answer natural language question…
arXiv:2609.20303v1 Announce Type: new Abstract: Classical philosophical corpora pose three compounding challenges for language resources: they exist i…
arXiv:2609.20232v1 Announce Type: new Abstract: We introduce the \textbf{Public Discourse Corpus (PDC)}, the first dataset of public-figure interview…
arXiv:2609.20223v1 Announce Type: new Abstract: We study weakly supervised incremental telecom fraud detection from raw telephone audio, where trainin…
arXiv:2609.20207v1 Announce Type: new Abstract: Large language models produce prompt-dependent probabilities over words, whereas scientific systems re…
arXiv:2609.20186v1 Announce Type: new Abstract: Speculative Decoding (SD) has significantly accelerated Large Language Model (LLM) inference, yet exis…
arXiv:2609.20169v1 Announce Type: new Abstract: Next-Generation Sequencing has revolutionized the study of genetic mutations, enabling large-scale inv…
arXiv:2609.20104v1 Announce Type: new Abstract: We describe the architecture, training methodology and inference speedups of Granite 5.0 Turbo CTC, a…
arXiv:2609.19989v1 Announce Type: new Abstract: The widespread adoption of LLMs has led to escalating content compliance risks. Prior works have contr…
arXiv:2609.19969v1 Announce Type: new Abstract: The widespread adoption of long-horizon agents has made model workloads increasingly input-heavy. Alth…