Latest AI/ML News

2784 articles · arXiv cs.CL

arXiv cs.CLSep 24, 2026

arXiv:2510.01354v2 Announce Type: replace-cross Abstract: Multiple prompt injection attacks have been proposed against web agents. At the same time, v…

arXiv cs.CLSep 24, 2026

arXiv:2312.17295v2 Announce Type: replace-cross Abstract: With the rise of large language models (LLMs) and concerns about potential misuse, watermark…

arXiv cs.CLSep 24, 2026

arXiv:2608.26706v2 Announce Type: replace Abstract: Expert-level financial question answering requires both grounded verification to catch numeric hal…

arXiv cs.CLSep 24, 2026

arXiv:2608.08090v2 Announce Type: replace Abstract: Although multilingual approaches to figurative language identification are not new, the shift beyo…

arXiv cs.CLSep 24, 2026

arXiv:2607.00848v4 Announce Type: replace Abstract: In this opinion paper, we propose MetaHOPE, an error severity-aware annotation framework for evalu…

arXiv cs.CLSep 24, 2026

arXiv:2606.11420v2 Announce Type: replace Abstract: Spoken factual claims often occur within multi-turn conversations, where surrounding dialogue can…

arXiv cs.CLSep 24, 2026

arXiv:2606.10338v2 Announce Type: replace Abstract: Machine unlearning is increasingly important for large language models, yet unlearning in Mixture-…

arXiv cs.CLSep 24, 2026

arXiv:2606.01322v2 Announce Type: replace Abstract: Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Res…

arXiv cs.CLSep 24, 2026

arXiv:2605.29313v2 Announce Type: replace Abstract: LLM multi-agent systems often coordinate through natural-language dialogue or loosely structured s…

arXiv cs.CLSep 24, 2026

arXiv:2605.28211v2 Announce Type: replace Abstract: SpeechLLMs are increasingly deployed in professional settings where domain customisation is standa…

arXiv cs.CLSep 24, 2026

arXiv:2604.26139v3 Announce Type: replace Abstract: Diffusion large language models generate text through iterative denoising, exposing hidden traject…

arXiv cs.CLSep 24, 2026

arXiv:2604.13061v3 Announce Type: replace Abstract: Large language models increasingly run in stateful pipelines that assemble each prompt from retrie…

arXiv cs.CLSep 24, 2026

arXiv:2603.18482v4 Announce Type: replace Abstract: Why does machine-generated text remain detectable? We investigate a mechanistic explanation at the…

arXiv cs.CLSep 24, 2026

arXiv:2603.17373v2 Announce Type: replace Abstract: Large language models are rapidly being deployed as AI tutors, yet current evaluation paradigms as…

arXiv cs.CLSep 24, 2026

arXiv:2603.09222v2 Announce Type: replace Abstract: Efficient context compression is critical for retrieval-augmented question answering in resource-c…

arXiv cs.CLSep 24, 2026

arXiv:2603.08166v2 Announce Type: replace Abstract: Automated Drug Combination Extraction (DCE) from large-scale biomedical literature is crucial for…

arXiv cs.CLSep 24, 2026

arXiv:2603.05308v4 Announce Type: replace Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and cl…

arXiv cs.CLSep 24, 2026

arXiv:2512.04457v3 Announce Type: replace Abstract: Machine unlearning for large language models (LLMs) remains challenging because full retraining is…

arXiv cs.CLSep 24, 2026

arXiv:2508.13680v5 Announce Type: replace Abstract: We introduce VMMU, a Vietnamese Multitask Multimodal Understanding and Reasoning Benchmark designe…

arXiv cs.CLSep 24, 2026

arXiv:2507.21112v4 Announce Type: replace Abstract: With the rapid rise of InsurTech, traditional insurance companies are increasingly exploring alter…

← PreviousPage 11 of 140Next →