Latest AI/ML News
2784 articles · arXiv cs.CL
arXiv:2601.21262v4 Announce Type: replace Abstract: Although Multimodal Large Language Models (MLLMs) have shown remarkable potential in Visual Docume…
arXiv:2509.13813v3 Announce Type: replace Abstract: Large language models are known to hallucinate, generating linguistically plausible but incorrect…
arXiv:2508.15526v2 Announce Type: replace Abstract: The rapid proliferation of large language models (LLMs) has intensified the requirement for reliab…
arXiv:2410.14948v2 Announce Type: replace Abstract: Medical multimodal large language models (MLLMs) can perform well on existing medical visual quest…
arXiv:2609.26481v1 Announce Type: cross Abstract: Social norms cannot be identified from behavior alone: the same cooperative equilibrium may reflect…
arXiv:2609.26218v1 Announce Type: cross Abstract: Structural graph analysis of the academic publishing network captures the topological relationships…
arXiv:2609.26061v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) models improve capacity with moderate overhead by sparsely activating exper…
arXiv:2609.25948v1 Announce Type: cross Abstract: Target-speaker and multi-speaker extraction are techniques for extracting speech from a desired spea…
arXiv:2609.25007v1 Announce Type: cross Abstract: The performance of state-of-the-art speaker verification (SV) systems severely degrades on short utt…
arXiv:2609.26796v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have recently emerged as a promising alternative to autoregres…
arXiv:2609.26781v1 Announce Type: new Abstract: A multi-agent system can reduce latency on complex tasks by executing work concurrently. Several pione…
arXiv:2609.26687v1 Announce Type: new Abstract: Distinguishing GPT-assisted from independently authored student writing has become a critical challeng…
arXiv:2609.26638v1 Announce Type: new Abstract: Autoregressive OCR vision-language models accurately convert document images into text and structured…
arXiv:2609.26634v1 Announce Type: new Abstract: We introduce Knowledge Pull Requests (KPRs), a framework for continual document authoring that makes e…
arXiv:2609.26629v1 Announce Type: new Abstract: Procedural character generation aims to populate games, simulations, and other virtual worlds with div…
arXiv:2609.26610v1 Announce Type: new Abstract: Despite their outstanding performance on many NLP tasks, LLMs face serious challenges related to seman…
arXiv:2609.26539v1 Announce Type: new Abstract: Large language models (LLMs) have increasingly been used to investigate how children acquire syntax at…
arXiv:2609.26536v1 Announce Type: new Abstract: In LLM-based speech translation, transcription-based chain-of-thought (CoT) suffers from a mismatch be…
arXiv:2609.26489v1 Announce Type: new Abstract: Calibration of language models -- the alignment between expressed or implicit confidence and empirical…
arXiv:2609.26488v1 Announce Type: new Abstract: While Chain-of-Thought (CoT) reasoning has improved the capability of language models, directly applyi…