Latest AI/ML News
2784 articles · arXiv cs.CL
arXiv:2606.05087v2 Announce Type: replace Abstract: Frequent verbs such as 'have' and 'make' can function either as collocates in light-verb construct…
arXiv:2606.03371v4 Announce Type: replace Abstract: Reliable proactive agents must choose an action and judge whether current evidence is sufficient t…
arXiv:2606.03027v2 Announce Type: replace Abstract: Text embeddings are fundamental to many downstream applications, making robustness important for r…
arXiv:2605.29791v2 Announce Type: replace Abstract: While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, th…
arXiv:2605.16023v3 Announce Type: replace Abstract: LLM-as-a-judge has become the dominant paradigm for grading model outputs at scale, yet the same m…
arXiv:2604.27846v2 Announce Type: replace Abstract: How people narrate their experiences offers a window into how the mind organizes them. Computation…
arXiv:2602.11391v5 Announce Type: replace Abstract: Objective: This study develops and validates a patient simulation framework that aligns with the N…
arXiv:2602.02219v3 Announce Type: replace Abstract: Large language models are widely employed as evaluators, a paradigm commonly referred to as LLM-as…
arXiv:2511.16811v2 Announce Type: replace Abstract: Building on the third-wave Extended Mind (EM) theory and radical enactivism, this article suggests…
arXiv:2504.00285v2 Announce Type: replace Abstract: Large Language Models (LLMs) are effective at deceiving when prompted to do so. Models that demons…
arXiv:2609.10397v1 Announce Type: cross Abstract: Exception Related Code (ERC), which includes throw statements, conditions (if statements) that guard…
arXiv:2609.10355v1 Announce Type: cross Abstract: Video understanding has rapidly evolved toward video large language models (VideoLLMs): systems that…
arXiv:2609.10123v1 Announce Type: cross Abstract: Large language models (LLMs) have become ubiquitous in software development, with LLM-based automate…
arXiv:2609.10022v1 Announce Type: cross Abstract: Modern TTS systems approach human quality for high-resource languages but degrade when clean speech…
arXiv:2609.09949v1 Announce Type: cross Abstract: Real-world detectors must often interpret functional or ambiguous prompts, yet conventional models s…
arXiv:2609.09662v1 Announce Type: cross Abstract: Deploying Large Language Models (LLMs) directly on mobile platforms at the edge is gaining traction…
arXiv:2609.09628v1 Announce Type: cross Abstract: Inferring speaker relationships from spoken conversations is an important step towards socially awar…
arXiv:2609.09551v1 Announce Type: cross Abstract: Recommender systems have become core infrastructure for modern online platforms, personalizing conte…
arXiv:2609.09372v1 Announce Type: cross Abstract: Although MMLU is widely adopted as a benchmark for calibrating general AI capabilities, we psychomet…
arXiv:2609.09206v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) often struggle with hallucinations, thus hindering their re…