Latest AI/ML News
2784 articles · arXiv cs.CL
arXiv:2609.24574v1 Announce Type: new Abstract: Computational social science increasingly relies on large language models for text annotation, and the…
arXiv:2609.24554v1 Announce Type: new Abstract: Concepts are commonly defined as abstract, compact representations of knowledge and treated as basic u…
arXiv:2609.24538v1 Announce Type: new Abstract: Functional annotation of newly sequenced proteins remains a bottleneck in molecular biology: the numbe…
arXiv:2609.24516v1 Announce Type: new Abstract: In recent years, large language models (LLMs) have emerged as a popular alternative for evaluation. Of…
arXiv:2609.24410v1 Announce Type: new Abstract: Speech-to-text engines are extremely needed nowadays for different applications, representing an essen…
arXiv:2609.24372v1 Announce Type: new Abstract: In-context learning (ICL) based on large language models (LLMs) has shown promising potential in allev…
arXiv:2609.24357v1 Announce Type: new Abstract: Cross-domain Named Entity Recognition (CD-NER) aims to transfer the rich knowledge in the source domai…
arXiv:2609.24275v1 Announce Type: new Abstract: Text-to-speech (TTS) corpora are costly to record, yet many utterances add little new phonetic informa…
arXiv:2609.24264v1 Announce Type: new Abstract: Tool-use agent traces identify messages and API calls, but procedural analyses also need explicit unit…
arXiv:2609.24246v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable progress in natural language understanding, yet the…
arXiv:2609.24238v1 Announce Type: new Abstract: We reproduce and stress-test the work of Yu et al. (2023), who characterize how language models (LMs)…
arXiv:2609.24219v1 Announce Type: new Abstract: Traditionally, the reliability of news publishers is assessed by expert organisations that evaluate ed…
arXiv:2609.24199v1 Announce Type: new Abstract: Evaluation benchmarks for Indian language automatic speech recognition (ASR) suffer from two systemati…
arXiv:2609.24196v1 Announce Type: new Abstract: Looped Language Models (LoopLMs) perform "latent reasoning" by recursively refining internal latent re…
arXiv:2609.24194v1 Announce Type: new Abstract: Evaluation scores used around LLM systems -- including reward models, rerankers, and LLM judges -- can…
arXiv:2609.24177v1 Announce Type: new Abstract: Legal information in Bangladesh is inaccessible to most citizens. Statutory text is English-only, trai…
arXiv:2609.24122v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) is hard to monitor in production: exhaustive relevance labels do…
arXiv:2609.24106v1 Announce Type: new Abstract: Questions scraped from the web are used across academia and industry as a proxy for what people want t…
arXiv:2609.24083v1 Announce Type: new Abstract: Generative AI enables scalable production of educational videos, but current systems largely focus on…
arXiv:2609.24066v1 Announce Type: new Abstract: Best-of-$N$ is a widely used inference strategy for complex reasoning, whose effectiveness depends on…