Skip to content

Category

natural language processing

2,394 papers

#artificial intelligence Preprint Apr 2026

Do We Still Need Humans in the Loop? Human vs. LLM Annotation in Active Learning for TikTok Hate Speech Detection

LLM annotation at scale outperforms human-supervised classifiers at roughly one-tenth the cost, for both a closed-source and an open-weight LLM, and the advantage is robust under soft-label evaluation.

Ahmad Dawar Hakimi, Lea Hirlimann, Isabelle Augenstein et al. · 0 citations
#natural language process... Preprint Apr 2026

A Robust Evaluation of Probe Robustness: Lessons for Reliable OOD Uncertainty Quantification

ProbeDrift is introduced, a systematic evaluation framework for supervised uncertainty probes covering a wide range of OOD settings across models, tasks, and distributional shifts, and it is argued that robust uncertainty estimation requires robust evaluation.

Joe Stacey, Hadas Orgad, Kentaro Inui et al. · 0 citations

Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation

This work systematically characterize what makes an effective multilingual teacher, and combines intrinsic measures of data quality with extrinsic student model performance in a metric the authors call Polyglot Score, which reveals that model scale alone does not significantly predict teacher effectiveness.

Lester James Validad Miranda, Ivan Vulic, Anna Korhonen · 1 citation

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

MedConceal, a benchmark with an interactive patient simulator for evaluating hidden-concern reasoning in medical dialogue, comprising 300 curated cases and 600 clinician-LLM interactions is presented, identifying hidden-concern reasoning under partial observability as a key unresolved challenge for medical dialogue systems.

Yikun Han, Joey Chan, Jingyuan Chen et al. · 1 citation

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

DLR is proposed, a reinforced latent reasoning framework that dynamically decomposes queries into textual premises, extracts premise-conditioned continuous visual latents, and deduces answers through grounded rationales to enable effective exploration in the latent space.

Mengdan Zhu, Senhao Cheng, Liang Zhao · 0 citations

Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation

A systematic analysis of expert routing patterns in MoE models reveals Language Routing Isolation, in which high- and low-resource languages tend to activate largely disjoint expert sets, and proposes RISE, a framework that exploits routing isolation to identify and adapt language-specific expert subnetworks.

Kening Zheng, Wei-Chieh Huang, Jiahao Huo et al. · 4 citations · ⚡2

GRADE: Probing Knowledge Gaps in LLMs through Gradient Subspace Dynamics

GRADE (GRAdient Dynamics for knowlEdge gap detection), which quantifies the knowledge gap via the cross-layer rank ratio of the gradient to that of the corresponding hidden state subspace, motivated by the property of gradients as estimators of the required knowledge updates for a given target.

Yujing Wang, Yuanbang Liang, Yu-Kun Lai et al. · 1 citation
#artificial intelligence Preprint Apr 2026

Where Does Robustness Live? Neuron-Guided Adaptation for Retrieval-Augmented Language Models

NeuRIT is proposed, a Neuron-guided Robust Instruction-Tuning framework built on a localization-first perspective that mines context-aware neurons associated with relevant and irrelevant context processing, and uses them as anchors to selectively adapt both the identified neuron groups and the layers in which they concentrate.

Jae Lee, Jaemin Kim, Sumyeong Ahn et al. · 0 citations

CARE: Privacy-Compliant Agentic Reasoning with Evidence Discordance

CARE is proposed: a multi-stage privacy-compliant agentic reasoning framework in which a proprietary LLM provides guidance by generating structured categories and transitions without accessing sensitive patient data, while a local LLM uses these categories and transitions to support evidence acquisition and final decision-making.

Hao Liu, Wei-En Li, Rui Song et al. · 0 citations
#natural language process... Preprint Mar 2026

Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models

This work explores a self-supervised framework that encourages models to predict concepts, approximated as sets of semantically equivalent tokens, suggesting that concepts enhance semantic alignment while preserving language modeling quality.

Christine Zhang, Daniel Jurafsky, Sha-Ni Chen · 1 citation · ⚡1

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

SwiAttn is a novel hybrid transformer that enables dynamic and fine-grained routing between full attention and sliding window attention, and dynamically routes the computation to either a full-attention branch for global information aggregation or a sliding-window branch for efficient local pattern matching.

Yusheng Zhao, Hourun Li, Bohan Wu et al. · 4 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.

MIT News · Artificial Intelligence Aug 20, 2026

Paving the way for greener ammonia production

New MIT research could lead to better materials for a fossil-fuel-free process for making the chemical that's essential to fertilizer and other products.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.