Skip to content

Category

natural language processing

2,491 papers

#artificial intelligence Review Jul 2026

Error Certificates for KV-Cache Eviction via Randomized Design

It is proved that no estimator computable from the information a deterministic scheme retains is consistent for its own eviction error: evicted values can be altered so that everything retained is unchanged while the true attention-output error grows without bound.

Peng Xie Amr Alanwar · 0 citations

Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation

Data-Adaptive Lower-Rank Adaptation (DALorRA), a simple and effective variational Bayesian sparse framework that shifts the paradigm of uncertainty quantification from the dense parameter space to the lightweight rank level of low-rank adaptation (LoRA).

Ji-Jie Zhang, Zhenjiang Ren, Quan Zhang et al. · 0 citations
#artificial intelligence Preprint May 2026

When the Strongest Teacher Is Not the Best Teacher: Student-Centric Answer Selection

Student-Centric Answer Sampling (SCAS) is proposed, a framework that selects from verified teacher-generated answers according to their estimated student-centric learning cost and is derived by a token-wise gradient decomposition and used to guide answer selection during training.

Zhengyu Hu, Zheyuan Xiao, Linxin Song et al. · 0 citations

OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling

A dynamic per-layer scalar derived by adapting the LARS/LAMB trust-ratio principle to the orthogonalized setting, where the standard denominator candidates---the raw momentum norm or the polar-factor norm---either live in the wrong unit space or carry no update-scale information.

Yuxuan Lou, Yang You · 1 citation

Personalized Group Relative Policy Optimization for Heterogenous Preference Alignment

Personalized GRPO is introduced, a novel alignment framework that decouples advantage estimation from immediate batch statistics and achieves faster convergence and higher rewards than standard GRPO, thereby enhancing its ability to recover and align with heterogeneous preference signals.

Jialu Wang, Heinrich Peters, A. Butt et al. · 1 citation

MUSE: A Run-Centric Platform for Multimodal Unified Safety Evaluation of Large Language Models

MUSE (Multimodal Unified Safety Evaluation), an open-source, browser-based, run-centric platform for multimodal safety evaluation, demonstrates the value of run-centric, fine-grained evaluation for characterizing multimodal safety behavior beyond a single binary success metric.

Zhongxi Wang, Yueqian Lin, Jingyang Zhang et al. · 0 citations

Constrained Group Relative Policy Optimization

This work introduces Constrained GRPO, a Lagrangian-based extension of GRPO for constrained policy optimization, and addresses the coupling induced by reward scalarization by scalarizing standardized advantages rather than rewards.

Roger Girgis, Rodrigue de Schaetzen, Luke Rowe et al. · 2 citations · ⚡1

From tech blogs

See all →
MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.

MIT News · Artificial Intelligence Aug 20, 2026

Paving the way for greener ammonia production

New MIT research could lead to better materials for a fossil-fuel-free process for making the chemical that's essential to fertilizer and other products.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.