Skip to content

Category

natural language processing

2,926 papers

#machine learning Preprint Aug 2026

TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories

TopoCompress is introduced, a training-free and model-agnostic framework that compresses long contexts by selecting coherent semantic spans by selecting coherent semantic spans and achieves performance comparable to the strongest baseline while using a 4x smaller compression budget.

Daniel Agyei Asante, Yang Li · 1 citation
#artificial intelligence Preprint Aug 2026

SingProbe Technical Report

SingProbe is introduced, a lightweight intrinsic runtime guard that directly reuses hidden states produced during LLM inference and operates alongside autoregressive decoding and extends this paradigm to medical generation through SingProbe-Med, which selectively activates risk-directed decoding interventions only when clinically relevant risks emerge.

Singg Team · 0 citations
#machine learning Preprint Aug 2026

What It Costs to Compose, Rebuild, and Correct Precomputed Memory

Both warm-rebuilding trained compressions of key-value caches and serving specifically-phrased updates beside a memory, as pasted text or injected cache state, show particular promise for keeping precomputed memories current, the latter as an interim measure between rebuilds.

Asa Shepard · 0 citations
#artificial intelligence Preprint Aug 2026

GMTS: Gradient Magnitude-based Token Selection Improves RLVR Training for LLM Reasoning

It is found that training on the top 20% tokens ranked by GMTS consistently outperforms entropy-based token selection across three reasoning domains and various model sizes, suggesting that GMTS provides a more fine-grained estimate of token contribution for RLVR training.

Outongyi Lv, Yuan-Wei Zhang, Xiao-Qun Zhang · 1 citation
#artificial intelligence Preprint Aug 2026

Reading the News: Adapting Large Language Models to Swedish Journalism Through Continued Pre-Training

This work investigates continued pre-training for adapting large language models to Swedish journalism, using a high-quality dataset that is curate from millions of news articles and demonstrates the importance of targeted evaluation in the adaptation process.

Lukas Borggren, Jenny Kunz, Marco Kuhlmann · 0 citations
#machine learning Preprint Aug 2026

Two Centuries of Sexism in British Parliament: A Computational Analysis of Women's Representation in the Hansard Corpus

This work analyzes 6,531 speeches over 200 years of UK parliamentary debate by using large language models to classify a speaker's perspective towards women's suffrage and political representation, as well as analyse sexist speech in parliament from the lens of the Ambivalent Sexism Inventory.

Mohammad Omar Khursheed, Mandira Sawkar, Ashiqur R. KhudaBukhsh · 0 citations
#artificial intelligence Preprint Aug 2026

Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions

MineAmongUs is introduced, a 3D multimodal Among Us sandbox where imposter agents must deceive crewmates through joint verbal and non-verbal action, and ARIA is proposed, a configurable VLM-agent harness that exposes five cognitive-component ablation axes and opens a new path for embodied VLM-agent alignment research.

Jaewoo Ahn, Junseo Kim, Hyunseo Kim et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.

MIT News · Artificial Intelligence Aug 20, 2026

Paving the way for greener ammonia production

New MIT research could lead to better materials for a fossil-fuel-free process for making the chemical that's essential to fertilizer and other products.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.