Skip to content

Category

natural language processing

2,491 papers

#artificial intelligence Preprint Aug 2026

Lot Machine: Multimodal Lot Extraction from Auction Catalogs

This work demonstrates that a VLM-based pipeline can successfully unlock historical auction catalogs for large-scale automated analysis, and benchmark the methods across different deployment modes ranging from commercial providers to locally hosted, quantized models.

Mathias Zinnen, Alisha Mund, Sabine Lang et al. · 0 citations
#artificial intelligence Preprint Aug 2026

EvoSkill Injection: Red-Teaming Autonomous Skill Generation and Evolution in Self-Evolving Agents

A red-teaming framework for evaluating this threat model targeting the autonomous skill generation and evolution pipeline of self-evolving agents and shows that SARGE induces malicious skill formation and that injected skills are persistently stored and repeatedly activated, highlighting the risk of persistent capability corruption.

Doyun Kim, Chanwoo Kim, Sugyeong Eo et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Ignorance or Incompetence? Constructing Knowledge-Gated, Verifiable Tasks for LLM Agents

A knowledge-gated task-construction protocol is introduced that separates a task instruction from a compact artefact containing private conventions, reference tables, and utility operators, and it is shown that the retained tasks improve post-training.

Han-Lin Tian, Min-Hao Li, Yuhan Mi et al. · 0 citations
#artificial intelligence Preprint Aug 2026

LaMoC: Loss-Aware Modular Compression for LLMs

LaMoC improves joint compression by selecting compression statistics that better align local module reconstruction error with the downstream loss, and reformulate joint modular compression as a two-tiered optimization problem that minimizes module reconstruction error while tuning the activation and gradient information blending rate.

Mohanad Odema, Jacob Song · 0 citations
#computer vision Preprint Aug 2026

Towards a Joint Khmer Text Recognition and Word Segmentation

Experimental results show that the proposed model can not only recognize characters in document images but also locate word boundaries, removing the need for an extra word segmentation step in a conventional sequential pipeline.

Marry Kong, Rina Buoy, Sovisal Chenda et al. · 0 citations
#artificial intelligence Preprint Aug 2026

A.X K2 Technical Report

To support long contexts efficiently, Sparse Gated Attention (SGA), which combines sparse attention with gated attention, and adopt Gated Norm (GN) to stabilize large-scale training is introduced, which keeps 4-bit NVFP4 serving within one point of FP8 accuracy.

Cheolseung Baek, Dhammiko Arya, Eunki Kim et al. · 0 citations
#natural language process... Preprint Aug 2026

The Language of the Question Selects the Market: Query Language and Exit IP as Separable Factors in Commercial Recommendations from a Generative Search Interface

A controlled probe of 234 runs against the logged-out ChatGPT web interface and the OpenAI API, collected on 29 and 30 August 2026 across four exit countries and six query languages, shows that language and location are separable and act on different things.

Dmitrij Żatuchin · 1 citation

Demand-Side Measurement for Generative Engine Optimization: Constructing and Validating a Million-Persona, Intent-Annotated Buyer Corpus

This work built and validated PersonaGen-1M, a corpus of 1,031,732 synthetic buyer personas spanning 511 industry labels and 4 market contexts, carrying 19,416,821 structured behavioral attributes, 5,160,046 of them search queries.

Dmitrij Żatuchin, Daniil Dzemesjuk · 1 citation
#computer vision Preprint Aug 2026

SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Models

SpanCalib-VLM is presented, a hybrid dual-system for the SHROOM-Visions Shared Task that combines a multimodal sequence tagger, consisting of XLM-RoBERTa-Large fused with a SigLIP vision encoder via cross-attention, with the fine-tuned generative VLM (Qwen3.5-4B-SHROOM-SFT).

A. Abebe, Yasmin Moslem · 0 citations
#natural language process... Preprint Aug 2026

Do LLMs Change Their Minds Like Humans? Diagnosing Human--LLM Divergence in Single-Turn Persuasion Judgments

A systematic comparison using a naturally occurring online persuasion corpus in which original posters explicitly verify whether a reply changed their view is conducted, highlighting the risk of treating LLM judgments as faithful proxies for human belief updating and point to structural differences in how LLMs and humans process persuasive discourse.

Lin Chen, Yi-Tong Chen, Yong Li · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.

MIT News · Artificial Intelligence Aug 20, 2026

Paving the way for greener ammonia production

New MIT research could lead to better materials for a fossil-fuel-free process for making the chemical that's essential to fertilizer and other products.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.