Bergson is an open source library that aims to enable faster progress in the field by providing a host of techniques that scale to very large language models and pre-training datasets, and provides quality of life tools for researchers.
Lucia Quirke, Louis Jaburi, David O. Johnston et al.· arXiv.org· 0 citations
The results suggest that benchmark composition, rather than numerical insufficiency, determines whether design rules appear to generalize, and that the Facebook-100 regime provides a concrete target for future adaptive aggregation methods.
This submission documents the divide-and-conquer modeling strategy developed for the CTF-4-Science Lorenz Chaotic Systems Challenge at AI-DEEDS 2026, which shows that bounded, scenario-specific updates can outperform broad model replacement on mixed chaotic forecasting benchmarks.
This work introduces a system that converts natural language queries into structured graphs and executes them via a deterministic planner, which uses depth-first search to resolve dependencies and combine results across tools, improving reliability and enabling queries beyond traditional keyword-based search.
A. Chakravarthy, Vidhi Kulkarni, Duen Horng Chau· 0 citations
Reach audiences
Advertise in front of researchers, engineers, and readers.
TriHead-GAN is proposed, a Transformer-based adversarial framework whose triple-head discriminator jointly supervises three complementary aspects of the joint distribution: distributional authenticity via a Wasserstein critic, cross-variable dependency via leakage-free regression of the target variable, and step-wise temporal smoothness via adjacent-difference prediction.
Ze-Sen Wang, Lijuan Lan, Yong-Gang Li et al.· arXiv.org· 0 citations
This work reframe attribution as subset-level counterfactual utility prediction and introduces GRASP, an interaction-aware surrogate that more than doubles the task-level rank correlation for counterfactual subset fidelity while reducing upfront artifact construction costs by nearly an order of magnitude.
The Mamba-Assisted Closure (MAC) framework is proposed, which employs a Mamba-based sequence model to predict the closure from the resolved trajectory and couples the learned closure with the reduced-order governing equations through a numerical integrator to advance the resolved variables in time.
Zhifei Wei, S. Qadeer, Panos Stinis· arXiv.org· 0 citations
GRZO is a Group-Relative Zeroth-Order optimizer that draws one pseudo-independent perturbation per mini-batch example and aggregates the per-example losses through group-relative normalization, raising the effective gradient-direction count from one to the batch size at no additional forward cost while preserving inference-level memory.
L. Tan, Yequan Zhao, Yifan Yang et al.· arXiv.org· 0 citations
The canonical Crawford-Sobel cheap-talk model is turned into a pre-specified benchmark for LLM honesty under preference misalignment, in which theory supplies an exact oracle.
This work proposes a novel block-diagonal Riemannian metric derived from the pullback of the Frobenius inner product and develops a Riemannian gradient descent algorithm that uses a tuning-free Gaussian step size and scales linearly in the number of observed entries per iteration.
This work proposes LaRA, a layer-wise representation analysis framework for detecting contamination in RL post-trained LLMs and finds that contamination produces progressive geometric deviations across layers, including amplified perturbation sensitivity, stronger directional collapse, and enhanced local rigidity.
Minju Gwak, Minseok Kwak, Dongseok Lee et al.· arXiv.org· 0 citations
The OISD framework is proposed, which improves reasoning by transferring on-policy predictive signals from the final layer to intermediate representations and employs signed advantage-weighted Jensen--Shannon alignment to distill informative intermediate representations while preserving policy consistency under a unified acting policy.
Xin-Yu Liu, Darryl C. Jacob, Yang Zhou et al.· arXiv.org· 0 citations
A weeklong summer workshop brought higher education faculty to campus to explore how AI and machine learning materials can be adapted for their classrooms.
A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.
MIT News · Artificial Intelligence· news.mit.eduAug 24, 2026